OpenAI团队的深度强化学习教程中文版
☆36May 16, 2020Updated 6 years ago
Alternatives and similar repositories for spinningup
Users that are interested in spinningup are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Codes for the paper "Multi-task Hierarchical Adversarial Inverse Reinforcement Learning"☆19May 20, 2023Updated 3 years ago
- Codebase for ReLMM☆23Apr 17, 2023Updated 3 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- PyTorch implementation of the paper Overcoming Exploration in Reinforcement Learning with Demonstrations in surgical robot manipulation t…☆12Aug 21, 2022Updated 3 years ago
- Modified Pytorch Lightning implementation of paper:-https://jcheminf.biomedcentral.com/track/pdf/10.1186/s13321-019-0407-y☆10Dec 22, 2020Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An easy to understand implementation of the paper "Model-Based Reinforcement Learning for Atari"☆18Sep 27, 2019Updated 6 years ago
- Code accompanying the paper "TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play" (AAMAS 2023) 足球游戏智能体☆14May 25, 2023Updated 3 years ago
- 臸娥粂陆亩竟☆10May 11, 2024Updated 2 years ago
- ☆15Oct 10, 2025Updated 9 months ago
- Using inverse kinematics (IK) solving to achieve partial hand movements for the Unitree G1 robot, based on unitree-rl-gym and IsaacGym.☆24May 8, 2025Updated last year
- Open source code for paper "Learning World Models with Identifiable Factorization"☆13Mar 4, 2024Updated 2 years ago
- opencv调用jetson/rk3588 mpp硬解码,重写了open与read函数,支持h264/h265☆14Nov 27, 2025Updated 8 months ago
- Collision Avoidance using Buffered Voronoi Cell☆14Feb 10, 2017Updated 9 years ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- GAMMA: A General Agent Motion Prediction Model for Autonomous Driving☆14Nov 17, 2021Updated 4 years ago
- android compose catalog☆17Jul 4, 2025Updated last year
- Ossian generic framework☆12Aug 25, 2021Updated 4 years ago
- ☆34Mar 24, 2023Updated 3 years ago
- ☆19Jun 30, 2024Updated 2 years ago
- SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores☆15Apr 24, 2024Updated 2 years ago
- An implementation of HOME: Heatmap Output for future Motion Estimation☆13Feb 7, 2022Updated 4 years ago
- Made for a reading group at the Center for Safe AGI.☆12Feb 23, 2026Updated 5 months ago
- Stanford CS234: Reinforcement Learning Winter 2020☆19Mar 24, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 📜 [NeurIPS 2022] "Symbolic Distillation for Learned TCP Congestion Control", S P Sharan, Wenqing Zheng, Kuo-Feng Hsu, Jiarong Xing, Ang …☆16Oct 13, 2022Updated 3 years ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- Codes for Complex-Valued Spectrum Estimation Network and Applications in Super-Resolution HRRPs Analysis with Wideband Radars☆12Dec 10, 2021Updated 4 years ago
- An reconstruction of RL Introduction and its course materials for a more efficient entry☆19Mar 4, 2026Updated 4 months ago
- Learning Kinematic Feasibility through Reinforcement Leanring: http://rl.uni-freiburg.de/research/kinematic-feasibility-rl☆24Jan 27, 2021Updated 5 years ago
- ☆29Oct 10, 2018Updated 7 years ago
- Reimplementation of Policy Optimization with Demonstrations (POfD) from ICML 2018.☆16Jun 5, 2019Updated 7 years ago
- fpv vehicle powered by esp32 cam☆10Aug 9, 2022Updated 3 years ago
- Maddpg_flight code☆10Jul 4, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- machine learning algorithms source code☆25Jun 8, 2021Updated 5 years ago
- Code of the paper "Universal Morphology Control via Contextual Modulation" at ICML 2023☆15Aug 3, 2023Updated 2 years ago
- PSO for Nash Equilibrium. This is the code for my undergraduate thesis.粒子群算法求解纳什均衡☆12Jan 5, 2023Updated 3 years ago
- ☆16Aug 30, 2022Updated 3 years ago
- object tracking learning☆13Aug 29, 2020Updated 5 years ago
- Flash Artifact for SIGCOMM22☆14Jun 14, 2022Updated 4 years ago
- Reinforcement learning based multi object tracker☆10Jan 29, 2018Updated 8 years ago