OpenAI团队的深度强化学习教程中文版
☆36May 16, 2020Updated 6 years ago
Alternatives and similar repositories for spinningup
Users that are interested in spinningup are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Codes for the paper "Multi-task Hierarchical Adversarial Inverse Reinforcement Learning"☆19May 20, 2023Updated 3 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- Modified Pytorch Lightning implementation of paper:-https://jcheminf.biomedcentral.com/track/pdf/10.1186/s13321-019-0407-y☆10Dec 22, 2020Updated 5 years ago
- A PyTorch implementation of SSINet.☆16Nov 10, 2020Updated 5 years ago
- OpenAI团队的深度强化学习教程中文版☆92May 21, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation for mSAC methods in PyTorch☆42Oct 10, 2021Updated 4 years ago
- 臸娥粂陆亩竟☆10May 11, 2024Updated 2 years ago
- Neural Message Passing for NMR Chemical Shift Prediction☆11Aug 10, 2022Updated 4 years ago
- ☆15Oct 10, 2025Updated 10 months ago
- Open source code for paper "Learning World Models with Identifiable Factorization"☆13Mar 4, 2024Updated 2 years ago
- 📖The Big-&-Extending-Repository-of-Transformers: Pretrained PyTorch models for Google's BERT, OpenAI GPT & GPT-2, Google/CMU Transformer…☆11May 30, 2019Updated 7 years ago
- Experimenting with meta-learning approaches to opponent modelling in MARL. Building upon previous public implementations of MADDPG and M3…☆14Apr 26, 2022Updated 4 years ago
- Collision Avoidance using Buffered Voronoi Cell☆14Feb 10, 2017Updated 9 years ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- GAMMA: A General Agent Motion Prediction Model for Autonomous Driving☆14Nov 17, 2021Updated 4 years ago
- android compose catalog☆17Jul 4, 2025Updated last year
- learning robust rewards with adversarial inverse reinforcement learning☆14Sep 13, 2020Updated 5 years ago
- ☆34Mar 24, 2023Updated 3 years ago
- Qt-like event loops, signals and slots for communication across threads and processes in Python☆14Mar 26, 2024Updated 2 years ago
- MindSpore implementations of deep reinforcement learning algorithms and environments☆17Sep 3, 2023Updated 2 years ago
- YEY Blog ->☆13Jun 26, 2025Updated last year
- SRL: Scaling Distributed Reinforcement Learning to Over Ten Thousand Cores☆15Apr 24, 2024Updated 2 years ago
- An implementation of HOME: Heatmap Output for future Motion Estimation☆13Feb 7, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Stanford CS234: Reinforcement Learning Winter 2020☆19Mar 24, 2023Updated 3 years ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- An reconstruction of RL Introduction and its course materials for a more efficient entry☆19Mar 4, 2026Updated 5 months ago
- Reimplementation of Policy Optimization with Demonstrations (POfD) from ICML 2018.☆16Jun 5, 2019Updated 7 years ago
- fpv vehicle powered by esp32 cam☆10Aug 9, 2022Updated 4 years ago
- Macro-Action Generator-Critic (MAGIC) - Learning Macro-actions for online POMDP planning☆17Feb 23, 2023Updated 3 years ago
- Code of the paper "Universal Morphology Control via Contextual Modulation" at ICML 2023☆15Aug 3, 2023Updated 3 years ago
- PSO for Nash Equilibrium. This is the code for my undergraduate thesis.粒子群算法求解纳什均衡☆12Jan 5, 2023Updated 3 years ago
- ☆16Aug 30, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Flash Artifact for SIGCOMM22☆15Jun 14, 2022Updated 4 years ago
- ☆21Jul 2, 2024Updated 2 years ago
- Reinforcement learning based multi object tracker☆10Jan 29, 2018Updated 8 years ago
- MLOT - A machine learning algorithm, written for use with MATLAB, in order to track in 3D moving particles based on a training data set. …☆12Dec 24, 2018Updated 7 years ago
- Project 1 of Udacity's Deep Reinforcement Learning nanodegree program☆13Dec 2, 2018Updated 7 years ago
- 《Reinforcement Learning: An Introduction》(第二版)中文翻译☆687Apr 9, 2022Updated 4 years ago
- EKF Indoor Tracker estimates static user position by using extended kalman filters☆16Nov 15, 2015Updated 10 years ago