Decentralized Reinforcment Learning: Global Decision-Making via Local Economic Transactions (ICML 2020)
☆43Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for decentralized-rl
Users that are interested in decentralized-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Receding Horizon Curiosity Algrithm☆13Mar 24, 2023Updated 3 years ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- [EMNLP 2024] Introducing Filtered Direct Preference Optimization (fDPO) that enhances language model alignment with human preferences by …☆16Nov 27, 2024Updated last year
- MuJoCo models for Unitree Robots☆12Nov 24, 2021Updated 4 years ago
- Official PyTorch Implementation for Metric Residual Networks for Sample Efficient Goal-Conditioned Reinforcement Learning☆21Jan 11, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- POMDP wrappers for OpenAI Gym☆15Nov 4, 2019Updated 6 years ago
- ☆18Feb 7, 2021Updated 5 years ago
- A small library for creating and manipulating custom JAX Pytree classes☆56Feb 26, 2023Updated 3 years ago
- ☆11Nov 3, 2022Updated 3 years ago
- Repo for the multi-agent PressurePlate environment☆20Feb 4, 2022Updated 4 years ago
- A set of environments utilizing pybullet for simulation of robotic manipulation tasks.☆29Mar 8, 2021Updated 5 years ago
- High-level Python Particle Sequential Convex Programming Model Predictive Control (SCP PMPC) interface☆16Oct 30, 2023Updated 2 years ago
- Library that provides environments for planning problems☆17Apr 24, 2026Updated 4 months ago
- ☆26Apr 27, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A brief JAX tutorial with examples from control theory☆12Nov 17, 2022Updated 3 years ago
- Implementation for "ROLL: Visual Self-Supervised Reinforcement Learning with Object Reasoning", CoRL 2020☆16Jun 22, 2022Updated 4 years ago
- Neural Fixed-Point Acceleration for Convex Optimization☆30Oct 6, 2022Updated 3 years ago
- Extending rllab to event-driven multiagent environments☆13Oct 1, 2018Updated 7 years ago
- ☆11Jun 4, 2021Updated 5 years ago
- Few-shot Bayesian Imitation Learning with Policies as Logic over Programs☆22Oct 19, 2025Updated 10 months ago
- Scaling scaling laws with board games.☆54Jul 17, 2023Updated 3 years ago
- ☆15Jun 8, 2023Updated 3 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Codebase for Offline Reinforcement Learning from Images with Latent Space Models☆31Apr 30, 2021Updated 5 years ago
- AGAC: Adversarially Guided Actor-Critic☆47Sep 16, 2021Updated 4 years ago
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Pytorch implementation of Stable Opponent Shaping (https://openreview.net/pdf?id=SyGjjsC5tQ).☆22Jan 15, 2020Updated 6 years ago
- Quasi-Newton Algorithm for Stochastic Optimization☆11May 20, 2022Updated 4 years ago
- Experiments in protein folding through language modeling☆10Dec 10, 2021Updated 4 years ago
- ☆17Mar 21, 2021Updated 5 years ago
- Pytorch code for "Learning Belief Representations for Imitation Learning in POMDPs" (UAI 2019)☆22Aug 4, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Differentiable Gaussian Process Motion Planning☆51Sep 1, 2021Updated 5 years ago
- Code for the paper Continual Learning from Demonstration of Robotic Skills☆37May 3, 2023Updated 3 years ago
- Code to reproduce the experimental results from the paper "Active Invariant Causal Prediction: Experiment Selection Through Stability", b…☆22Jul 6, 2023Updated 3 years ago
- ☆22Nov 8, 2021Updated 4 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- Agent Zero RL Framework☆15Nov 22, 2024Updated last year
- Jax implementation of Proximal Policy Optimization (PPO) specifically tuned for Procgen, with benchmarked results and saved model weights…☆63Aug 4, 2022Updated 4 years ago