From search engines, to science, to robotics, this reposity is meant to showcase the use of reinforcement learning in the world..
☆294Aug 9, 2026Updated this week
Alternatives and similar repositories for DeepRLInTheWorld
Users that are interested in DeepRLInTheWorld are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mitigating Partial Observability in Sequential Decision Processes via the Lambda Discrepancy☆24Oct 28, 2024Updated last year
- Official Code for "Relative Entropy Pathwise Policy Optimization"☆59May 6, 2026Updated 3 months ago
- Source code for the paper "Policy Architectures for Compositional Generalization in Control"☆30May 19, 2022Updated 4 years ago
- Scalable Computation of Hessian Diagonals☆14Jun 2, 2024Updated 2 years ago
- [NeurIPS'21 Outstanding Paper] Library for reliable evaluation on RL and ML benchmarks, even with only a handful of seeds.☆880Aug 12, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- On the model-based stochastic value gradient for continuous reinforcement learning☆58Mar 6, 2026Updated 5 months ago
- Code for the paper "Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning". Great performance in many environments…☆39Oct 24, 2025Updated 9 months ago
- Docker containers of baseline agents for the Crafter environment☆30Dec 14, 2021Updated 4 years ago
- RL Environments in JAX 🌍☆912Apr 2, 2026Updated 4 months ago
- Library for Model Based RL☆1,064Jul 12, 2024Updated 2 years ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- 📴 OffCon^3: SOTA PyTorch SAC and TD3 Implementations (arxiv: 2101.11331)☆25Jun 20, 2021Updated 5 years ago
- JAX (Flax) implementation of algorithms for Deep Reinforcement Learning with continuous action spaces.☆757Oct 26, 2022Updated 3 years ago
- 🕹️ A diverse suite of scalable reinforcement learning environments in JAX☆855Aug 4, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Release code for ICML2020 Knowing The What But Not The Where in Bayesian Optimization☆15Mar 7, 2023Updated 3 years ago
- High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, T…☆10,254Apr 20, 2026Updated 3 months ago
- Jax/Flax Implementation of TD-MPC2☆81Jul 28, 2026Updated 2 weeks ago
- ☆333Dec 19, 2024Updated last year
- Simple and easily configurable grid world environments for reinforcement learning☆2,496Aug 6, 2026Updated last week
- Evaluating long-term memory of reinforcement learning algorithms☆181Jun 23, 2023Updated 3 years ago
- JAX-accelerated Meta-Reinforcement Learning Environments Inspired by XLand and MiniGrid 🏎️☆343Dec 16, 2025Updated 7 months ago
- Deep reinforcement learning without experience replay, target networks, or batch updates.☆294Mar 18, 2025Updated last year
- Standardized Minecraft Diamond Environment for Reinforcement Learning☆40May 19, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- DrQ-v2: Improved Data-Augmented Reinforcement Learning☆439May 31, 2022Updated 4 years ago
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Really Fast End-to-End Jax RL Implementations☆1,093Sep 9, 2024Updated last year
- A framework for evaluating LLMs in Atari games☆15Apr 21, 2025Updated last year
- An implementation of DreamerV2 written in JAX, with support for running multiple random seeds of an experiment on a single GPU.☆18Jan 16, 2023Updated 3 years ago
- Official codebase for LEAP: Planning with Goal Conditioned Policies☆51Sep 30, 2022Updated 3 years ago
- Code for the paper "Inference via Interpolation: Contrastive Representations Provably Enable Planning and Inference"☆44Jul 10, 2024Updated 2 years ago
- C++-based high-performance parallel environment execution engine (vectorized env) for general RL environments.☆1,494Jul 17, 2026Updated 3 weeks ago
- A collection of reference environments for offline reinforcement learning☆1,699Nov 18, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Simple JAX Graphics Library.☆38Nov 3, 2024Updated last year
- [AutoML'22] Bayesian Generational Population-based Training (BG-PBT)☆31Sep 16, 2022Updated 3 years ago
- A System for Morphology-Task Generalization via Unified Representation and Behavior Distillation (ICLR2023)☆14Feb 3, 2023Updated 3 years ago
- Implementations of robust Dual Curriculum Design (DCD) algorithms for unsupervised environment design.☆141Aug 20, 2024Updated last year
- Code for the paper Learning Visible Connectivity Dynamics for Cloth Smoothing☆49Nov 13, 2022Updated 3 years ago
- Simplifying Model-based RL: Learning Representations, Latent-space Models and Policies with One Objective☆82Mar 9, 2023Updated 3 years ago
- code for CoRL 2020 paper "Contrastive Variational Model-Based Reinforcement Learning for Complex Observations"☆24Dec 29, 2021Updated 4 years ago