Accompanying code for "Discovering State-of-the-art Reinforcement Algorithms" Nature publication
☆738Dec 2, 2025Updated 9 months ago
Alternatives and similar repositories for disco_rl
Users that are interested in disco_rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- streaming deep reinforcement learning but 4x faster with jax!☆19Jan 4, 2026Updated 8 months ago
- Unified Implementations of Offline Reinforcement Learning Algorithms☆232Dec 19, 2025Updated 9 months ago
- Official implementation for "How Should We Meta-Learn Reinforcement Learning Algorithms?"☆23Sep 7, 2025Updated last year
- Mastering Diverse Domains through World Models☆3,825May 25, 2026Updated 4 months ago
- (Crafter + NetHack) in JAX. ICML 2024 Spotlight.☆455Jun 20, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Hydra sweeper integration of our favorite optimization packages, utilizing ask-and-tell interfaces.☆16Nov 14, 2025Updated 10 months ago
- The official implementation of "Horizon Reduction Makes RL Scalable"☆205Aug 2, 2025Updated last year
- Official implementation of Stackelberg PPO for morphology–control co-design.☆19Mar 17, 2026Updated 6 months ago
- Deep reinforcement learning without experience replay, target networks, or batch updates.☆300Updated this week
- JAX-accelerated Meta-Reinforcement Learning Environments Inspired by XLand and MiniGrid 🏎️☆346Dec 16, 2025Updated 9 months ago
- Simple single-file baselines for Q-Learning in pure-GPU setting☆245Nov 24, 2025Updated 10 months ago
- Code for Scalable Offline Model-Based RL with Action chunking☆34Feb 20, 2026Updated 7 months ago
- Multi-Agent Reinforcement Learning with JAX☆854Sep 10, 2026Updated 2 weeks ago
- A simple, performant and scalable JAX-based world modeling codebase.☆164Jan 15, 2026Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for "SimbaV2: Hyperspherical Normalization for Scalable Deep Reinforcement Learning"☆112Nov 4, 2025Updated 10 months ago
- Implementation of Dreamer v3 in pytorch.☆890Mar 8, 2026Updated 6 months ago
- Reinforcement learning on general 2D physics environments in JAX. ICLR 2025 Oral.☆274Updated this week
- Code for "Transitive RL: Value Learning via Divide and Conquer"☆61Oct 31, 2025Updated 10 months ago
- SBX: Stable Baselines Jax (SB3 + Jax) RL algorithms☆610Sep 17, 2026Updated last week
- Implementation of Danijar's latest iteration for his Dreamer line of work☆220Sep 3, 2026Updated 3 weeks ago
- Official implementation of DiscoGen, for "Procedural Generation of Algorithm Discovery Tasks in Machine Learning"☆50Sep 16, 2026Updated last week
- HPO and Architecture Benchmarking for RL: Dynamically, Reactive and Efficient☆34Jun 16, 2026Updated 3 months ago
- ☆101Jan 21, 2026Updated 8 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 🏛️A research-friendly codebase for fast experimentation of single-agent reinforcement learning in JAX • End-to-End JAX RL☆422Mar 18, 2026Updated 6 months ago
- RL Environments in JAX 🌍☆931Aug 30, 2026Updated 3 weeks ago
- High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, T…☆10,457Apr 20, 2026Updated 5 months ago
- ☆131Feb 25, 2025Updated last year
- [NeurIPS 2025] TTRL: Test-Time Reinforcement Learning☆1,124Apr 15, 2026Updated 5 months ago
- Official implementation of the δ-model presented in the ICML 2024 paper "A Distributional Analogue to the Successor Representation".☆23Nov 8, 2024Updated last year
- Implementations of Multi-Task and Meta-Learning baselines for the Metaworld benchmark☆39May 20, 2026Updated 4 months ago
- Code for "Reversal Q-Learning (RQL)" for Flow RL from Prior Data☆37Jun 17, 2026Updated 3 months ago
- ☆17Apr 23, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ⚡ Flashbax: Accelerated Replay Buffers in JAX☆282Aug 27, 2026Updated last month
- Really Fast End-to-End Jax RL Implementations☆1,110Sep 9, 2024Updated 2 years ago
- Evolution Pretraining Fully in Int Formats☆183Feb 25, 2026Updated 7 months ago
- A benchmark for offline goal-conditioned RL and offline RL☆490Jan 14, 2026Updated 8 months ago
- [NeurIPS'21 Outstanding Paper] Library for reliable evaluation on RL and ML benchmarks, even with only a handful of seeds.☆878Aug 12, 2024Updated 2 years ago
- LeanRL is a fork of CleanRL, where selected PyTorch scripts optimized for performance using compile and cudagraphs.☆703Aug 22, 2025Updated last year
- Official code release for "CrossQ: Batch Normalization in Deep Reinforcement Learning for Greater Sample Efficiency and Simplicity"☆98Jun 4, 2024Updated 2 years ago