Deep reinforcement learning without experience replay, target networks, or batch updates.
☆293Mar 18, 2025Updated last year
Alternatives and similar repositories for streaming-drl
Users that are interested in streaming-drl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Action Value Gradient Algorithm☆29May 18, 2025Updated last year
- Algorithms for Gradient TD updates☆19Feb 21, 2026Updated 5 months ago
- ☆65Jan 30, 2026Updated 5 months ago
- Simple single-file baselines for Q-Learning in pure-GPU setting☆242Nov 24, 2025Updated 8 months ago
- ☆34Jun 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning". Great performance in many environments…☆39Oct 24, 2025Updated 9 months ago
- Code for "SimbaV2: Hyperspherical Normalization for Scalable Deep Reinforcement Learning"☆108Nov 4, 2025Updated 8 months ago
- ☆13Aug 15, 2020Updated 5 years ago
- Flax Implementation of DreamerV3 on Crafter☆19Nov 29, 2025Updated 8 months ago
- ☆128Feb 25, 2025Updated last year
- Online Goal-Conditioned Reinforcement Learning in JAX. ICLR 2025 Spotlight.☆274Jun 6, 2026Updated last month
- Real-Time RTUs☆12Mar 20, 2026Updated 4 months ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago
- Code for AAAI 2023 paper "Hypernetworks for Zero-shot Transfer in Reinforcement Learning"☆24Apr 26, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A framework for Reinforcement Learning research.☆268Updated this week
- Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning.☆44Feb 9, 2026Updated 5 months ago
- Minimal Decision Transformer Implementation written in Jax (Flax).☆18Aug 8, 2022Updated 3 years ago
- Official implementation for "How Should We Meta-Learn Reinforcement Learning Algorithms?"☆23Sep 7, 2025Updated 10 months ago
- Agar.io for Continual Reinforcement Learning☆24Jul 24, 2025Updated last year
- Clean single-file implementation of offline RL algorithms in JAX☆182Jun 5, 2026Updated last month
- (Crafter + NetHack) in JAX. ICML 2024 Spotlight.☆432Jun 20, 2026Updated last month
- Simple single file implementations of Reinforcement Learning algorithms in Julia☆24Feb 15, 2025Updated last year
- PyTorch implementation for "Discovery of Incremental Skills" (DISk) algorithm from ICLR 2022 paper "One After Another: Learning Increment…☆21Mar 22, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆98Jan 21, 2026Updated 6 months ago
- DrQ-v2: Improved Data-Augmented Reinforcement Learning☆439May 31, 2022Updated 4 years ago
- Really Fast End-to-End Jax RL Implementations☆1,093Sep 9, 2024Updated last year
- [ICML 2025 oral] Network Sparsity Unlocks the Scaling Potential of Deep Reinforcement Learning☆41Jun 5, 2025Updated last year
- Code for the paper "FIRE: Frobenius-Isometry Reinitialization for Balancing the Stability–Plasticity Tradeoff" (ICLR 2026 Oral)☆30Apr 27, 2026Updated 3 months ago
- Reinforcement Learning inside a 3D soccer simulation☆37Sep 15, 2024Updated last year
- Official release of the DMControl Generalization Benchmark 2 (DMC-GB2)☆22Jul 21, 2025Updated last year
- ☆96Feb 16, 2026Updated 5 months ago
- JAX implementation of deep RL agents with resets from the paper "The Primacy Bias in Deep Reinforcement Learning"☆106May 17, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆63Aug 3, 2023Updated 2 years ago
- ☆28May 11, 2026Updated 2 months ago
- ☆20Oct 27, 2025Updated 9 months ago
- ☆23Aug 19, 2022Updated 3 years ago
- Implements the Messenger environment and EMMA model.☆25Jun 14, 2023Updated 3 years ago
- ☆26Jan 26, 2024Updated 2 years ago
- The official implementation of "Horizon Reduction Makes RL Scalable"☆200Aug 2, 2025Updated 11 months ago