☆45Jul 1, 2026Updated 2 months ago
Alternatives and similar repositories for MTBench
Users that are interested in MTBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code for "Relative Entropy Pathwise Policy Optimization"☆61Aug 23, 2026Updated last week
- Code for "SimbaV2: Hyperspherical Normalization for Scalable Deep Reinforcement Learning"☆111Nov 4, 2025Updated 9 months ago
- ☆17Apr 23, 2026Updated 4 months ago
- ☆462May 16, 2026Updated 3 months ago
- ☆35Mar 26, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- The official implementation of "Horizon Reduction Makes RL Scalable"☆205Aug 2, 2025Updated last year
- Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning.☆46Feb 9, 2026Updated 6 months ago
- Official code repository for the paper "Learning Massively Multitask World Models for Continuous Control".☆138Jan 9, 2026Updated 7 months ago
- The official implementation of Value Flows☆56Feb 27, 2026Updated 6 months ago
- ☆18Oct 1, 2025Updated 11 months ago
- ☆31Jun 30, 2026Updated 2 months ago
- MR.Q is a general-purpose model-free reinforcement learning algorithm.☆155Apr 7, 2026Updated 4 months ago
- Implementations of Multi-Task and Meta-Learning baselines for the Metaworld benchmark☆39May 20, 2026Updated 3 months ago
- JAX implementation of WSRL and RL baselines | ICLR 2025☆149Feb 26, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official implementation for "How Should We Meta-Learn Reinforcement Learning Algorithms?"☆23Sep 7, 2025Updated 11 months ago
- ☆51Sep 18, 2025Updated 11 months ago
- Official Implementation of `An Optimisation Framework for Unsupervised Environment Design` from RLC 2025☆17Nov 24, 2025Updated 9 months ago
- ☆129Feb 25, 2025Updated last year
- A framework for Reinforcement Learning research.☆280Jul 28, 2026Updated last month
- [ICLR 2024] Official code of the paper "Multi-Task Reinforcement Learning with Mixture of Orthogonal Experts.☆51Nov 18, 2024Updated last year
- Code for SAPG: Split and Aggregate Policy Gradients (ICML 2024)☆87Sep 17, 2024Updated last year
- DEAS + Isaac-GR00T + RoboCasa☆19Nov 22, 2025Updated 9 months ago
- Official Repository for "Eurekaverse: Environment Curriculum Generation via Large Language Models" (CoRL 2024)☆117May 26, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official repo for paper "TD-M(PC)^2: Improving Temporal Difference MPC Through Policy Constraint"☆88Feb 11, 2025Updated last year
- Decoupled Q-Chunking☆75May 3, 2026Updated 3 months ago
- [ICLR 2024] Adaptive Replay Ratio implementation from 'Revisiting Plasticity in Visual RL: Data, Modules and Training Stages'.☆13Oct 9, 2024Updated last year
- ☆402Feb 5, 2026Updated 6 months ago
- ☆28May 11, 2026Updated 3 months ago
- Bottom-Up Skill Discovery from Unsegmented Demonstrations for Long-Horizon Robot Manipulation (BUDS)☆57Dec 2, 2021Updated 4 years ago
- Code for the paper "Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning". Great performance in many environments…☆39Oct 24, 2025Updated 10 months ago
- ☆30Nov 26, 2025Updated 9 months ago
- ☆12Apr 9, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Action Value Gradient Algorithm☆30May 18, 2025Updated last year
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆17Jul 3, 2026Updated 2 months ago
- Official implementation of the BRO algorithm☆62Jan 29, 2025Updated last year
- ☆30Aug 29, 2024Updated 2 years ago
- Efficiently send large arrays across machines☆15Jul 24, 2024Updated 2 years ago
- FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control☆468Apr 9, 2026Updated 4 months ago
- Code Release for floq: Training Critics via Flow-Matching for Scaling Compute In Value-Based RL☆46Apr 7, 2026Updated 4 months ago