Building and Scaling RL environments in the age of LLMs
☆165Aug 6, 2026Updated this week
Alternatives and similar repositories for RL_Envs_101
Users that are interested in RL_Envs_101 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The best ChatGPT that $100 can buy.☆58Updated this week
- ☆15Dec 4, 2025Updated 8 months ago
- An interface library for RL post training with environments.☆2,490Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,138Updated this week
- A new substrate for browser agents: files, actions, diffs, logs, and artifacts.☆22May 26, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 7 months ago
- ☆35Apr 28, 2025Updated last year
- Agentic RL Training at Scale☆1,878Updated this week
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆432Updated this week
- KV Cache & LoRA for minGPT☆62Mar 4, 2026Updated 5 months ago
- Our library for RL environments + evals☆4,485Updated this week
- A movie journal coupled with open IMDb data, and a Flask web-app for easy movie insertion.☆11Apr 13, 2026Updated 3 months ago
- ☆42Nov 7, 2025Updated 9 months ago
- torch_remat fine-grained activation checkpointing API☆15Jul 28, 2026Updated 2 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆29Jul 23, 2025Updated last year
- Exploratory stuff on RLMs for video workflows☆35Apr 20, 2026Updated 3 months ago
- Framework for evaluating and improving agents☆4,076Updated this week
- Carla worldsim envrionment for decision-making evaluation and RL☆19Apr 22, 2026Updated 3 months ago
- A Python SDK for Open Reward Standard servers and clients☆17Mar 24, 2026Updated 4 months ago
- Agentic RL on Any Harness at Scale☆760Updated this week
- ☆896Updated this week
- Meta-Reinforcement Learning with Self-Reflection☆33Mar 26, 2026Updated 4 months ago
- Convert any Repo into an RL Environment☆490Jul 27, 2026Updated 2 weeks ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆228Apr 30, 2026Updated 3 months ago
- ☆42Mar 26, 2025Updated last year
- Tora: Torchtune-LoRA for RL☆87Dec 2, 2025Updated 8 months ago
- ☆13Mar 2, 2023Updated 3 years ago
- Provide performance insight capabilities for RL frameworks.☆53Updated this week
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,945Updated this week
- 🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models☆10,726Jul 30, 2026Updated last week
- VisualToolAgent (VisTA): A Reinforcement Learning Framework for Visual Tool Selection☆27May 31, 2025Updated last year
- OmniDocs📄 - One stop visual document processing framework☆144Mar 19, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Training terminal-agents☆274Updated this week
- A Gym for Agentic LLMs☆504Jan 21, 2026Updated 6 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 3 months ago
- GLM-5.2 for ZGX GB10☆21Jul 3, 2026Updated last month
- Modular Rust transformer/LLM library using Candle☆40May 5, 2024Updated 2 years ago
- A super quick way to update npm dependencies, via a code lens.☆18Feb 21, 2026Updated 5 months ago
- Textbook on reinforcement learning from human feedback☆2,274Updated this week