Building and Scaling RL environments in the age of LLMs
☆154May 6, 2026Updated 2 months ago
Alternatives and similar repositories for RL_Envs_101
Users that are interested in RL_Envs_101 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The best ChatGPT that $100 can buy.☆54Updated this week
- An interface library for RL post training with environments.☆2,439Updated this week
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,085Updated this week
- 🌱 A little course on Reinforcement Learning Environments for evaluating and training Language Models☆216May 27, 2026Updated last month
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Agentic RL Training at Scale☆1,702Updated this week
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆427Updated this week
- Our library for RL environments + evals☆4,390Updated this week
- torch_remat fine-grained activation checkpointing API☆15Updated this week
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆29Jul 23, 2025Updated 11 months ago
- Framework for evaluating and improving agents☆3,348Updated this week
- A Python SDK for Open Reward Standard servers and clients☆17Mar 24, 2026Updated 3 months ago
- Carla worldsim envrionment for decision-making evaluation and RL☆19Apr 22, 2026Updated 2 months ago
- Hugging Face's take home challenge for post-training internships, now open for you to try!☆48Apr 23, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆546Updated this week
- ☆128Jun 1, 2026Updated last month
- Meta-Reinforcement Learning with Self-Reflection☆33Mar 26, 2026Updated 3 months ago
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆225Apr 30, 2026Updated 2 months ago
- Spellbound - your multilingual AI-powered writing assistant☆13May 12, 2025Updated last year
- ☆42Mar 26, 2025Updated last year
- Dump your codebase into single file, so you can use it as an input to LLMs, like ChatGPT, Google Gemini (directly or through or Google AI…☆14Feb 11, 2025Updated last year
- Tora: Torchtune-LoRA for RL☆87Dec 2, 2025Updated 7 months ago
- Slides for the System Design series sessions.☆12Aug 31, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- My Solution to Assignments of CS234(Stanford / Fall 2019)☆15Sep 3, 2020Updated 5 years ago
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆1,761Updated this week
- Dynamo Workshop☆19Nov 7, 2025Updated 8 months ago
- OmniDocs📄 - One stop visual document processing framework☆141Mar 19, 2026Updated 4 months ago
- Training terminal-agents☆238Updated this week
- A Gym for Agentic LLMs☆502Jan 21, 2026Updated 6 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 2 months ago
- SimLab is the data layer for creating simulations to QA, evaluate, hillclimb, and refine agents.☆23Updated this week
- Modular Rust transformer/LLM library using Candle☆39May 5, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 🤗 ml-intern: an open-source ML engineer that reads papers, trains models, and ships ML models☆10,672Updated this week
- Textbook on reinforcement learning from human feedback☆2,171Updated this week
- Lab Cookbook☆37Updated this week
- A super quick way to update npm dependencies, via a code lens.☆18Feb 21, 2026Updated 5 months ago
- ☆25Oct 10, 2025Updated 9 months ago
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆17Jul 13, 2026Updated last week
- My own experimental AGI (similar to AutoGPT).☆14Mar 16, 2024Updated 2 years ago