FineEnvs — RL Environments 101: building and scaling RL environments in the age of LLMs
☆225Sep 21, 2026Updated this week
Alternatives and similar repositories for FineEnvs
Users that are interested in FineEnvs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The best ChatGPT that $100 can buy.☆58Updated this week
- An interface library for RL post training with environments.☆2,602Updated this week
- ☆15Aug 22, 2026Updated last month
- SkyRL: A Modular Full-stack RL Library for LLMs☆2,339Updated this week
- Agentic RL Training at Scale☆2,068Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Optimizing diffusion for production-ready speeds☆40Jan 10, 2026Updated 8 months ago
- A scalable asynchronous reinforcement learning implementation with in-flight weight updates.☆434Aug 5, 2026Updated last month
- KV Cache & LoRA for minGPT☆61Mar 4, 2026Updated 6 months ago
- Our library for RL environments + evals☆4,642Updated this week
- A movie journal coupled with open IMDb data, and a Flask web-app for easy movie insertion.☆11Updated this week
- An agentic-first and HuggingFace-native RL framework for research (9k lines).☆1,149Updated this week
- torch_remat fine-grained activation checkpointing API☆20Sep 9, 2026Updated last week
- Learning to route instances for Human vs AI Feedback (ACL Main '25)☆30Jul 23, 2025Updated last year
- Carla worldsim envrionment for decision-making evaluation and RL☆19Apr 22, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Framework for evaluating and improving agents☆5,496Updated this week
- Agentic RL on Any Harness at Scale☆843Aug 13, 2026Updated last month
- Hugging Face's take home challenge for post-training internships, now open for you to try!☆51Apr 23, 2026Updated 4 months ago
- ☆134Jun 1, 2026Updated 3 months ago
- Meta-Reinforcement Learning with Self-Reflection☆36Mar 26, 2026Updated 5 months ago
- [ICML 2026] RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments☆234Apr 30, 2026Updated 4 months ago
- Χ-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows?☆63Updated this week
- Convert any Repo into an RL Environment☆665Updated this week
- Environments by the Prime Intellect Research Team☆130Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Tora: Torchtune-LoRA for RL☆87Dec 2, 2025Updated 9 months ago
- Slides for the System Design series sessions.☆12Aug 31, 2022Updated 4 years ago
- My Solution to Assignments of CS234(Stanford / Fall 2019)☆14Sep 3, 2020Updated 6 years ago
- Archived — ML Intern is no longer maintained. Continue with HuggingChat.☆10,823Sep 14, 2026Updated last week
- Textbook on reinforcement learning from human feedback☆2,404Sep 11, 2026Updated last week
- Training terminal-agents☆311Updated this week
- A Gym for Agentic LLMs☆510Jan 21, 2026Updated 8 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 4 months ago
- ☆30May 9, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SimLab is the data layer for creating simulations to QA, evaluate, hillclimb, and refine agents.☆24Aug 5, 2026Updated last month
- Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.☆2,969Updated this week
- Modular Rust transformer/LLM library using Candle☆40May 5, 2024Updated 2 years ago
- Lab Cookbook☆42Aug 5, 2026Updated last month
- ☆25Oct 10, 2025Updated 11 months ago
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆18Updated this week
- My own experimental AGI (similar to AutoGPT).☆14Mar 16, 2024Updated 2 years ago