A curated list of reinforcement learning with verifiable rewards (continually updated)
☆265Jun 1, 2026Updated last month
Alternatives and similar repositories for awesome-RLVR
Users that are interested in awesome-RLVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025 SynthData Workshop Spotlight] Empowering LLMs in Decision Games through Algorithmic Data Synthesis☆50Apr 27, 2025Updated last year
- [NeurIPS 2025 AI for Music Workshop] Vocal Reaction Model and Benchmark☆46Dec 10, 2025Updated 7 months ago
- [ICCV 2025] Pretrained Reversible Generation as Unsupervised Visual Representation Learning☆47Nov 5, 2025Updated 8 months ago
- Decision Intelligence Adventure for Beginners☆112Dec 9, 2022Updated 3 years ago
- A curated list of of awesome UI agents resources, encompassing Web, App, OS, and beyond (continually updated)☆313Jun 17, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- LightRFT: Light, Efficient, Omni-modal & Reward-model Driven Reinforcement Fine-Tuning Framework☆404Apr 29, 2026Updated 2 months ago
- Collection of latest papers and materials in the area of RLVR!☆135Updated this week
- OpenDILab RL HPC OP Lib, including CUDA and Triton kernel☆266Jul 4, 2024Updated 2 years ago
- A curated list of Multi-Modal Reinforcement Learning resources (continually updated)☆617May 30, 2026Updated last month
- Decision Intelligence platform for Biological Sequence Searching☆153Oct 10, 2022Updated 3 years ago
- 羊了个羊 + 深度强化学习(Deep Reinforcement Learning + 3 Tiles Game)☆518Mar 10, 2025Updated last year
- Open-Source Reproduction/Demo of the LLM Riddles Game☆587Jul 30, 2024Updated last year
- Auxiliary code for pulling, loading reinforcement learning models based on DI-engine from the Huggingface Hub, or pushing them onto Huggi…☆72Dec 12, 2023Updated 2 years ago
- [CVPR 2024] LMDrive: Closed-Loop End-to-End Driving with Large Language Models☆924Apr 14, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A curated list of Decision Transformer resources (continually updated)☆914May 21, 2026Updated 2 months ago
- [EMNLP 2024 Tutorial] Language Agents: Foundations, Prospects, and Risks☆10Nov 27, 2024Updated last year
- OpenDILab RL Object Store☆197Apr 20, 2022Updated 4 years ago
- A curated list of awesome exploration RL resources (continually updated)☆715May 21, 2026Updated 2 months ago
- [CoRL 2022] InterFuser: Safety-Enhanced Autonomous Driving Using Interpretable Sensor Fusion Transformer☆653Jan 4, 2026Updated 6 months ago
- [ICLR 2025] "Noisy Test-Time Adaptation in Vision-Language Models"☆16Feb 22, 2025Updated last year
- Official implementation of ICML 2025 paper "Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach"☆12May 27, 2025Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆89Dec 12, 2025Updated 7 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆22,654Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models☆564Updated this week
- Our library for RL environments + evals☆4,400Updated this week
- ☆15Apr 17, 2026Updated 3 months ago
- A curated collection of papers and resources on On-Policy Distillation for Large Language Models.☆470Updated this week
- ☆19Jul 1, 2026Updated 3 weeks ago
- ☆15Updated this week
- ☆11Jan 16, 2020Updated 6 years ago
- ☆1,846Jun 18, 2026Updated last month
- A curated list of papers and resources on Reward Hacking, Emergent Misalignment, and Proxy Exploitation in Large Models☆42Apr 17, 2026Updated 3 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICLR 2026] Constructive Distortion: Multimodal LLMs with Attention‑Aware Image Warping☆23Feb 9, 2026Updated 5 months ago
- High-quality and streaming Speech-to-Speech interactive agent in a single file. 只用一个文件实现的流式全双工语音交互原型智能体!☆534Apr 7, 2026Updated 3 months ago
- A Survey of Reinforcement Learning for Large Reasoning Models☆2,469Nov 9, 2025Updated 8 months ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆592Jun 12, 2026Updated last month
- [EMNLP 2024] ”ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models“☆27Jun 24, 2024Updated 2 years ago
- [ICCV 2023] Efficient Joint Optimization of Layer-Adaptive Weight Pruning in Deep Neural Networks☆25Oct 31, 2023Updated 2 years ago
- Accompanying code for our NeurIPS 2019 paper☆11Nov 7, 2019Updated 6 years ago