A curated list of reinforcement learning with verifiable rewards (continually updated)
☆324Jun 1, 2026Updated 3 months ago
Alternatives and similar repositories for awesome-RLVR
Users that are interested in awesome-RLVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2025 Tokenization Workshop] HH-Codec: High Compression High-fidelity Discrete Neural Codec for Spoken Language Modeling☆108Sep 28, 2025Updated 11 months ago
- Decision Intelligence Adventure for Beginners☆112Dec 9, 2022Updated 3 years ago
- Python library for solving reinforcement learning (RL) problems using generative models (e.g. Diffusion Models).☆223Feb 18, 2025Updated last year
- A curated list of of awesome UI agents resources, encompassing Web, App, OS, and beyond (continually updated)☆319Jun 17, 2026Updated 3 months ago
- Collection of latest papers and materials in the area of RLVR!☆142Sep 7, 2026Updated 2 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- PsyDI: Towards a Personalized and Progressively In-depth Chatbot for Psychological Measurements. (e.g. MBTI Measurement Agent)☆227Aug 4, 2025Updated last year
- OpenDILab RL Kubernetes Custom Resource and Operator Lib☆280Jan 9, 2023Updated 3 years ago
- A curated list of Multi-Modal Reinforcement Learning resources (continually updated)☆620May 30, 2026Updated 3 months ago
- Decision Intelligence platform for Biological Sequence Searching☆153Oct 10, 2022Updated 3 years ago
- 1024 + 深度强化学习(Deep Reinforcement Learning + 1024 Game/ 2048 Game)☆158Jul 23, 2024Updated 2 years ago
- 羊了个羊 + 深度强化学习(Deep Reinforcement Learning + 3 Tiles Game)☆519Mar 10, 2025Updated last year
- Here are the most awesome tree structure computing solutions, make your life easier. (这里有目前性能最优的树形结构计算解决方案)☆270Oct 17, 2024Updated last year
- Open-Source Reproduction/Demo of the LLM Riddles Game☆586Jul 30, 2024Updated 2 years ago
- Auxiliary code for pulling, loading reinforcement learning models based on DI-engine from the Huggingface Hub, or pushing them onto Huggi…☆72Dec 12, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆195Dec 26, 2022Updated 3 years ago
- A curated list of Decision Transformer resources (continually updated)☆917May 21, 2026Updated 4 months ago
- [ICML 2024] "Envisioning Outlier Exposure by Large Language Models for Out-of-Distribution Detection"☆13Feb 15, 2025Updated last year
- [CVPR 2023] ReasonNet: End-to-End Driving with Temporal and Global Reasoning☆194Jun 29, 2023Updated 3 years ago
- [ICLR 2025] "Noisy Test-Time Adaptation in Vision-Language Models"☆16Feb 22, 2025Updated last year
- PPO x Family DRL Tutorial Course(决策智能入门级公开课:8节课帮你盘清算法理论,理顺代码逻辑,玩转决策AI应用实践 )☆2,621Mar 13, 2025Updated last year
- Official implementation of ICML 2025 paper "Understanding Multimodal LLMs Under Distribution Shifts: An Information-Theoretic Approach"☆12May 27, 2025Updated last year
- A curated list of resources on Reinforcement Learning with Verifiable Rewards (RLVR) and the reasoning capability boundary of Large Langu…☆93Dec 12, 2025Updated 9 months ago
- verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework☆23,583Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A curated collection of papers, technical reports, frameworks, and tools for on-policy distillation (OPD) of large language models☆843Aug 30, 2026Updated 3 weeks ago
- [ICML 2024] Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models☆19Mar 23, 2026Updated 6 months ago
- Code of ICLR 2025 paper "DynaPrompt: Dynamic Test-Time Prompt Tuning"☆22Jan 29, 2025Updated last year
- ☆15Apr 17, 2026Updated 5 months ago
- A curated collection of papers and resources on On-Policy Distillation for Large Language Models.☆558Aug 12, 2026Updated last month
- ☆20Jul 1, 2026Updated 2 months ago
- ☆11Jan 16, 2020Updated 6 years ago
- ☆1,899Jun 18, 2026Updated 3 months ago
- A curated list of papers and resources on Reward Hacking, Emergent Misalignment, and Proxy Exploitation in Large Models☆53Apr 17, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- High-quality and streaming Speech-to-Speech interactive agent in a single file. 只用一个文件实现的流式全双工语音交互原型智能体!☆541Apr 7, 2026Updated 5 months ago
- PRD: Peer Rank and Discussion Improve Large Language Model based Evaluations☆14Apr 21, 2024Updated 2 years ago
- A Survey of Reinforcement Learning for Large Reasoning Models☆2,489Sep 14, 2026Updated last week
- A curated list of Diffusion Model in RL resources (continually updated)☆1,642May 30, 2026Updated 3 months ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆641Jun 12, 2026Updated 3 months ago
- [AAAI 2023] Official PyTorch implementation of paper "ACE: Cooperative Multi-agent Q-learning with Bidirectional Action-Dependency".☆261Dec 7, 2022Updated 3 years ago
- [EMNLP 2024] ”ESC-Eval: Evaluating Emotion Support Conversations in Large Language Models“☆27Jun 24, 2024Updated 2 years ago