A curated list of reinforcement learning (RL) for agents.
☆111Aug 24, 2026Updated last week
Alternatives and similar repositories for awesome-rl-for-agents
Users that are interested in awesome-rl-for-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2024] CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition☆16Nov 12, 2025Updated 9 months ago
- ☆512Oct 11, 2025Updated 10 months ago
- [IROS 2023] Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition☆21Jul 12, 2025Updated last year
- ☆1,878Jun 18, 2026Updated 2 months ago
- [ICCV2023] Chaotic World: A Large and Challenging Benchmark for Human Behavior Understanding in Chaotic Events☆10Dec 7, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Implementation of the paper: VG4D: Vision-Language Model Goes 4D Video Recognition(ICRA 2024)☆15Apr 23, 2024Updated 2 years ago
- Awesome List for Agentic RL☆1,817Updated this week
- [CVIU2026] Implementation of the paper: ModelNet-O: A Large-Scale Synthetic Dataset for Occlusion-Aware Point Cloud Classification☆13Aug 31, 2024Updated 2 years ago
- The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Eva…☆290Jul 30, 2026Updated last month
- 🌟 A curated list of papers, methods, and resources on long-horizon credit assignment for agentic RL.☆61Aug 6, 2026Updated 3 weeks ago
- RLAnything (ICML 2026) & AutoTool (ICML 2026), DemyAgent: Open-Source RL for LLMs and Agentic Scenarios☆625Jun 12, 2026Updated 2 months ago
- 从 幻觉翻译 获取基于 LaTex 源码翻译的arXiv文章☆21Jul 20, 2026Updated last month
- [IJCV2026/NeurIPS2023] Implementation of the paper: Explore In-Context Learning for 3D Point Cloud Understanding☆74Mar 18, 2026Updated 5 months ago
- ☆17Apr 10, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- model compression and optimization for deployment for Pytorch, including knowledge distillation, quantization and pruning.(知识蒸馏,量化,剪枝)☆21Sep 10, 2024Updated last year
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervision☆106Jul 23, 2026Updated last month
- The Notion Citation Updater is a Python script designed to automate the process of updating citation counts for academic papers stored in…☆15Oct 28, 2024Updated last year
- Repository about single/multi-agent, robotics, llm/vlm/vla, scientific discovery, etc.☆20Jul 10, 2025Updated last year
- Unofficial implementation of Chain of Hindsight (https://arxiv.org/abs/2302.02676) using pytorch and huggingface Trainers.☆11Apr 5, 2023Updated 3 years ago
- ☆49Mar 15, 2025Updated last year
- ☆24Jul 8, 2023Updated 3 years ago
- [ICML 2025] LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models☆16Nov 4, 2025Updated 9 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,111Aug 20, 2026Updated last week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- PreAct: Prediction Enhances Agent's Planning Ability (Coling2025)☆31Dec 12, 2024Updated last year
- MedSoft-Diffusion was early accepted to MICCAI 2025 (top 9%, scores: 5/4/4).☆43Mar 1, 2025Updated last year
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- ☆192Aug 13, 2026Updated 2 weeks ago
- Official Implementation of Papar CM2☆26Apr 21, 2026Updated 4 months ago
- ☆16Nov 22, 2025Updated 9 months ago
- A Survey of Reinforcement Learning for Large Reasoning Models☆2,481Aug 20, 2026Updated last week
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆111Aug 18, 2026Updated last week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,640Aug 24, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆27Jul 18, 2023Updated 3 years ago
- [EMNLP 2024 Findings] Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information☆13Oct 1, 2024Updated last year
- slime is an LLM post-training framework for RL Scaling.☆8,313Updated this week
- R1V, trained with AI feedback, answers open-ended visual questions.☆14Apr 12, 2025Updated last year
- World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).☆493Updated this week
- The official repository for Trust-Region Adaptive Policy Optimization (TRAPO) – a novel hybrid framework designed to enhance large langua…☆16Mar 2, 2026Updated 5 months ago
- [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models☆31Mar 21, 2026Updated 5 months ago