A curated list of reinforcement learning (RL) for agents.
☆111Jul 27, 2026Updated last week
Alternatives and similar repositories for awesome-rl-for-agents
Users that are interested in awesome-rl-for-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding☆23Jul 6, 2026Updated 3 weeks ago
- [NeurIPS 2024] CHASE: Learning Convex Hull Adaptive Shift for Skeleton-based Multi-Entity Action Recognition☆16Nov 12, 2025Updated 8 months ago
- ☆511Oct 11, 2025Updated 9 months ago
- [MMAsia 2023] Official PyTorch implementation of the paper " Cross-Modal Retrieval for Motion and Text via DropTriple Loss "☆38Nov 30, 2024Updated last year
- [IROS 2023] Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition☆21Jul 12, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆1,851Jun 18, 2026Updated last month
- [IEEE TIP 2024] Facial Prior Guided Micro-Expression Generation☆13Nov 8, 2024Updated last year
- Awesome List for Agentic RL☆1,747Jul 23, 2026Updated last week
- The official repository of "A Comprehensive Survey on Reinforcement Learning-based Agentic Search: Foundations, Roles, Optimizations, Eva…☆281Updated this week
- Enhanced version of original AutoGPTQ (https://github.com/PanQiWei/AutoGPTQ).☆10Nov 2, 2023Updated 2 years ago
- 从 幻觉翻译 获取基于 LaTex 源码翻译的arXiv文章☆19Jul 20, 2026Updated 2 weeks ago
- ☆17Apr 10, 2024Updated 2 years ago
- ☆14May 21, 2024Updated 2 years ago
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervision☆105Jul 23, 2026Updated last week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Repository about single/multi-agent, robotics, llm/vlm/vla, scientific discovery, etc.☆20Jul 10, 2025Updated last year
- Unofficial implementation of Chain of Hindsight (https://arxiv.org/abs/2302.02676) using pytorch and huggingface Trainers.☆11Apr 5, 2023Updated 3 years ago
- ☆48Mar 15, 2025Updated last year
- [ICML 2025] LaCache: Ladder-Shaped KV Caching for Efficient Long-Context Modeling of Large Language Models☆16Nov 4, 2025Updated 8 months ago
- [ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)☆1,098Jul 13, 2026Updated 3 weeks ago
- MedSoft-Diffusion was early accepted to MICCAI 2025 (top 9%, scores: 5/4/4).☆43Mar 1, 2025Updated last year
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- ☆176Jul 2, 2026Updated last month
- This is an agent (including contextual prompts) that queries your CSV☆10Jun 8, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Jun 11, 2025Updated last year
- A curated list of resources (surveys, papers, benchmarks, and opensource projects) on Rubrics☆103Updated this week
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,593Jul 27, 2026Updated last week
- [EMNLP 2024 Findings] Wrong-of-Thought: An Integrated Reasoning Framework with Multi-Perspective Verification and Wrong Information☆13Oct 1, 2024Updated last year
- slime is an LLM post-training framework for RL Scaling.☆7,742Updated this week
- World model reasoning RL for multi-turn VLM agents☆491Jul 23, 2026Updated last week
- A comprehensive paper list of Table-based Question Answering.☆40Sep 1, 2023Updated 2 years ago
- [ICLR 2026] Mixing Importance with Diversity: Joint Optimization for KV Cache Compression in Large Vision-Language Models☆29Mar 21, 2026Updated 4 months ago
- Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcemen…☆830Feb 15, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 🔍 OpenSearch-VL provides a fully open recipe for training strong multimodal deep search agents through high-quality data curation, diver…☆260Updated this week
- The implementation for SIGIR 2026: Learning to Retrieve from Agent Trajectories.☆56Jul 14, 2026Updated 2 weeks ago
- ☆22Jun 16, 2026Updated last month
- Summaries of ICML 2024 papers☆12Jul 31, 2024Updated 2 years ago
- A curated list of awesome resources about reward construction for AI agents. This repository covers cutting-edge research, and practical …☆60Sep 1, 2025Updated 11 months ago
- GRU-PPO for stable-baselines3.☆13Apr 24, 2024Updated 2 years ago
- Code repo for "CritiPrefill: A Segment-wise Criticality-based Approach for Prefilling Acceleration in LLMs".☆17Sep 15, 2024Updated last year