Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.
☆562Nov 17, 2025Updated 10 months ago
Alternatives and similar repositories for LLM-RL-Papers
Users that are interested in LLM-RL-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A collection of LLM with RL papers☆282Apr 24, 2024Updated 2 years ago
- ☆21Apr 12, 2024Updated 2 years ago
- A comprehensive list of PAPERS, CODEBASES, and, DATASETS on Decision Making using Foundation Models including LLMs and VLMs.☆384Apr 24, 2024Updated 2 years ago
- We perform functional grounding of LLMs' knowledge in BabyAI-Text☆277Oct 27, 2025Updated 11 months ago
- [ICLR 2024 Spotlight] Text2Reward: Reward Shaping with Language Models for Reinforcement Learning☆213Dec 17, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation of Zero-Hero paper☆32Feb 13, 2025Updated last year
- Implementation of TWOSOME☆81Jan 11, 2025Updated last year
- ☆30Nov 7, 2025Updated 10 months ago
- LLM-Empowered State Representation for Reinforcement Learning (ICML2024 Accepted paper)☆42Jun 14, 2024Updated 2 years ago
- ☆91Aug 21, 2023Updated 3 years ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆20Aug 9, 2024Updated 2 years ago
- Python code to implement LLM4Teach, a policy distillation approach for teaching reinforcement learning agents with Large Language Model☆55Apr 19, 2024Updated 2 years ago
- Natural Language Reinforcement Learning☆104Jul 30, 2025Updated last year
- Lamorel is a Python library designed for RL practitioners eager to use Large Language Models (LLMs).☆248Dec 11, 2025Updated 9 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A curated list of reinforcement learning with human feedback resources (continually updated)☆4,431May 20, 2026Updated 4 months ago
- A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites☆4,476Jul 17, 2026Updated 2 months ago
- ☆66Jan 30, 2026Updated 8 months ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- A Survey on Large Language Model-Based Game Agents (ACM CSUR)☆966Jun 7, 2026Updated 3 months ago
- ☆65Nov 15, 2024Updated last year
- TextStarCraft2,a pure language env which support llms play starcraft2☆361Apr 25, 2025Updated last year
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 7 years ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,810Aug 23, 2026Updated last month
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆51Jun 27, 2025Updated last year
- A curated list of awesome model based RL resources (continually updated)☆1,400May 21, 2026Updated 4 months ago
- Large Language Models and Robotics.☆22Apr 27, 2024Updated 2 years ago
- ☆17Apr 14, 2026Updated 5 months ago
- [ICANN 2022] ''An Improved Lightweight YOLOv5 Model Based on Attention Mechanism for Face Mask Detection'' Official Code☆10Feb 27, 2024Updated 2 years ago
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges (In IJCAI 2024)☆1,321Aug 21, 2026Updated last month
- Unofficial PyTorch implementation (replicating paper results) of Implicit Q-Learning (In-sample Q-Learning) for offline RL☆24Nov 4, 2024Updated last year
- BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL alg…☆668Feb 7, 2026Updated 7 months ago
- [NeurIPS 2025] The official implementation of "RF-Agent: Automated Reward Function Design via Language Agent Tree Search"☆17Feb 27, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository contains a collection of resources and papers on Diffusion Models for RL, accompanying the paper "Diffusion Models for Re…☆671Nov 29, 2024Updated last year
- ☆15Mar 26, 2024Updated 2 years ago
- Official Implementation of "Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models" (accep…☆23Dec 14, 2025Updated 9 months ago
- A curated list of Diffusion Model in RL resources (continually updated)☆1,644May 30, 2026Updated 4 months ago
- Official Code Repository for EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents (COLM 2024)☆41Jul 13, 2024Updated 2 years ago
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆10,057Sep 17, 2026Updated 2 weeks ago
- Train transformer language models with reinforcement learning.☆19,422Updated this week