Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.
☆562Nov 17, 2025Updated 9 months ago
Alternatives and similar repositories for LLM-RL-Papers
Users that are interested in LLM-RL-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A collection of LLM with RL papers☆281Apr 24, 2024Updated 2 years ago
- ☆21Apr 12, 2024Updated 2 years ago
- A comprehensive list of PAPERS, CODEBASES, and, DATASETS on Decision Making using Foundation Models including LLMs and VLMs.☆383Apr 24, 2024Updated 2 years ago
- We perform functional grounding of LLMs' knowledge in BabyAI-Text☆275Oct 27, 2025Updated 10 months ago
- [ICLR 2024 Spotlight] Text2Reward: Reward Shaping with Language Models for Reinforcement Learning☆212Dec 17, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official implementation of Zero-Hero paper☆32Feb 13, 2025Updated last year
- Implementation of TWOSOME☆81Jan 11, 2025Updated last year
- ☆30Nov 7, 2025Updated 10 months ago
- LLM-Empowered State Representation for Reinforcement Learning (ICML2024 Accepted paper)☆42Jun 14, 2024Updated 2 years ago
- ☆91Aug 21, 2023Updated 3 years ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆20Aug 9, 2024Updated 2 years ago
- Python code to implement LLM4Teach, a policy distillation approach for teaching reinforcement learning agents with Large Language Model☆55Apr 19, 2024Updated 2 years ago
- Natural Language Reinforcement Learning☆103Jul 30, 2025Updated last year
- Lamorel is a Python library designed for RL practitioners eager to use Large Language Models (LLMs).☆248Dec 11, 2025Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [IROS2023]Learning to Solve Tasks with Exploring Prior Behaviours☆13Mar 3, 2024Updated 2 years ago
- A curated list of reinforcement learning with human feedback resources (continually updated)☆4,425May 20, 2026Updated 3 months ago
- A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites☆4,461Jul 17, 2026Updated last month
- ☆66Jan 30, 2026Updated 7 months ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- A Survey on Large Language Model-Based Game Agents (ACM CSUR)☆959Jun 7, 2026Updated 3 months ago
- ☆64Nov 15, 2024Updated last year
- ☆49Jun 27, 2025Updated last year
- TextStarCraft2,a pure language env which support llms play starcraft2☆357Apr 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- Agent RL framework for LLM agents: multi-turn reinforcement learning with StarPO and reasoning-collapse diagnostics☆2,797Aug 23, 2026Updated 2 weeks ago
- A curated list of awesome model based RL resources (continually updated)☆1,394May 21, 2026Updated 3 months ago
- Large Language Models and Robotics.☆22Apr 27, 2024Updated 2 years ago
- ☆17Apr 14, 2026Updated 4 months ago
- [ICANN 2022] ''An Improved Lightweight YOLOv5 Model Based on Attention Mechanism for Face Mask Detection'' Official Code☆10Feb 27, 2024Updated 2 years ago
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges (In IJCAI 2024)☆1,312Aug 21, 2026Updated 2 weeks ago
- Unofficial PyTorch implementation (replicating paper results) of Implicit Q-Learning (In-sample Q-Learning) for offline RL☆24Nov 4, 2024Updated last year
- BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL alg…☆659Feb 7, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS 2025] The official implementation of "RF-Agent: Automated Reward Function Design via Language Agent Tree Search"☆16Feb 27, 2026Updated 6 months ago
- This repository contains a collection of resources and papers on Diffusion Models for RL, accompanying the paper "Diffusion Models for Re…☆671Nov 29, 2024Updated last year
- ☆15Mar 26, 2024Updated 2 years ago
- Official Implementation of "Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models" (accep…☆21Dec 14, 2025Updated 8 months ago
- A curated list of Diffusion Model in RL resources (continually updated)☆1,638May 30, 2026Updated 3 months ago
- Official Code Repository for EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents (COLM 2024)☆41Jul 13, 2024Updated 2 years ago
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,994Updated this week