Monitoring recent cross-research on LLM & RL on arXiv for control. If there are good papers, PRs are welcome.
☆559Nov 17, 2025Updated 9 months ago
Alternatives and similar repositories for LLM-RL-Papers
Users that are interested in LLM-RL-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A collection of LLM with RL papers☆281Apr 24, 2024Updated 2 years ago
- ☆21Apr 12, 2024Updated 2 years ago
- A comprehensive list of PAPERS, CODEBASES, and, DATASETS on Decision Making using Foundation Models including LLMs and VLMs.☆383Apr 24, 2024Updated 2 years ago
- We perform functional grounding of LLMs' knowledge in BabyAI-Text☆275Oct 27, 2025Updated 9 months ago
- [ICLR 2024 Spotlight] Text2Reward: Reward Shaping with Language Models for Reinforcement Learning☆210Dec 17, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of Zero-Hero paper☆31Feb 13, 2025Updated last year
- Implementation of TWOSOME☆82Jan 11, 2025Updated last year
- ☆29Nov 7, 2025Updated 9 months ago
- LLM-Empowered State Representation for Reinforcement Learning (ICML2024 Accepted paper)☆42Jun 14, 2024Updated 2 years ago
- ☆91Aug 21, 2023Updated 3 years ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆19Aug 9, 2024Updated 2 years ago
- Python code to implement LLM4Teach, a policy distillation approach for teaching reinforcement learning agents with Large Language Model☆55Apr 19, 2024Updated 2 years ago
- Natural Language Reinforcement Learning☆101Jul 30, 2025Updated last year
- Lamorel is a Python library designed for RL practitioners eager to use Large Language Models (LLMs).☆249Dec 11, 2025Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [IROS2023]Learning to Solve Tasks with Exploring Prior Behaviours☆13Mar 3, 2024Updated 2 years ago
- A curated list of reinforcement learning with human feedback resources (continually updated)☆4,422May 20, 2026Updated 3 months ago
- A comprehensive list of papers using large language/multi-modal models for Robotics/RL, including papers, codes, and related websites☆4,453Jul 17, 2026Updated last month
- ☆66Jan 30, 2026Updated 6 months ago
- This repository is an implementation of "MASER: Multi-Agent Reinforcement Learning with Subgoals Generated from Experience Replay Buffer"…☆23Jul 6, 2023Updated 3 years ago
- A Survey on Large Language Model-Based Game Agents (ACM CSUR)☆952Jun 7, 2026Updated 2 months ago
- ☆64Nov 15, 2024Updated last year
- ☆46Jun 27, 2025Updated last year
- TextStarCraft2,a pure language env which support llms play starcraft2☆354Apr 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.☆2,772Jul 24, 2026Updated 3 weeks ago
- A curated list of awesome model based RL resources (continually updated)☆1,392May 21, 2026Updated 3 months ago
- Large Language Models and Robotics.☆22Apr 27, 2024Updated 2 years ago
- ☆16Apr 14, 2026Updated 4 months ago
- [ICANN 2022] ''An Improved Lightweight YOLOv5 Model Based on Attention Mechanism for Face Mask Detection'' Official Code☆10Feb 27, 2024Updated 2 years ago
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges (In IJCAI 2024)☆1,304Jul 12, 2026Updated last month
- Unofficial PyTorch implementation (replicating paper results) of Implicit Q-Learning (In-sample Q-Learning) for offline RL☆24Nov 4, 2024Updated last year
- BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL alg…☆654Feb 7, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains a collection of resources and papers on Diffusion Models for RL, accompanying the paper "Diffusion Models for Re…☆670Nov 29, 2024Updated last year
- ☆15Mar 26, 2024Updated 2 years ago
- Official Implementation of "Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models" (accep…☆20Dec 14, 2025Updated 8 months ago
- A curated list of Diffusion Model in RL resources (continually updated)☆1,633May 30, 2026Updated 2 months ago
- Official Code Repository for EnvGen: Generating and Adapting Environments via LLMs for Training Embodied Agents (COLM 2024)☆40Jul 13, 2024Updated 2 years ago
- An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Asy…☆9,937Aug 13, 2026Updated last week
- Train transformer language models with reinforcement learning.☆19,111Updated this week