This is the reading list for the survey "A Survey on the Optimization of LLM-based Agents ". We will keep adding papers and improving the list. Any suggestions and PRs are welcome!
☆242Aug 12, 2026Updated this week
Alternatives and similar repositories for Awesome-LLM-Agent-Optimization-Papers
Users that are interested in Awesome-LLM-Agent-Optimization-Papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 这是对基于大模型的多智能体系统论文的总结☆10Jun 23, 2024Updated 2 years ago
- ☆87May 14, 2026Updated 3 months ago
- A collection of papers and libraries for performing multi-agent optimization☆21Jul 12, 2026Updated last month
- A repo lists papers related to LLM based agent☆2,334Jul 12, 2025Updated last year
- ☆34Sep 19, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Latest Advances on Reasoning of Multimodal Large Language Models (Multimodal R1 \ Visual R1) ) 🍓☆36Apr 3, 2025Updated last year
- MPO: Boosting LLM Agents with Meta Plan Optimization (EMNLP 2025 Findings)☆82Aug 20, 2025Updated 11 months ago
- Implementation for "Surrogate Losses for Online Learning of Stepsizes in Stochastic Non-Convex Optimization"☆10Aug 3, 2022Updated 4 years ago
- An reconstruction of RL Introduction and its course materials for a more efficient entry☆19Mar 4, 2026Updated 5 months ago
- Contains implementation of the DoubIL and ResiduIL algorithms from the ICML '22 paper Causal Imitation Learning under Temporally Correlat…☆11Dec 9, 2022Updated 3 years ago
- ☆506Jul 28, 2025Updated last year
- Survey on LLM Agents (Published on CoLing 2025)☆518Oct 3, 2025Updated 10 months ago
- ☆21Jun 9, 2025Updated last year
- verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in…☆2,218Jun 9, 2026Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- This repository contains code for the paper "Learning Decision Trees as Amortized Structure Inference"☆16Mar 25, 2025Updated last year
- ☆15Oct 28, 2024Updated last year
- Code and implementations for the ACL 2025 paper "AgentGym: Evolving Large Language Model-based Agents across Diverse Environments" by Zhi…☆829May 30, 2026Updated 2 months ago
- ☆1,282Oct 15, 2025Updated 9 months ago
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- An RL-Friendly Vision-Language Model for Minecraft☆41Oct 17, 2024Updated last year
- ☆28May 30, 2026Updated 2 months ago
- ☆54Sep 6, 2025Updated 11 months ago
- Autonomous Agents (LLMs) research papers. Updated Daily.☆1,366Jun 24, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR☆21Apr 7, 2026Updated 4 months ago
- Automated pipeline for generating, verifying, and preparing high-quality function calling datasets for model fine-tuning. Reduces manual …☆34Jul 24, 2026Updated 2 weeks ago
- Official implementation of the paper "Chain-of-Experts: When LLMs Meet Complex Operation Research Problems"☆121Feb 6, 2026Updated 6 months ago
- ☆1,864Jun 18, 2026Updated last month
- ☆61May 21, 2025Updated last year
- Quality Diversity through Human Feedback: Towards Open-Ended Diversity-Driven Optimization (ICML 2024)☆20Apr 6, 2025Updated last year
- [AAAI 2026] Benchmarking Language Model Agents in Algorithm Search for Combinatorial Optimization☆53May 17, 2026Updated 2 months ago
- Code to reproduce key results accompanying "SAEs (usually) Transfer Between Base and Chat Models"☆13Jul 18, 2024Updated 2 years ago
- A compilation of the best multi-agent papers☆1,645Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official code for paper "SPA-RL: Reinforcing LLM Agent via Stepwise Progress Attribution"☆92Sep 13, 2025Updated 11 months ago
- ☆166Jan 21, 2025Updated last year
- [ICML 2025] Improving Planning of Agents for Long-Horizon Tasks☆45Oct 2, 2025Updated 10 months ago
- An interactive playbook for learning Claude Code through source analysis, architecture breakdowns, and guided code-reading paths.☆18Apr 1, 2026Updated 4 months ago
- Sys2Bench is a benchmarking suite designed to evaluate reasoning and planning capabilities of large language models across algorithmic, l…☆31Mar 5, 2025Updated last year
- Agent-R1: Training Powerful LLM Agents with End-to-End Reinforcement Learning☆1,610Updated this week
- This is the repo of developing reasoning models in the specific domain of financial, aim to enhance models capabilities in handling finan…☆79Jun 23, 2025Updated last year