The code for paper "EPO: Entropy-regularized Policy Optimization for LLM Agents Reinforcement Learning"
☆40Jul 13, 2026Updated last month
Alternatives and similar repositories for EPO
Users that are interested in EPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆21Jun 2, 2026Updated 3 months ago
- ☆33May 9, 2025Updated last year
- [ICLR 2026] RPG: KL-Regularized Policy Gradient (https://arxiv.org/abs/2505.17508)☆76Jun 29, 2026Updated 2 months ago
- From Commands to Prompts: LLM-based Semantic File System for AIOS☆55Mar 9, 2025Updated last year
- MemEye: A Visual-Centric Evaluation Framework for Multimodal Agent Memory