The official repository of paper: MemPO: Self-Memory Policy Optimization for Long-Horizon Agents
☆26Apr 10, 2026Updated 4 months ago
Alternatives and similar repositories for MemPO
Users that are interested in MemPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Math24o: 高中奥林匹克数学竞赛测评集 High School Olympiad Mathematics Chinese Benchmark☆14Mar 27, 2025Updated last year
- Design for Error Detection in Deep-Research Agents Trajectories.☆22Jun 4, 2026Updated 2 months ago
- ☆34Jan 26, 2026Updated 6 months ago
- [ICLR 2026] Adaptive Social Learning via Mode Policy Optimization for Language Agents☆51Feb 2, 2026Updated 6 months ago
- Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner☆31Jun 27, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆28Mar 31, 2026Updated 4 months ago
- [COLING 2024 (Oral)] PromISe:Releasing the Capabilities of LLMs with Prompt Introspective Search☆23Aug 26, 2024Updated 2 years ago
- Mem-T: Densifying Rewards for Long-Horizon Memory Agents☆42Mar 22, 2026Updated 5 months ago
- [arXiv:2605.19952] "Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory"☆17May 20, 2026Updated 3 months ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 4 months ago
- [CVPR 2026 Highlight] DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding☆19Jun 4, 2026Updated 2 months ago
- LLM-Check: Investigating Detection of Hallucinations in Large Language Models (NeurIPS 2024)☆40Dec 8, 2024Updated last year
- Marathon: A Multiple-choice Long Context Evaluation Benchmark for Large Language Models.☆10May 16, 2024Updated 2 years ago
- PyTorch code for NeurIPSW 2020 paper (4th Workshop on Meta-Learning) "Few-Shot Unsupervised Continual Learning through Meta-Examples"☆20Nov 2, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- ☆11Jul 14, 2023Updated 3 years ago
- Benchmarking Social Intelligence of Language Agents through Interactive Scenarios☆13Jan 4, 2025Updated last year
- Code for AttriBoT from "AttriBoT: A Bag of Tricks for Efficiently Approximating Leave-One-Out Context Attribution"☆15Apr 21, 2025Updated last year
- Code for Linguistic Structure Guided Context Modeling for Referring Image Segmentation, ECCV2020.☆16Oct 2, 2020Updated 5 years ago
- GisPy: A Tool for Measuring Gist Inference Score in Text https://aclanthology.org/2022.wnu-1.5/☆13Jul 1, 2024Updated 2 years ago
- ☆18Jan 9, 2025Updated last year
- MedSafetyBench: Evaluating and Improving the Medical Safety of LLMs, NeurIPS 2024☆50Dec 4, 2025Updated 8 months ago
- Official implementation for "Mixture of In-Context Experts Enhance LLMs’ Awareness of Long Contexts" (Accepted by Neurips2024)☆14Jan 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Graph-based experience memory for LLM reward prediction with limited labels. 20% labels → 97.3% Oracle.☆19Mar 24, 2026Updated 5 months ago
- a brief repo about paper research☆15Sep 4, 2024Updated last year
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆51May 12, 2026Updated 3 months ago
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆224Dec 25, 2025Updated 8 months ago
- ☆10Mar 24, 2023Updated 3 years ago
- A supervised fine-tuning method for controllable reasoning length in large language models (一种通过有监督微调实现大语言模型思考长度可控的方法)☆11May 8, 2025Updated last year
- The guideline for pod.☆10Jun 19, 2020Updated 6 years ago
- [EMNLP 2025🔥] UNComp: Can Matrix Entropy Uncover Sparsity? -- A Compressor Design from an Uncertainty-Aware Perspective☆20Jan 7, 2026Updated 7 months ago
- create timer videos at any speed.☆15Sep 25, 2023Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆331Jan 3, 2026Updated 7 months ago
- ☆19Nov 24, 2025Updated 9 months ago
- Dimentionality reduction framework with autoencoders for mineral exploration☆18Feb 23, 2025Updated last year
- ☆48Nov 26, 2025Updated 9 months ago
- BigBang-Proton is a LLM pretrained on cross-scale, cross-structure, cross-discipline real-world scientific tasks to construct a scienti…☆21Nov 8, 2025Updated 9 months ago
- ☆10Dec 6, 2019Updated 6 years ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 4 months ago