Open source code for ICLR 2026 Paper: Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
☆451Aug 20, 2026Updated 3 weeks ago
Alternatives and similar repositories for MemoryAgentBench
Users that are interested in MemoryAgentBench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆1,166Aug 13, 2024Updated 2 years ago
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆227Dec 25, 2025Updated 8 months ago
- Benchmarking Chat Assistants on Long-Term Interactive Memory (ICLR 2025)☆1,083May 11, 2026Updated 4 months ago
- Membenchmark repository☆59Nov 27, 2025Updated 9 months ago
- ☆335Jan 3, 2026Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Agent Memory Benchmark☆80Sep 2, 2026Updated last week
- ☆60Jun 1, 2026Updated 3 months ago
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,105May 12, 2026Updated 4 months ago
- [ICML 26] An evaluation framework assessing long-context retention and long-horizon memory performance for agentic applications (AMA-benc…☆80Jun 15, 2026Updated 2 months ago
- The code for NeurIPS 2025 paper "A-Mem: Agentic Memory for LLM Agents"☆964Mar 5, 2026Updated 6 months ago
- [ICLR 2026] LightMem: Lightweight and Efficient Memory-Augmented Generation☆1,140Sep 5, 2026Updated last week
- The paper list of "Memory in the Age of AI Agents: A Survey"☆2,377Mar 4, 2026Updated 6 months ago
- Benchmarking LLMs in Real-World Memory-Driven Interaction☆50Apr 7, 2026Updated 5 months ago
- [ICML'26] MemEvolve & EvolveLab☆266May 5, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- EvaLearn is a pioneering benchmark designed to evaluate large language models (LLMs) on their learning capability and efficiency in chall…☆431May 12, 2026Updated 4 months ago
- ☆37Feb 13, 2026Updated 7 months ago
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆160Sep 3, 2026Updated last week
- ☆277Apr 29, 2025Updated last year
- Mem-T: Densifying Rewards for Long-Horizon Memory Agents☆43Mar 22, 2026Updated 5 months ago
- A-MEM: Agentic Memory for LLM Agents☆1,176Dec 12, 2025Updated 9 months ago
- MemGen: Weaving Generative Latent Memory for Self-Evolving Agents☆415Jun 10, 2026Updated 3 months ago
- ☆506Jul 28, 2025Updated last year
- The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"☆16May 18, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A-MEM: Agentic Memory for LLM Agents☆394Mar 15, 2026Updated 5 months ago
- [ICML 2025] "SepLLM: Accelerate Large Language Models by Compressing One Segment into One Separator"☆572Jul 29, 2025Updated last year
- ☆90Jul 23, 2024Updated 2 years ago
- ☆343Jul 4, 2025Updated last year
- ☆590Oct 11, 2025Updated 11 months ago
- [ICML 2025] A pytorch implementation of the paper "TreeLoRA: Efficient Continual Learning via Layer-Wise LoRAs Guided by a Hierarchical G…☆350Dec 15, 2025Updated 8 months ago
- GENERanno: A Genomic Foundation Model for Metagenomic Annotation☆317Jun 15, 2026Updated 2 months ago
- [Up-To-Date] [TMLR 2026] Awesome Agent Memory Paper Resource☆229Jul 23, 2026Updated last month
- ☆250Feb 11, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Code for MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems☆91Updated this week
- Neobanker FinTalk-AI: A Grounded Orchestration Framework for Multi-Agent Collaboration on Financial Tasks Leveraging the OSWorld Environm…☆39Aug 23, 2026Updated 3 weeks ago
- Unified benchmark for evaluating conversational memory and RAG across multiple datasets☆314Aug 24, 2026Updated 2 weeks ago
- 自然语言处理、深度学习、机器学习的一些个人博客☆35Nov 19, 2022Updated 3 years ago
- Mirix is a multi-agent personal assistant designed to track on-screen activities and answer user questions intelligently. By capturing re…☆3,441Updated this week
- 以jax为后端的类似keras的框架☆98Jan 13, 2023Updated 3 years ago
- AI Group is a powerful mobile intelligent assistant application that integrates multiple large language models (LLMs) and AI services, pr…☆1,101Sep 10, 2025Updated last year