Membenchmark repository
☆57Nov 27, 2025Updated 8 months ago
Alternatives and similar repositories for Membench
Users that are interested in Membench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Comprehensive Library for Memory of LLM-based Agents.☆113May 13, 2025Updated last year
- The official repository for "MemSim: A Bayesian Simulator for Evaluating Memory of LLM-based Personal Assistants".☆17Oct 10, 2024Updated last year
- Code for MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems☆89Jun 27, 2026Updated last month
- Benchmarking Chat Assistants on Long-Term Interactive Memory (ICLR 2025)☆1,001May 11, 2026Updated 3 months ago
- ☆1,095Aug 13, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official repository of DialSim☆33Oct 31, 2025Updated 9 months ago
- An LLM leaderboard for stateful agents☆21Oct 16, 2025Updated 9 months ago
- The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"☆15May 18, 2025Updated last year
- This respository is used for time reasoning task for mult-session dialogue system.☆17Feb 7, 2026Updated 6 months ago
- Mem-T: Densifying Rewards for Long-Horizon Memory Agents☆39Mar 22, 2026Updated 4 months ago
- Source code and demo for memory bank and SiliconFriend☆444May 24, 2023Updated 3 years ago
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆221Dec 25, 2025Updated 7 months ago
- ☆328Jan 3, 2026Updated 7 months ago
- ☆47Apr 7, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A novel system that unifies LLM serving with query optimization to efficiently process batch agentic workflows.☆16Jun 14, 2026Updated last month
- The code for NeurIPS 2025 paper "A-Mem: Agentic Memory for LLM Agents"☆941Mar 5, 2026Updated 5 months ago
- ☆34Oct 13, 2025Updated 10 months ago
- ☆25Apr 3, 2026Updated 4 months ago
- A Multi-domain Benchmark for Personalized Search Evaluation☆12Sep 7, 2023Updated 2 years ago
- Benchmarking Long-Term Memory for AI Clones☆30Apr 7, 2026Updated 4 months ago
- ☆234Dec 20, 2024Updated last year
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆44Apr 13, 2026Updated 4 months ago
- [ICML 2025] Official repository for paper "OR-Bench: An Over-Refusal Benchmark for Large Language Models"☆31Mar 4, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A minimalist MVP demonstrating a simple yet profound insight: aligning AI memory with human episodic memory granularity. Shows how this s…☆209Apr 16, 2026Updated 3 months ago
- Code repo for "LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners"☆96May 30, 2025Updated last year
- [NeurIPS 2025 D&B (Spotlight🌟)] TIME: A Multi-level Benchmark for Temporal Reasoning of LLMs in Real-World Scenario☆32Oct 5, 2025Updated 10 months ago
- ☆506Jul 28, 2025Updated last year
- Evaluate your agent memory on real-world dialogues, not LLM-simulated dialogues.☆48Jul 3, 2025Updated last year
- ☆14Aug 7, 2024Updated 2 years ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆11Sep 4, 2025Updated 11 months ago
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆152Apr 30, 2026Updated 3 months ago
- ☆11May 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Jun 5, 2024Updated 2 years ago
- [ICML'26] MemEvolve & EvolveLab☆259May 5, 2026Updated 3 months ago
- LLMs + Persona-Plug = Personalized LLMs☆16Oct 16, 2024Updated last year
- Official Implementation of "Democratizing Large Language Models via Personalized Parameter-Efficient Fine-tuning" at EMNLP 2024 Main Conf…☆51Jul 31, 2025Updated last year
- This repository introduce a comprehensive paper list, datasets, methods and tools for memory research.☆357Dec 29, 2025Updated 7 months ago
- A MemAgent framework that can be extrapolated to 3.5M, along with a training framework for RL training of any agent workflow.☆1,093May 12, 2026Updated 3 months ago
- Implementation of "RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm"☆18Apr 11, 2025Updated last year