Repo of Paper: delta-Mem: Efficient Online Memory for Large Language Models
☆55May 27, 2026Updated 4 months ago
Alternatives and similar repositories for delta-Mem
Users that are interested in delta-Mem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repo of the paper: $delta$-mem: Efficient Online Memory for Large Language Models☆263Sep 25, 2026Updated last week
- Python toolkit for MinT, the open infrastructure for experiential intelligence and LoRA RL.☆78Sep 21, 2026Updated last week
- Official repo for PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks☆17Feb 17, 2026Updated 7 months ago
- Dynamic weight generation for recursive transformers via input-conditioned LoRA modulation☆36Apr 2, 2026Updated 6 months ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆47Jul 24, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Repo for Error-Free Linear Attention is a Free Lunch: Exact Solution from Continuous-Time Dynamics☆77Sep 25, 2026Updated last week
- Object-storage-native KV cache for LLM inference & RL. Cross-restart, cross-conversation, cross-engine via shared S3 bucket.☆19Aug 10, 2026Updated last month
- [🏆ECCV'26] Official Repo for SlowBA: An efficiency backdoor attack towards VLM-based GUI agents☆20Sep 14, 2026Updated 2 weeks ago
- mobile handoff for Codex TUI sessions☆21Sep 24, 2026Updated last week
- ☆13Aug 12, 2026Updated last month
- a fast implementation of BM25☆10Sep 15, 2022Updated 4 years ago
- An LLM inference engine, written in C++☆20Mar 30, 2026Updated 6 months ago
- [COLM 2025] Code for Paper: Learning Adaptive Parallel Reasoning with Language Models☆144Dec 17, 2025Updated 9 months ago
- Open MinT training runtime on AReaL☆27Jul 1, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures☆35Jan 29, 2026Updated 8 months ago
- ☆14Mar 11, 2025Updated last year
- A Multi-Agent Approach Integrating Socratic Guidance for Automated Prompt Optimization☆18Dec 15, 2025Updated 9 months ago
- ☆15Feb 26, 2026Updated 7 months ago
- The paper list of "Memory in the Age of AI Agents: A Survey"☆14Dec 19, 2025Updated 9 months ago
- Code and Data for Evaluating the Evaluators☆16Aug 20, 2025Updated last year
- Service-aware KV-cache compression for bandwidth-efficient disaggregated LLM serving.☆46Updated this week
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆23Jun 2, 2026Updated 4 months ago
- pip install continualcode☆48Feb 10, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The repo for SHINE: A Scalable In-Context Hypernetwork for Mapping Context to LoRA in a Single Pass☆117Aug 18, 2026Updated last month
- Open format for specifying structured assumptions and requirements about code.☆93Sep 11, 2026Updated 3 weeks ago
- some tutorials for blog: simonjisu.github.io☆23Mar 25, 2021Updated 5 years ago
- Code repo for paper: Effective Strategies for Asynchronous Software Engineering Agents☆73Apr 2, 2026Updated 6 months ago
- ☆33Oct 15, 2025Updated 11 months ago
- Sniff: Misalignment detection in Vibe Coding loops☆23Jul 30, 2025Updated last year
- [NeurIPS 2024 poster] Cross-model Control: Improving Multiple Large Language Models in One-time Training☆15Oct 25, 2024Updated last year
- Orchard-Agentic is a collection of open-source work on agentic modeling☆535Updated this week
- An implementation of torchngp + semantic-nerf☆13Sep 10, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official Project Page for HLA: Higher-order Linear Attention (https://arxiv.org/abs/2510.27258)☆105Jun 15, 2026Updated 3 months ago
- Materials for "Multi-property Steering of Large Language Models with Dynamic Activation Composition"☆14Nov 22, 2024Updated last year
- Model merging is a highly efficient approach for long-to-short reasoning.☆104Oct 15, 2025Updated 11 months ago
- MCP Server for Ghidra. Exposes tools to be used by AI-powered reverse engineers.☆17Mar 29, 2025Updated last year
- Self-hosted file/code/media sharing website☆14Aug 8, 2026Updated last month
- 💦 A codebase for data-driven hydrological time-series forecasting, with official implementation of FloodDAN.☆20Mar 8, 2023Updated 3 years ago
- The implementation of paper "LLM Critics Help Catch Bugs in Mathematics: Towards a Better Mathematical Verifier with Natural Language Fee…☆38Jul 25, 2024Updated 2 years ago