[ACL'25] We propose a novel fine-tuning method, Separate Memory and Reasoning, which combines prompt tuning with LoRA.
☆87Nov 2, 2025Updated 8 months ago
Alternatives and similar repositories for Disentangling-Memory-and-Reasoning
Users that are interested in Disentangling-Memory-and-Reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [FCS'24] LVLM Safety paper☆19Jan 4, 2025Updated last year
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆26Jul 21, 2025Updated last year
- [KDD Explore'24]Time Series Forecasting with LLMs: Understanding and Enhancing Model Capabilities☆17May 7, 2025Updated last year
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆30Feb 7, 2026Updated 5 months ago
- [ACL 2025] The official code for "AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection".☆42Aug 4, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆18Jun 2, 2026Updated last month
- ☆22Oct 12, 2024Updated last year
- Official code for Guiding Language Model Math Reasoning with Planning Tokens☆19Feb 29, 2024Updated 2 years ago
- [ICONIP'24]Mingyu.Jin's final year project☆30Aug 23, 2024Updated last year
- [preprint] sparsity☆22May 6, 2026Updated 2 months ago
- TrustAgent: Towards Safe and Trustworthy LLM-based Agents☆58Feb 7, 2025Updated last year
- ☆31Feb 10, 2025Updated last year
- ☆23Sep 2, 2025Updated 10 months ago
- VAEGAN, I Love u☆16Aug 15, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆20Mar 25, 2026Updated 3 months ago
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆148Apr 30, 2026Updated 2 months ago
- Demonstration Agents for AIOS☆19Dec 25, 2024Updated last year
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆33Feb 26, 2026Updated 4 months ago
- Code for "In-Context Former: Lightning-fast Compressing Context for Large Language Model" (Findings of EMNLP 2024)☆21Nov 21, 2024Updated last year
- Official implementation of Latent-SFT: teaching LLMs to reason with vocabulary-space latent chains.☆55May 18, 2026Updated 2 months ago
- This repository contains the code for the paper: SirLLM: Streaming Infinite Retentive LLM☆60May 28, 2024Updated 2 years ago
- The original Shared Recurrent Memory Transformer implementation☆36Jul 11, 2025Updated last year
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The code repo for the paper "Differentiable Evolutionary Reinforcement Learning"☆18Jan 6, 2026Updated 6 months ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆31Jun 30, 2025Updated last year
- ☆23Dec 17, 2024Updated last year
- ☆16Jul 23, 2024Updated 2 years ago
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- ☆43Jul 16, 2025Updated last year
- ☆151Sep 12, 2025Updated 10 months ago
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆218Dec 25, 2025Updated 6 months ago
- [COLM 2024] JailBreakV-28K: A comprehensive benchmark designed to evaluate the transferability of LLM jailbreak attacks to MLLMs, and fur…☆96May 9, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"☆11Jan 10, 2025Updated last year
- ☆48Mar 15, 2025Updated last year
- [ICLR 2026] Geometric-Mean Policy Optimization☆104Jan 26, 2026Updated 5 months ago
- ☆89Sep 11, 2024Updated last year
- From Commands to Prompts: LLM-based Semantic File System for AIOS☆55Mar 9, 2025Updated last year
- Repository for NPHardEval, a quantified-dynamic benchmark of LLMs☆64Mar 26, 2024Updated 2 years ago
- The code for "AttentionPredictor: Temporal Pattern Matters for Efficient LLM Inference", Qingyue Yang, Jie Wang, Xing Li, Zhihai Wang, Ch…☆29Jul 15, 2025Updated last year