[ACL'25] We propose a novel fine-tuning method, Separate Memory and Reasoning, which combines prompt tuning with LoRA.
☆89Nov 2, 2025Updated 10 months ago
Alternatives and similar repositories for Disentangling-Memory-and-Reasoning
Users that are interested in Disentangling-Memory-and-Reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [FCS'24] LVLM Safety paper☆19Jan 4, 2025Updated last year
- [ACL'24] Chain of Thought (CoT) is significant in improving the reasoning abilities of large language models (LLMs). However, the correla…☆47May 11, 2025Updated last year
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆27Jul 21, 2025Updated last year
- [ICML'25] Our study systematically investigates massive values in LLMs' attention mechanisms. First, we observe massive values are concen…☆87Jun 20, 2025Updated last year
- [KDD Explore'24]Time Series Forecasting with LLMs: Understanding and Enhancing Model Capabilities☆17May 7, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆30Feb 7, 2026Updated 6 months ago
- [ACL 2025] The official code for "AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection".☆44Aug 12, 2026Updated 3 weeks ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆21Jun 2, 2026Updated 3 months ago
- Official code for Guiding Language Model Math Reasoning with Planning Tokens☆21Feb 29, 2024Updated 2 years ago
- [preprint] sparsity☆23Jul 26, 2026Updated last month
- [ICML 2026] Heima☆76May 20, 2026Updated 3 months ago
- ☆31Feb 10, 2025Updated last year
- ☆23Sep 2, 2025Updated last year
- VAEGAN, I Love u☆16Aug 15, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆20Mar 25, 2026Updated 5 months ago
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆159Updated this week
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆32Aug 4, 2026Updated 3 weeks ago
- Code for "In-Context Former: Lightning-fast Compressing Context for Large Language Model" (Findings of EMNLP 2024)☆21Nov 21, 2024Updated last year
- ☆12Sep 23, 2023Updated 2 years ago
- This repository contains the code for the paper: SirLLM: Streaming Infinite Retentive LLM☆60May 28, 2024Updated 2 years ago
- The original Shared Recurrent Memory Transformer implementation☆37Aug 24, 2026Updated last week
- The code repo for the paper "Differentiable Evolutionary Reinforcement Learning"☆18Jan 6, 2026Updated 7 months ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆33Jun 30, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆23Dec 17, 2024Updated last year
- ☆16Jul 23, 2024Updated 2 years ago
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆226Dec 25, 2025Updated 8 months ago
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"☆11Jan 10, 2025Updated last year
- ☆49Mar 15, 2025Updated last year
- Repository for NPHardEval, a quantified-dynamic benchmark of LLMs☆66Mar 26, 2024Updated 2 years ago
- ☆22May 14, 2025Updated last year
- Official code and data repository of MathChat: MathChat: Benchmarking Mathematical Reasoning and Instruction Following in Multi-Turn Inte…☆22Jun 3, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Self-Questioning Language Models☆56Mar 30, 2026Updated 5 months ago
- Code for ICML 2024 paper☆34Sep 18, 2025Updated 11 months ago
- ☆51Jul 3, 2026Updated 2 months ago
- ☆80Nov 19, 2024Updated last year
- [COLING'25] Exploring Concept Depth: How Large Language Models Acquire Knowledge at Different Layers?☆82Jan 22, 2025Updated last year
- A framework to study AI models in Reasoning, Alignment, and use of Memory (RAM).☆382Jun 25, 2026Updated 2 months ago
- [NeurIPS '25] Multi-Token Prediction Needs Registers☆32Dec 14, 2025Updated 8 months ago