[ACL'25] We propose a novel fine-tuning method, Separate Memory and Reasoning, which combines prompt tuning with LoRA.
☆89Nov 2, 2025Updated 10 months ago
Alternatives and similar repositories for Disentangling-Memory-and-Reasoning
Users that are interested in Disentangling-Memory-and-Reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL'24] Chain of Thought (CoT) is significant in improving the reasoning abilities of large language models (LLMs). However, the correla…☆47May 11, 2025Updated last year
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆27Jul 21, 2025Updated last year
- [KDD Explore'24]Time Series Forecasting with LLMs: Understanding and Enhancing Model Capabilities☆17May 7, 2025Updated last year
- [ICLR 2026] The official code for "Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models"☆31Feb 7, 2026Updated 7 months ago
- [ACL 2025] The official code for "AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection".☆45Aug 12, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆22Oct 12, 2024Updated last year
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆23Jun 2, 2026Updated 3 months ago
- Official code for Guiding Language Model Math Reasoning with Planning Tokens☆21Feb 29, 2024Updated 2 years ago
- EmojiCrypt: Prompt Encryption for Secure Communication with Large Language Models☆27Feb 21, 2024Updated 2 years ago
- [preprint] sparsity☆23Jul 26, 2026Updated last month
- [ICML 2026] Heima☆77May 20, 2026Updated 4 months ago
- TrustAgent: Towards Safe and Trustworthy LLM-based Agents☆61Feb 7, 2025Updated last year
- ☆31Feb 10, 2025Updated last year
- ☆23Sep 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆20Mar 25, 2026Updated 5 months ago
- HaluMem is the first operation level hallucination evaluation benchmark tailored to agent memory systems.☆163Sep 3, 2026Updated 2 weeks ago
- Demonstration Agents for AIOS☆19Dec 25, 2024Updated last year
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆32Aug 4, 2026Updated last month
- Code for "In-Context Former: Lightning-fast Compressing Context for Large Language Model" (Findings of EMNLP 2024)☆21Nov 21, 2024Updated last year
- Official implementation of Latent-SFT: teaching LLMs to reason with vocabulary-space latent chains.☆60Aug 31, 2026Updated 3 weeks ago
- ☆12Sep 23, 2023Updated 3 years ago
- This repository contains the code for the paper: SirLLM: Streaming Infinite Retentive LLM☆60May 28, 2024Updated 2 years ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆33Jun 30, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆23Dec 17, 2024Updated last year
- ☆20Nov 13, 2025Updated 10 months ago
- ☆16Jul 23, 2024Updated 2 years ago
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- The official implementation of the paper "Mem-α: Learning Memory Construction via Reinforcement Learning"☆228Dec 25, 2025Updated 8 months ago
- [COLM 2024] JailBreakV-28K: A comprehensive benchmark designed to evaluate the transferability of LLM jailbreak attacks to MLLMs, and fur…☆99May 9, 2025Updated last year
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"☆11Jan 10, 2025Updated last year
- ☆49Mar 15, 2025Updated last year
- [ICLR 2026] Geometric-Mean Policy Optimization☆105Jan 26, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 2025] RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios☆31Jul 31, 2026Updated last month
- ☆90Sep 11, 2024Updated 2 years ago
- From Commands to Prompts: LLM-based Semantic File System for AIOS☆55Mar 9, 2025Updated last year
- Repository for NPHardEval, a quantified-dynamic benchmark of LLMs☆66Mar 26, 2024Updated 2 years ago
- The code for "AttentionPredictor: Temporal Pattern Matters for Efficient LLM Inference", Qingyue Yang, Jie Wang, Xing Li, Zhihai Wang, Ch…☆29Jul 15, 2025Updated last year
- [arxiv: 2503.23895] Dynamic Parametric Retrieval Augmented Generation for Test-time Knowledge Enhancement☆184Aug 14, 2025Updated last year
- Official code and data repository of MathChat: MathChat: Benchmarking Mathematical Reasoning and Instruction Following in Multi-Turn Inte…☆22Jun 3, 2024Updated 2 years ago