An experiment that applies Google Research's `ReasoningBank` technique to Small Language Models. This experiment hopes to show that the same gains from the ReasoningBank paper also applies to much smaller, less capable models.
☆108Oct 14, 2025Updated 11 months ago
Alternatives and similar repositories for reasoning-bank-slm
Users that are interested in reasoning-bank-slm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ReasoningBank Implementation based on Google Paper https://arxiv.org/abs/2509.25140☆17Oct 19, 2025Updated 11 months ago
- An AI agent memory framework that converts an agent’s own interaction traces—both successes and failures—into reusable, high-level reason…☆65Feb 9, 2026Updated 8 months ago
- ☆12May 30, 2025Updated last year
- A PyTorch implementation of gradient-free optimization for directly optimizing NDCG (Normalized Discounted Cumulative Gain) in neural inf…☆19Dec 21, 2025Updated 9 months ago
- Control Civilization VI using natural voice commands. You make the strategy — the agent executes it.☆30May 17, 2026Updated 4 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- HiAgent: Hierarchical Working Memory Management for Solving Long-Horizon Agent Tasks with Large Language Model☆76Apr 15, 2026Updated 5 months ago
- Code for "FactKB: Generalizable Factuality Evaluation using Language Models Enhanced with Factual Knowledge". EMNLP 2023.☆20Dec 25, 2023Updated 2 years ago
- A lightweight chat interface for interacting with local models, featuring persistent memory using a seamless SQLite database to store you…☆34Sep 15, 2025Updated last year
- ☆10Jul 4, 2024Updated 2 years ago
- ☆24Dec 6, 2025Updated 10 months ago
- Our EMNLP 2022 paper on VIP-Based Prompting for Parameter-Efficient Learning☆10Oct 22, 2022Updated 3 years ago
- Developing K - a language model to generate OPENSCAD code from prompt☆19Dec 3, 2025Updated 10 months ago
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆50Aug 7, 2026Updated 2 months ago
- Real-time offline speech-to-text transcription script on macOS using parakeet-mlx☆17Oct 28, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Bloat Free, Portable and Lightweight LLM Frontend (Single HTML file). With Lorebook, Web Search, Macro Engine etc.☆22Sep 24, 2026Updated 2 weeks ago
- Accurate and fast KV cache compression with a gating mechanism☆31Jul 27, 2026Updated 2 months ago
- Code and data for the paper: On the Reliability of Psychological Scales on Large Language Models☆31Dec 15, 2025Updated 9 months ago
- ☆35Feb 8, 2026Updated 8 months ago
- Self-Hinting Language Models Enhance Reinforcement Learning☆28Mar 28, 2026Updated 6 months ago
- This repositorie es the code of the paper Optimizing Reusable Knowledge for Continual Learning via Metalearning.☆11Oct 12, 2021Updated 4 years ago
- AWM: Agent Workflow Memory☆477Dec 22, 2025Updated 9 months ago
- ☆28Feb 20, 2026Updated 7 months ago
- Official implementation of PolySkill, a framework that enables web agents to learn generalizable and compositional skills through polymor…☆19Jul 6, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV 2023] Black Box Few-Shot Adaptation for Vision-Language models☆28May 14, 2024Updated 2 years ago
- ☆11Oct 15, 2024Updated last year
- ☆14Aug 30, 2023Updated 3 years ago
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cpp☆89Jul 12, 2026Updated 2 months ago
- Some thoughts about writing scientific papers☆23Nov 8, 2024Updated last year
- Local modular AI assistant with speech, vision, and robotics support. Uses Qwen3-VL-4B in LM Studio.☆54Jan 9, 2026Updated 9 months ago
- Reinforcement Learning from Text Feedback☆49Feb 17, 2026Updated 7 months ago
- SCoRe: Training Language Models to Self-Correct via Reinforcement Learning☆16May 14, 2026Updated 4 months ago
- A PyTorch native platform for training generative AI models☆17Jun 30, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A realtime speech to text diarization system to gather and interleave speech from multiple speaker audio.☆61Jan 29, 2026Updated 8 months ago
- ☆50Mar 15, 2025Updated last year
- Apex: An advanced autonomous coding agent for VS Code featuring total autonomy modes, recursive chain-of-thought reasoning, council-of-cr…☆33Feb 11, 2026Updated 7 months ago
- Benchmarking long-horizon chain-of-thought reasoning.☆44Apr 20, 2026Updated 5 months ago
- 中英文信息抽取数据集整理☆20May 15, 2022Updated 4 years ago
- The official implementation of the paper "Large Scale Knowledge Washing"☆10Jun 12, 2024Updated 2 years ago
- codebase for the SIMAT dataset and evaluation☆39Feb 16, 2022Updated 4 years ago