An experiment that applies Google Research's `ReasoningBank` technique to Small Language Models. This experiment hopes to show that the same gains from the ReasoningBank paper also applies to much smaller, less capable models.
☆108Oct 14, 2025Updated 9 months ago
Alternatives and similar repositories for reasoning-bank-slm
Users that are interested in reasoning-bank-slm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ReasoningBank Implementation based on Google Paper https://arxiv.org/abs/2509.25140☆17Oct 19, 2025Updated 9 months ago
- An AI agent memory framework that converts an agent’s own interaction traces—both successes and failures—into reusable, high-level reason…☆62Feb 9, 2026Updated 5 months ago
- ☆12May 30, 2025Updated last year
- ☆16Feb 24, 2025Updated last year
- Control Civilization VI using natural voice commands. You make the strategy — the agent executes it.☆28May 17, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [AAAI'25] SPRING: Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models☆26Sep 24, 2025Updated 10 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- ☆18Jan 17, 2024Updated 2 years ago
- ☆10Jul 4, 2024Updated 2 years ago
- ☆24Dec 6, 2025Updated 7 months ago
- Our EMNLP 2022 paper on VIP-Based Prompting for Parameter-Efficient Learning☆10Oct 22, 2022Updated 3 years ago
- (ICML 2026) Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search☆49May 1, 2026Updated 3 months ago
- Accurate and fast KV cache compression with a gating mechanism☆28Jul 27, 2026Updated last week
- Bloat Free, Portable and Lightweight LLM Frontend (Single HTML file). With Lorebook, Web Search, Macro Engine etc.☆22Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Container-free RL framework for training software engineering agents☆72Jun 24, 2026Updated last month
- ☆35Feb 8, 2026Updated 5 months ago
- Self-Hinting Language Models Enhance Reinforcement Learning☆27Mar 28, 2026Updated 4 months ago
- ☆16Jun 1, 2025Updated last year
- ☆26Feb 20, 2026Updated 5 months ago
- ☆20Jul 4, 2025Updated last year
- [ICML 2026] InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem☆28Jun 21, 2026Updated last month
- This repositorie es the code of the paper Optimizing Reusable Knowledge for Continual Learning via Metalearning.☆11Oct 12, 2021Updated 4 years ago
- AWM: Agent Workflow Memory☆450Dec 22, 2025Updated 7 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"☆19Feb 20, 2025Updated last year
- Official implementation of PolySkill, a framework that enables web agents to learn generalizable and compositional skills through polymor…☆16Jul 6, 2026Updated 3 weeks ago
- An AI agent to use Ghidra with any AI.☆28Mar 31, 2025Updated last year
- The official code of FineRMoE.☆21Mar 17, 2026Updated 4 months ago
- ☆12Sep 28, 2023Updated 2 years ago
- [ICCV 2023] Black Box Few-Shot Adaptation for Vision-Language models☆27May 14, 2024Updated 2 years ago
- ☆17Dec 8, 2023Updated 2 years ago
- Some thoughts about writing scientific papers☆23Nov 8, 2024Updated last year
- Interactive Jacobian-Lens visualizer and live steerer for GGUF models on llama.cpp☆80Jul 12, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official implementation of Language Models as Compilers: Simulating the Execution Of Pseudocode Improves Algorithmic Reasoning in Languag…☆23Apr 8, 2024Updated 2 years ago
- About Official PyTorch implementation of "Query-Efficient Black-Box Red Teaming via Bayesian Optimization" (ACL'23)☆15Jul 9, 2023Updated 3 years ago
- Local modular AI assistant with speech, vision, and robotics support. Uses Qwen3-VL-4B in LM Studio.☆53Jan 9, 2026Updated 6 months ago
- A Comprehensive Speech Processing Algorithms Library for research and production use☆18Oct 25, 2025Updated 9 months ago
- Using PCA, Autoencoder and Fisher linear discriminant to extract the effective representations from the face images. Do the reconstructio…☆12Apr 23, 2019Updated 7 years ago
- MemSyco-Bench: Benchmarking Sycophancy in Agent Memory☆17Updated this week
- Reinforcement Learning from Text Feedback☆48Feb 17, 2026Updated 5 months ago