KV Cache Steering for Controlling Frozen LLMs
☆52Aug 18, 2026Updated last month
Alternatives and similar repositories for cache-steering
Users that are interested in cache-steering are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Bitune: Bidirectional Instruction-Tuning☆27Jun 19, 2025Updated last year
- This is a repository that implements the Dense NN Retrieval Evaluation used for evaluating the In-Context Learning Capabilities of Vision…☆32Nov 3, 2025Updated 10 months ago
- ☆35May 24, 2024Updated 2 years ago
- "Near, far: Patch-ordering enhances vision foundation models' scene understanding": A New SSL Post-Training Approach for Improving DINOv2…☆34Apr 20, 2025Updated last year
- Official implementation of "SentenceKV: Efficient LLM Inference via Sentence-Level Semantic KV Caching" (COLM 2025). A novel KV cache com…☆15Sep 29, 2025Updated 11 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- This repo contains the official implementation of ICCV 2025 paper "MoSiC: Optimal-Transport Motion Trajectory for Dense Self-Supervised L…☆23Sep 12, 2025Updated last year
- transformer layers behavior as painters🧑🎨☆15May 6, 2025Updated last year
- This is the official code for OThink-R1 project.☆21Jun 19, 2025Updated last year
- Learning to Count without Annotations☆24May 24, 2024Updated 2 years ago
- ☆17Jun 10, 2025Updated last year
- ☆15Dec 12, 2024Updated last year
- ☆18Aug 4, 2025Updated last year
- The official code implementation for paper "PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs"☆30May 24, 2025Updated last year
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- TTRV: Test-Time Reinforcement Learning for Vision–Language Models (CVPR 2026)☆48Mar 8, 2026Updated 6 months ago
- Official repo for paper: Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs☆20Nov 26, 2025Updated 10 months ago
- Reasoning Activation in LLMs via Small Model Transfer (NeurIPS 2025)☆22Oct 16, 2025Updated 11 months ago
- ☆30Sep 11, 2026Updated 2 weeks ago
- ☆17Feb 4, 2026Updated 7 months ago
- Resources related to the model cards for ML☆11Mar 16, 2021Updated 5 years ago
- ROS2 catestian_impedance_controller from PdZ☆13Oct 22, 2025Updated 11 months ago
- Starter notebook and utilities for the Clevr-4 dataset☆18Nov 1, 2023Updated 2 years ago
- Fork of Flame repo for training of some new stuff in development☆20Aug 27, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Apr 11, 2024Updated 2 years ago
- Official Code for "Learning to Reason via Mixture-of-Thought for Logical Reasoning"☆33Nov 20, 2025Updated 10 months ago
- 这是我的博客《不用框架,使用Python搭建基于numpy的卷积神经网络来进行cifar-10分类的深度学习系统》的代码实现。☆10Jul 1, 2019Updated 7 years ago
- Jet tagging in the Lund plane with graph networks☆10Dec 16, 2021Updated 4 years ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆28Sep 20, 2026Updated last week
- ☆13Jul 2, 2025Updated last year
- ☆13Apr 17, 2025Updated last year
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆17Mar 2, 2026Updated 6 months ago
- [CVPR 2023] SeSDF: Self-evolved Signed Distance Field for Implicit 3D Clothed Human Reconstruction☆17Apr 6, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Feb 16, 2024Updated 2 years ago
- The official implementation of the paper "A Dual-Space Framework for General Knowledge Distillation of Large Language Models".☆17Jan 4, 2026Updated 8 months ago
- ☆19Aug 16, 2025Updated last year
- The official repository of "Document Image Machine Translation with Dynamic Multi-pre-trained Models Assembling"☆14Nov 26, 2025Updated 10 months ago
- ☆15Apr 6, 2026Updated 5 months ago
- [COLM 2025: 1st Workshop on the Application of LLM Explainability to Reasoning and Planning] Latent Chain-of-Thought? Decoding the Depth-…☆22Aug 19, 2026Updated last month
- Self-Hinting Language Models Enhance Reinforcement Learning☆28Mar 28, 2026Updated 5 months ago