A collection of lightweight interpretability scripts to understand how LLMs think
☆91Mar 18, 2026Updated 4 months ago
Alternatives and similar repositories for llm-interp
Users that are interested in llm-interp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Negation Neglect☆16May 22, 2026Updated 2 months ago
- H-Net Dynamic Hierarchical Architecture☆81Sep 11, 2025Updated 10 months ago
- smolLM with Entropix sampler on pytorch☆148Oct 31, 2024Updated last year
- ADAG: Transluce's MLP neuron-level circuit tracing library☆37Apr 10, 2026Updated 4 months ago
- Official PyTorch implementation for "TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors" [ACL 2026]☆48Apr 14, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for NLP project. Name to be changed when we decide on a project☆16Apr 19, 2022Updated 4 years ago
- MLX implementation of Hierarchical Reasoning Model (HRM) - Adaptive computation for complex reasoning tasks☆29Aug 27, 2025Updated 11 months ago
- ☆21Jun 12, 2025Updated last year
- As barebones as you can get the GPT, now accelerated on Macs thanks to tinygrad.☆16Oct 22, 2025Updated 9 months ago
- ☆10Nov 18, 2024Updated last year
- Training LLMs to Report Their Learned Behaviors☆28Apr 28, 2026Updated 3 months ago
- Evolution Pretraining Fully in Int Formats☆179Feb 25, 2026Updated 5 months ago
- Training framework with a goal to explore the frontier of sample efficiency of small language models☆101Jan 25, 2026Updated 6 months ago
- ☆19Nov 24, 2025Updated 8 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 3 months ago
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 10 months ago
- ☆63Jul 10, 2025Updated last year
- A lightweight, user-friendly data-plane for LLM training.☆40Sep 10, 2025Updated 11 months ago
- Streamline on-policy/off-policy distillation workflows in a few lines of code☆109Updated this week
- ☆15Apr 26, 2025Updated last year
- X Developer Challenge☆12Apr 25, 2024Updated 2 years ago
- Transformers for Mathematics Tutorial | Simons/SLMath Workshop on AI for Mathematics 2025☆43Jan 18, 2026Updated 6 months ago
- Stack based programming language in C☆20Apr 12, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for "Training Language Models To Explain Their Own Computations"☆32Jul 7, 2026Updated last month
- Official implementation for DenseMixer: Improving MoE Post-Training with Precise Router Gradient☆68Aug 3, 2025Updated last year
- [ICML 2025] CommVQ: Commutative Vector Quantization for KV Cache Compression☆28Sep 2, 2025Updated 11 months ago
- ☆17Feb 23, 2026Updated 5 months ago
- A Python reimplementation + extension of "Planning with Large Language Models for Code Generation" (https://arxiv.org/abs/2303.05510)☆17Dec 1, 2023Updated 2 years ago
- alternative way to calculating self attention☆18May 25, 2024Updated 2 years ago
- A collection of various llm pruning implementations, training code for GPUs & TPUs, and evaluation script.☆69Apr 20, 2026Updated 3 months ago
- Open source interpretability artefacts for R1.☆183Apr 21, 2025Updated last year
- Engine for collecting, uploading, and downloading model activations☆30Apr 2, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Entropy Based Sampling and Parallel CoT Decoding☆17Oct 9, 2024Updated last year
- A toolkit for embedding text datasets with sparse autoencoders☆30Mar 24, 2026Updated 4 months ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆20Jun 2, 2026Updated 2 months ago
- Formalization of Arithmetization of Mathematics/Metamathematics☆14Mar 8, 2025Updated last year
- Low-Rank Llama Custom Training☆23Mar 27, 2024Updated 2 years ago
- A graph visualization of attention☆56May 20, 2025Updated last year
- Course on Flash-attention in Triton☆100Feb 9, 2026Updated 6 months ago