A collection of lightweight interpretability scripts to understand how LLMs think
☆90Mar 18, 2026Updated 4 months ago
Alternatives and similar repositories for llm-interp
Users that are interested in llm-interp are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Negation Neglect☆16May 22, 2026Updated last month
- H-Net Dynamic Hierarchical Architecture☆81Sep 11, 2025Updated 10 months ago
- ADAG: Transluce's MLP neuron-level circuit tracing library☆33Apr 10, 2026Updated 3 months ago
- Official PyTorch implementation for "TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors" [ACL 2026]☆47Apr 14, 2026Updated 3 months ago
- Repository for NLP project. Name to be changed when we decide on a project☆16Apr 19, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MLX implementation of Hierarchical Reasoning Model (HRM) - Adaptive computation for complex reasoning tasks☆29Aug 27, 2025Updated 10 months ago
- ☆21Jun 12, 2025Updated last year
- PyTorch memory allocation visualizer☆76Mar 6, 2026Updated 4 months ago
- ☆10Nov 18, 2024Updated last year
- prompt engineering experiments with DSPy GEPA and TextGrad☆70Sep 2, 2025Updated 10 months ago
- Training LLMs to Report Their Learned Behaviors☆27Apr 28, 2026Updated 2 months ago
- Game of Life on a Toroidal Surface☆16Aug 5, 2025Updated 11 months ago
- Evolution Pretraining Fully in Int Formats☆177Feb 25, 2026Updated 4 months ago
- Training framework with a goal to explore the frontier of sample efficiency of small language models☆101Jan 25, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆19Nov 24, 2025Updated 7 months ago
- Training and evaluating with OpenReward☆33Apr 28, 2026Updated 2 months ago
- Stack based programming language in C☆21Apr 12, 2023Updated 3 years ago
- Efficient non-uniform quantization with GPTQ for GGUF☆64Sep 17, 2025Updated 10 months ago
- ☆63Jul 10, 2025Updated last year
- A lightweight, user-friendly data-plane for LLM training.☆40Sep 10, 2025Updated 10 months ago
- X Developer Challenge☆12Apr 25, 2024Updated 2 years ago
- ☆33Jan 7, 2025Updated last year
- Repository for "Training Language Models To Explain Their Own Computations"☆23Jul 7, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation for DenseMixer: Improving MoE Post-Training with Precise Router Gradient☆67Aug 3, 2025Updated 11 months ago
- A Python reimplementation + extension of "Planning with Large Language Models for Code Generation" (https://arxiv.org/abs/2303.05510)☆17Dec 1, 2023Updated 2 years ago
- Engine for collecting, uploading, and downloading model activations☆30Apr 2, 2025Updated last year
- A collection of various llm pruning implementations, training code for GPUs & TPUs, and evaluation script.☆69Apr 20, 2026Updated 3 months ago
- alternative way to calculating self attention☆18May 25, 2024Updated 2 years ago
- Open source interpretability artefacts for R1.☆183Apr 21, 2025Updated last year
- A toolkit for embedding text datasets with sparse autoencoders☆30Mar 24, 2026Updated 3 months ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆18Jun 2, 2026Updated last month
- Entropy Based Sampling and Parallel CoT Decoding☆17Oct 9, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Low-Rank Llama Custom Training☆23Mar 27, 2024Updated 2 years ago
- A graph visualization of attention☆56May 20, 2025Updated last year
- Repo for Paper: Discovering Interpretable Algorithms by Decompiling Transformers to RASP☆15May 25, 2026Updated last month
- Code for the experiments and websites of the paper "Same Task, Different Circuits"☆36Jun 9, 2026Updated last month
- Course on Flash-attention in Triton☆100Feb 9, 2026Updated 5 months ago
- FlashSampling: Fast and Memory-Efficient Exact Sampling (https://huggingface.co/papers/2603.15854)☆76Jun 15, 2026Updated last month
- ☆27Jan 31, 2026Updated 5 months ago