πͺ Interpreto is an interpretability toolbox for LLMs
β204Sep 23, 2026Updated this week
Alternatives and similar repositories for interpreto
Users that are interested in interpreto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build and train Lipschitz-constrained networks: PyTorch implementation of 1-Lipschitz layers. For TensorFlow/Keras implementation, see htβ¦β44Mar 23, 2026Updated 6 months ago
- π Influenciae is a Tensorflow Toolbox for Influence Functionsβ67Aug 24, 2026Updated last month
- π Overcomplete is a Vision-based SAE Toolboxβ152Dec 4, 2025Updated 9 months ago
- β40Sep 15, 2025Updated last year
- Build and train Lipschitz constrained networks: TensorFlow implementation of k-Lipschitz layersβ102Mar 14, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Easy-to-use MIRAGE code for faithful answer attribution in RAG applications. Paper: https://aclanthology.org/2024.emnlp-main.347/β25Mar 10, 2025Updated last year
- π CODS - Conformal Object Detection and Segmentationβ22Jul 24, 2026Updated 2 months ago
- π Xplique is a Neural Networks Explainability Toolboxβ754Updated this week
- [NeurIPS 2023] and [ICLR 2024] for robustness certification.β10Nov 30, 2024Updated last year
- β16Nov 14, 2025Updated 10 months ago
- β15Jan 2, 2023Updated 3 years ago
- Unified access to Large Language Model modules using NNsightβ120Updated this week
- π Puncc is a python library for predictive uncertainty quantification using conformal prediction.β410Updated this week
- The nnsight package enables interpreting and manipulating the internals of deep learned models.β1,108Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for "Training Language Models To Explain Their Own Computations"β37Jul 7, 2026Updated 2 months ago
- DiffuLab is designed to provide a simple and flexible way to train diffusion models while allowing full customization of its core componeβ¦β43Jan 11, 2026Updated 8 months ago
- A runway dataset and a generator of synthetic aerial images with automatic labeling.β140Sep 3, 2026Updated 3 weeks ago
- π Code for : "CRAFT: Concept Recursive Activation FacTorization for Explainability" (CVPR 2023)β76Jul 20, 2023Updated 3 years ago
- β14May 6, 2025Updated last year
- Repository for DISRPT2021 shared taskβ16Sep 5, 2022Updated 4 years ago
- Layer-wise Relevance Propagation for Large Language Models and Vision Transformers [ICML 2024]β251Aug 11, 2026Updated last month
- π¬ Interpretability for Leela Chess Zero networks.β22Sep 3, 2026Updated 3 weeks ago
- β17Updated this week
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code to enable layer-level steering in LLMs using sparse auto encodersβ35Sep 18, 2025Updated last year
- β17May 19, 2026Updated 4 months ago
- Interpretability for sequence generation models π πβ476Apr 25, 2026Updated 4 months ago
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.β31Jul 8, 2026Updated 2 months ago
- Code for the "Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning" paper.β19Jul 6, 2026Updated 2 months ago
- Attribution-based Parameter Decompositionβ35Jun 11, 2025Updated last year
- π Code for the paper: "Look at the Variance! Efficient Black-box Explanations with Sobol-based Sensitivity Analysis" (NeurIPS 2021)β33Jul 18, 2022Updated 4 years ago
- Code for Evaluating Explanations for Reading Comprehension with Realistic Counterfactuals.β17Apr 25, 2021Updated 5 years ago
- LENS Projectβ53Feb 22, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ADAG: Transluce's MLP neuron-level circuit tracing libraryβ39Apr 10, 2026Updated 5 months ago
- Engine for collecting, uploading, and downloading model activationsβ31Apr 2, 2025Updated last year
- IR module for experimaestroβ14Sep 16, 2026Updated last week
- ICLR 2024: Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretationsβ24May 1, 2025Updated last year
- Code for "On Measuring Faithfulness of Natural Language Explanations"β23Jul 14, 2026Updated 2 months ago
- Erasing conceptual knowledge from language models through low-rank fine-tuningβ25Mar 27, 2025Updated last year
- https://footprints.baulab.infoβ17Oct 4, 2024Updated last year