πͺ Interpreto is an interpretability toolbox for LLMs
β192Jul 22, 2026Updated this week
Alternatives and similar repositories for interpreto
Users that are interested in interpreto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Build and train Lipschitz-constrained networks: PyTorch implementation of 1-Lipschitz layers. For TensorFlow/Keras implementation, see htβ¦β44Mar 23, 2026Updated 4 months ago
- π Influenciae is a Tensorflow Toolbox for Influence Functionsβ67May 20, 2026Updated 2 months ago
- π Overcomplete is a Vision-based SAE Toolboxβ148Dec 4, 2025Updated 7 months ago
- Simple, compact, and hackable post-hoc deep OOD detection for already trained tensorflow or pytorch image classifiers.β61May 19, 2026Updated 2 months ago
- β40Sep 15, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Build and train Lipschitz constrained networks: TensorFlow implementation of k-Lipschitz layersβ102Mar 14, 2025Updated last year
- Easy-to-use MIRAGE code for faithful answer attribution in RAG applications. Paper: https://aclanthology.org/2024.emnlp-main.347/β25Mar 10, 2025Updated last year
- CoFmuPy is a Python library designed for rapid prototyping of digital twins through the co-simulation of Functional Mock-up Units (FMUs)β23Mar 10, 2026Updated 4 months ago
- π Xplique is a Neural Networks Explainability Toolboxβ747Mar 31, 2026Updated 3 months ago
- β16Nov 14, 2025Updated 8 months ago
- Unified access to Large Language Model modules using NNsightβ116Updated this week
- π Puncc is a python library for predictive uncertainty quantification using conformal prediction.β403Jul 10, 2026Updated 2 weeks ago
- The nnsight package enables interpreting and manipulating the internals of deep learned models.β998Updated this week
- Generic Engine for Multi-disciplinary Scenarios, Exploration and Optimization. This is a MIRROR of our gitlab repository, the developmentβ¦β34Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Repository for "Training Language Models To Explain Their Own Computations"β23Jul 7, 2026Updated 2 weeks ago
- Repository for DISRPT2021 shared taskβ16Sep 5, 2022Updated 3 years ago
- A research toolkit for decomposing and explaining text similarity across neural, structured, and symbolic levels.β30Apr 11, 2026Updated 3 months ago
- Layer-wise Relevance Propagation for Large Language Models and Vision Transformers [ICML 2024]β241Jul 11, 2025Updated last year
- Code for Spectral Norm of Convolutional Layers with Circular and Zero Paddings and Efficient Bound of Lipschitz Constant for Convolutionaβ¦β15Feb 2, 2024Updated 2 years ago
- mETRICS - rEproducible sofTware peRformance analysIs in perfeCt Simplicityβ11May 14, 2025Updated last year
- π¬ Interpretability for Leela Chess Zero networks.β20Apr 29, 2026Updated 2 months ago
- β16Updated this week
- [CVPRW 2024] Conformal prediction for uncertainty quantification in image segmentationβ27Dec 9, 2024Updated last year
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code to enable layer-level steering in LLMs using sparse auto encodersβ34Sep 18, 2025Updated 10 months ago
- β17May 19, 2026Updated 2 months ago
- Masked Omics Modeling for Multimodal Representation Learning across Histopathology and Molecular Profilesβ17May 5, 2026Updated 2 months ago
- Interpretability for sequence generation models π πβ471Apr 25, 2026Updated 3 months ago
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.β31Jul 8, 2026Updated 2 weeks ago
- β33Apr 8, 2026Updated 3 months ago
- Code for the "Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning" paper.β17Jul 6, 2026Updated 2 weeks ago
- Attribution-based Parameter Decompositionβ35Jun 11, 2025Updated last year
- π Code for the paper: "Look at the Variance! Efficient Black-box Explanations with Sobol-based Sensitivity Analysis" (NeurIPS 2021)β33Jul 18, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for Evaluating Explanations for Reading Comprehension with Realistic Counterfactuals.β17Apr 25, 2021Updated 5 years ago
- LENS Projectβ52Feb 22, 2024Updated 2 years ago
- ADAG: Transluce's MLP neuron-level circuit tracing libraryβ34Apr 10, 2026Updated 3 months ago
- Engine for collecting, uploading, and downloading model activationsβ30Apr 2, 2025Updated last year
- IR module for experimaestroβ14Updated this week
- ICLR 2024: Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretationsβ24May 1, 2025Updated last year
- Code for "On Measuring Faithfulness of Natural Language Explanations"β23Jul 14, 2026Updated last week