πͺ Interpreto is an interpretability toolbox for LLMs
β196Aug 13, 2026Updated this week
Alternatives and similar repositories for interpreto
Users that are interested in interpreto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- New implementations of old orthogonal layers unlock large scale training.β32Sep 19, 2025Updated 10 months ago
- Build and train Lipschitz-constrained networks: PyTorch implementation of 1-Lipschitz layers. For TensorFlow/Keras implementation, see htβ¦β44Mar 23, 2026Updated 4 months ago
- π Influenciae is a Tensorflow Toolbox for Influence Functionsβ67Jul 29, 2026Updated 2 weeks ago
- π Overcomplete is a Vision-based SAE Toolboxβ148Dec 4, 2025Updated 8 months ago
- Easy-to-use MIRAGE code for faithful answer attribution in RAG applications. Paper: https://aclanthology.org/2024.emnlp-main.347/β25Mar 10, 2025Updated last year
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- π CODS - Conformal Object Detection and Segmentationβ22Jul 24, 2026Updated 3 weeks ago
- CoFmuPy is a Python library designed for rapid prototyping of digital twins through the co-simulation of Functional Mock-up Units (FMUs)β24Mar 10, 2026Updated 5 months ago
- π Xplique is a Neural Networks Explainability Toolboxβ750Aug 6, 2026Updated last week
- [NeurIPS 2023] and [ICLR 2024] for robustness certification.β10Nov 30, 2024Updated last year
- β15Jan 2, 2023Updated 3 years ago
- Unified access to Large Language Model modules using NNsightβ119Jul 28, 2026Updated 2 weeks ago
- π Puncc is a python library for predictive uncertainty quantification using conformal prediction.β402Jul 10, 2026Updated last month
- The nnsight package enables interpreting and manipulating the internals of deep learned models.β1,023Updated this week
- Generic Engine for Multi-disciplinary Scenarios, Exploration and Optimization. This is a MIRROR of our gitlab repository, the developmentβ¦β34Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for "Training Language Models To Explain Their Own Computations"β32Jul 7, 2026Updated last month
- DiffuLab is designed to provide a simple and flexible way to train diffusion models while allowing full customization of its core componeβ¦β43Jan 11, 2026Updated 7 months ago
- A runway dataset and a generator of synthetic aerial images with automatic labeling.β135Updated this week
- π Code for : "CRAFT: Concept Recursive Activation FacTorization for Explainability" (CVPR 2023)β76Jul 20, 2023Updated 3 years ago
- β14May 6, 2025Updated last year
- Repository for DISRPT2021 shared taskβ16Sep 5, 2022Updated 3 years ago
- A research toolkit for decomposing and explaining text similarity across neural, structured, and symbolic levels.β30Apr 11, 2026Updated 4 months ago
- Layer-wise Relevance Propagation for Large Language Models and Vision Transformers [ICML 2024]β248Updated this week
- Code for Spectral Norm of Convolutional Layers with Circular and Zero Paddings and Efficient Bound of Lipschitz Constant for Convolutionaβ¦β15Feb 2, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- π¬ Interpretability for Leela Chess Zero networks.β21Aug 8, 2026Updated last week
- β16Updated this week
- [CVPRW 2024] Conformal prediction for uncertainty quantification in image segmentationβ27Dec 9, 2024Updated last year
- Code to enable layer-level steering in LLMs using sparse auto encodersβ34Sep 18, 2025Updated 10 months ago
- β17May 19, 2026Updated 2 months ago
- Interpretability for sequence generation models π πβ474Apr 25, 2026Updated 3 months ago
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.β31Jul 8, 2026Updated last month
- β33Apr 8, 2026Updated 4 months ago
- FastClassification is a tensorflow toolbox for class classification. It provides a training module with various backbones and training trβ¦β15Jun 26, 2021Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code for the "Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning" paper.β18Jul 6, 2026Updated last month
- Attribution-based Parameter Decompositionβ35Jun 11, 2025Updated last year
- π Code for the paper: "Look at the Variance! Efficient Black-box Explanations with Sobol-based Sensitivity Analysis" (NeurIPS 2021)β33Jul 18, 2022Updated 4 years ago
- Code for Evaluating Explanations for Reading Comprehension with Realistic Counterfactuals.β17Apr 25, 2021Updated 5 years ago
- LENS Projectβ53Feb 22, 2024Updated 2 years ago
- Engine for collecting, uploading, and downloading model activationsβ30Apr 2, 2025Updated last year
- IR module for experimaestroβ14Updated this week