Repository for PURE: Turning Polysemantic Neurons Into Pure Features by Identifying Relevant Circuits, accepted at CVPR 2024 XAI4CV Workshop (spotlight)
☆20Aug 30, 2026Updated 3 weeks ago
Alternatives and similar repositories for PURE
Users that are interested in PURE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Prototypical Concept-based Explanations, accepted at SAIAD workshop at CVPR 2024.☆17Aug 30, 2026Updated 3 weeks ago
- PyPI package for DualXDA for efficient data attribution and feature-level explanations of training data influence☆22Mar 6, 2026Updated 6 months ago
- ☆18Jun 3, 2026Updated 3 months ago
- Reveal to Revise: An Explainable AI Life Cycle for Iterative Bias Correction of Deep Models. Paper presented at MICCAI 2023 conference.☆19Jan 17, 2024Updated 2 years ago
- A toolkit for quantitative evaluation of data attribution methods.☆60Aug 28, 2026Updated 3 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for "Don't trust your eyes: on the (un)reliability of feature visualizations" (ICML 2024)☆33Nov 15, 2023Updated 2 years ago
- [TMLR 25] An automated method for explaining complex neuron behaviors in deep vision models using large language models☆12Feb 20, 2025Updated last year
- Layer-wise Relevance Propagation for Large Language Models and Vision Transformers [ICML 2024]☆251Aug 11, 2026Updated last month
- Pruning CNN using CNN with toy example☆23Jun 21, 2021Updated 5 years ago
- CoRelAy is a tool to compose small-scale (single-machine) analysis pipelines.☆32Apr 30, 2026Updated 4 months ago
- An eXplainable AI toolkit with Concept Relevance Propagation and Relevance Maximization☆143Jan 14, 2026Updated 8 months ago
- Code for CVPR 2024 Oral "Neural Lineage"☆17Jun 18, 2024Updated 2 years ago
- A tiny easily hackable implementation of a feature dashboard.☆18Oct 21, 2025Updated 11 months ago
- Zennit is a high-level framework in Python using PyTorch for explaining/exploring neural networks using attribution methods like LRP.☆249May 13, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 24] A novel automated neuron explanation framework that can accurately describe poly-semantic concepts in deep neural networks☆14May 2, 2025Updated last year
- ViT Prisma is a mechanistic interpretability library for Vision and Video Transformers (ViTs).☆390Jul 23, 2025Updated last year
- Subliminal learning in LLMs: language models can transmit hidden preferences through seemingly unrelated training data.☆25Nov 9, 2025Updated 10 months ago
- Official code for "Can We Talk Models Into Seeing the World Differently?" (ICLR 2025).☆30Jan 26, 2025Updated last year
- This is the official repository for the "Towards Vision-Language Mechanistic Interpretability: A Causal Tracing Tool for BLIP" paper acce…☆25Feb 16, 2026Updated 7 months ago
- [NeurIPS 2025] Interpreting vision transformers via residual replacement model☆22Nov 3, 2025Updated 10 months ago
- [CVPR2025 Highlight] ICE: Intrinsic Concept Extraction from a Single Image via Diffusion Models☆20Mar 3, 2026Updated 6 months ago
- Implementation of the paper "Improving the Accuracy-Robustness Trade-off of Classifiers via Adaptive Smoothing".☆10Feb 6, 2024Updated 2 years ago
- Fast RDF serialization for Python☆19Aug 24, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Arrakis is a library to conduct, track and visualize mechanistic interpretability experiments.☆31Jul 8, 2026Updated 2 months ago
- [NeurIPS XAIA & Springer] Code and notebooks to paper "A Fresh Look at Sanity Checks for Saliency Maps"☆25Jul 12, 2024Updated 2 years ago
- An AI agent that collaborates with historians to extract structured datasets from primary sources, adapt to heterogeneous documents, and …☆109Jul 29, 2026Updated last month
- Explainable AI in Julia.☆118Jul 13, 2026Updated 2 months ago
- Accompanying codebase for neuroscope.io, a website for displaying max activating dataset examples for language model neurons☆15Feb 13, 2023Updated 3 years ago
- Linear Relational Embeddings (LREs) and Linear Relational Concepts (LRCs) for LLMs in PyTorch☆11Aug 7, 2024Updated 2 years ago
- Code accompanying "Dynamic Predictive Coding: A Model of Hierarchical Sequence Learning and Prediction in the Neocortex"☆10Mar 2, 2025Updated last year
- Understanding Rare Spurious Correlations in Neural Network☆12Jun 5, 2022Updated 4 years ago
- NeurIPS Reproducbility Challenge 2019☆10Feb 25, 2020Updated 6 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆14Mar 4, 2024Updated 2 years ago
- Official implemention of the paper High-Resolution and Precise Counterfactual Medical Image Generation using Language-guided Stable Diffu…☆24Jul 8, 2025Updated last year
- Tools for exploring Transformer neuron behaviour, including input pruning and diversification.☆11Jun 6, 2023Updated 3 years ago
- ☆12Jan 10, 2023Updated 3 years ago
- minimalistic AI library that resembles HF's transformers☆13Dec 31, 2024Updated last year
- [NeurIPS 2023] "Learning to Augment Distributions for Out-of-distribution Detection"☆11Nov 14, 2023Updated 2 years ago
- Official Code for What Makes and Breaks Safety Fine-tuning? A Mechanistic Study (NeurIPS 2024)☆11Oct 31, 2024Updated last year