[NeurIPS 2025] Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models
☆91Jun 5, 2026Updated 2 months ago
Alternatives and similar repositories for sae-for-vlm
Users that are interested in sae-for-vlm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code and data for the paper "LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs"☆48Mar 31, 2026Updated 4 months ago
- [ICCV 2025] Auto Interpretation Pipeline and many other functionalities for Multimodal SAE Analysis.☆199Sep 26, 2025Updated 10 months ago
- [NeurIPS 2025] This is the official repository for VL-SAE: Interpreting and Enhancing Vision-Language Alignment with a Unified Concept Se…☆15Oct 29, 2025Updated 9 months ago
- Sparse autoencoders for vision☆66Updated this week
- Interpreting CLIP with Hierarchical Sparse Autoencoders (ICML 2025)☆28Jan 17, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A framework that allows you to apply Sparse AutoEncoder on any models☆53Jul 11, 2025Updated last year
- [ICLR '25] Official Pytorch implementation of "Interpreting and Editing Vision-Language Representations to Mitigate Hallucinations"☆105Nov 30, 2025Updated 8 months ago
- Implementation of PatchSAE as presented in "Sparse autoencoders reveal selective remapping of visual concepts during adaptation"☆33Apr 22, 2026Updated 3 months ago
- ☆16Jun 14, 2025Updated last year
- 👋 Overcomplete is a Vision-based SAE Toolbox☆148Dec 4, 2025Updated 8 months ago
- Code for the experiments and websites of the paper "Same Task, Different Circuits"☆37Jul 21, 2026Updated 3 weeks ago
- Localization of Knowledge in Text-to-Image Models☆11Oct 8, 2024Updated last year
- This repository contains the code used for the experiments in the paper "Language Models use Lookbacks to Track Beliefs".☆17Mar 14, 2026Updated 5 months ago
- If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions☆17Apr 4, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [CVPR Findings 2026] "Circuit Tracing in Vision-Language Models"☆29Jul 14, 2026Updated 3 weeks ago
- [ICML 2026] The Latent Color Subspace: Emergent Order in High-Dimensional Chaos☆27Jun 9, 2026Updated 2 months ago
- [NeurIPS 2025] Official Implementation of paper "Sherlock: Self-Correcting Reasoning in Vision-Language Models"☆31Jun 4, 2026Updated 2 months ago
- [ACL 2024] FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model☆17Apr 28, 2025Updated last year
- [CVPR 2026 Oral] FINER: MLLMs Hallucinate under Fine-grained Negative Queries☆18Jul 6, 2026Updated last month
- ☆72Jan 17, 2025Updated last year
- Source code for EMNLP2022 paper "Finding Skill Neurons in Pre-trained Transformers via Prompt Tuning".☆18Mar 13, 2023Updated 3 years ago
- [ICML 24] A novel automated neuron explanation framework that can accurately describe poly-semantic concepts in deep neural networks☆14May 2, 2025Updated last year
- ViT Prisma is a mechanistic interpretability library for Vision and Video Transformers (ViTs).☆383Jul 23, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ICML 2025] Unlearning in Diffusion Models using Sparse Autoencoders☆62Oct 16, 2025Updated 9 months ago
- ☆22Jun 4, 2025Updated last year
- Training Sparse Autoencoders on Language Models☆1,501Updated this week
- Code for the multi-agent computer use project.☆21Jul 3, 2026Updated last month
- Efficient Dictionary Learning with Switch Sparse Autoencoders (SAEs)☆25Dec 1, 2024Updated last year
- Code for Stochastic Concept Bottleneck Models☆16Aug 5, 2026Updated last week
- This is the official implementation of the Concept Discovery Models paper.☆15Aug 27, 2023Updated 2 years ago
- ☆430Aug 21, 2025Updated 11 months ago
- Code for reproducing our paper "Are Sparse Autoencoders Useful? A Case Study in Sparse Probing"☆34Mar 31, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- 🔥 [NeurIPS 2025] Official implementation of "Generate, but Verify: Reducing Visual Hallucination in Vision-Language Models with Retrospe…☆58Jan 22, 2026Updated 6 months ago
- v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning☆21Updated this week
- ☆87Nov 5, 2024Updated last year
- The official repository for Deformable ProtoPNet, as described in "Deformable ProtoPNet: An Interpretable Image Classifier Using Deformab…☆54Dec 3, 2024Updated last year
- Code for Negation Neglect☆16May 22, 2026Updated 2 months ago
- [NeurIPS 2024] CoSy is an automatic evaluation framework for textual explanations of neurons.☆20Jan 28, 2026Updated 6 months ago
- [AAAI 2025] Official Implementation of I-HallA v1.0☆16Feb 2, 2025Updated last year