[ICML 2025 Poster] SAE-V: Interpreting Multimodal Models for Enhanced Alignment
☆17Jun 5, 2025Updated last year
Alternatives and similar repositories for SAE-V
Users that are interested in SAE-V are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2025] This is the official repository for VL-SAE: Interpreting and Enhancing Vision-Language Alignment with a Unified Concept Se…☆15Oct 29, 2025Updated 8 months ago
- [NeurIPS 2025] Sparse Autoencoders Learn Monosemantic Features in Vision-Language Models☆88Jun 5, 2026Updated last month
- Code to enable layer-level steering in LLMs using sparse auto encoders☆34Sep 18, 2025Updated 10 months ago
- ☆72Jan 17, 2025Updated last year
- A framework that allows you to apply Sparse AutoEncoder on any models☆53Jul 11, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 25] A novel framework for building intrinsically interpretable LLMs with human-understandable concepts to ensure safety, reliabilit…☆33Feb 5, 2026Updated 5 months ago
- Codebase for Temporal SAEs paper☆24Nov 14, 2025Updated 8 months ago
- XL-VLMs: General Repository for eXplainable Large Vision Language Models☆52Sep 8, 2025Updated 10 months ago
- ☆10Apr 17, 2024Updated 2 years ago
- Code for my NeurIPS 2024 ATTRIB paper titled "Attribution Patching Outperforms Automated Circuit Discovery"☆48May 31, 2024Updated 2 years ago
- ☆12Feb 14, 2026Updated 5 months ago
- Qualitative evaluation of automatic chord extraction results: analysis of the musical relationships between predicted chords and target c…☆10Oct 25, 2021Updated 4 years ago
- ☆10Jul 5, 2023Updated 3 years ago
- Implementation of the experiments for "Semi-supervised Neural Chord Estimation Based on a Variational Autoencoder with Latent Chord Label…☆11Dec 3, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official repository for "Piano score rearrangement into multiple difficulty levels via notation-to-notation approach" incl. ST+ token…☆13Feb 26, 2024Updated 2 years ago
- Unofficial implementation of "Explorative Inbetweening of Time and Space"☆13Jul 10, 2024Updated 2 years ago
- [ICCV 2025] Auto Interpretation Pipeline and many other functionalities for Multimodal SAE Analysis.☆199Sep 26, 2025Updated 9 months ago
- [ICLR25] Official Implementation of "Decoupling Angles and Strength in Low-rank Adaptation"☆15Dec 12, 2025Updated 7 months ago
- My tests and experiments with some popular dl frameworks.☆17Sep 11, 2025Updated 10 months ago
- Training Sparse Autoencoders on Language Models☆1,476Updated this week
- Generate beats out of given samples☆16Jan 11, 2023Updated 3 years ago
- ☆47Mar 12, 2025Updated last year
- Instituto de Telecomunicações Deep Learning-based Point Cloud Codec☆11Jun 18, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DynAuditClaw — A security audit skill that dynamically discovers your OpenClaw agent's real configuration, designs targeted attack scenar…☆15Apr 6, 2026Updated 3 months ago
- CVPR2023: Vector Quantization with Self-Attention for Quality-Independent Representation Learning.☆15May 17, 2024Updated 2 years ago
- Resa: Transparent Reasoning Models via SAEs☆50Sep 23, 2025Updated 9 months ago
- ☆18Nov 8, 2024Updated last year
- Interpreting Chest X-rays Like a Radiologist: A Benchmark with Clinical Reasoning, release the dataset and the model weight☆13May 26, 2025Updated last year
- Code for paper "Fast and Complete: Enabling Complete Neural Network Verification with Rapid and Massively Parallel Incomplete Verifiers"☆17Jan 27, 2023Updated 3 years ago
- Official code for the paper "Does CLIP's Generalization Performance Mainly Stem from High Train-Test Similarity?" (ICLR 2024)☆11Aug 26, 2024Updated last year
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆17Jul 1, 2026Updated 2 weeks ago
- [EMNLP 2025] AutoSteer: Automating Steering for Safe Multimodal Large Language Models☆15Aug 21, 2025Updated 11 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICML'21 Oral] Improving Lossless Compression Rates via Monte Carlo Bits-Back Coding☆14Jun 10, 2021Updated 5 years ago
- ☆17Sep 23, 2024Updated last year
- Code for the "Overcoming Sparsity Artifacts in Crosscoders to Interpret Chat-Tuning" paper.☆17Jul 6, 2026Updated 2 weeks ago
- [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective☆21Jun 13, 2025Updated last year
- ☆13Jan 25, 2024Updated 2 years ago
- PET reconstruction with deep learning☆15Jun 28, 2022Updated 4 years ago
- ☆21Jun 25, 2026Updated 3 weeks ago