Official repo for the paper "Bilinear MLPs enable weight-based mechanistic interpretability".
☆42Jun 2, 2026Updated 2 months ago
Alternatives and similar repositories for bilinear-decomposition
Users that are interested in bilinear-decomposition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation: Large Language Models are Interpretable Learners - Google☆13Jun 29, 2024Updated 2 years ago
- A toolkit for embedding text datasets with sparse autoencoders☆30Mar 24, 2026Updated 5 months ago
- ☆33Feb 11, 2025Updated last year
- Forkit Core is an open source passport layer for AI models and agents with GitHub CI validation, local verification, and Hugging Face-com…☆17Jun 9, 2026Updated 2 months ago
- [NeurIPS'25] Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders☆16May 28, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆24Jun 18, 2024Updated 2 years ago
- Drug Target Interaction Prediction Using Protein Binding Sites And Drug Fragments☆12Aug 11, 2025Updated last year
- Official repository of NeurIPS 2025 paper "LoMix: Learnable Weighted Multi‑Scale Logits Mixing for Medical Image Segmentation"☆19Feb 4, 2026Updated 6 months ago
- ☆20Jun 2, 2026Updated 2 months ago
- [ICML 2025 Spotlight] Raptor computes expressive embeddings of medical volumes with no training.☆18Jun 23, 2026Updated 2 months ago
- 3cb: Catastrophic Cyber Capabilities Benchmarking of Large Language Models☆17Oct 30, 2024Updated last year
- Code for "A Principled Framework for Multi-View Contrastive Learning"☆20Jul 10, 2025Updated last year
- code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models☆52Nov 17, 2024Updated last year
- DiWA: Diverse Weight Averaging for Out-of-Distribution Generalization☆31Jan 31, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official codebase for "Analyzing the Generalization and Reliability of Steering Vectors"☆22Dec 14, 2024Updated last year
- R3: Robust Rubric-Agnostic Reward Models☆23Jul 12, 2025Updated last year
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 7 years ago
- ☆19May 20, 2025Updated last year
- ☆64Apr 25, 2020Updated 6 years ago
- ☆66Jan 13, 2022Updated 4 years ago
- ☆14Nov 15, 2022Updated 3 years ago
- [NeurIPS 2025 MechInterp Workshop - Spotlight] Official implementation of the paper "RelP: Faithful and Efficient Circuit Discovery in La…☆29Nov 3, 2025Updated 9 months ago
- 👄🇧🇷 Alinhamento fonético forçado em Português Brasileiro☆13Jul 18, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- PyTorch Implementation of [AudioLCM]: a efficient and high-quality text-to-audio generation with latent consistency model.☆13Jun 15, 2024Updated 2 years ago
- Continuous LLM governance monitoring for regulated environments - EU AI Act, GDPR, ANSSI. Self-hosted, profile-driven, no data leaves you…☆27Updated this week
- Code Release for the 2023 NeurIPS Paper How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained langua…☆17Dec 6, 2024Updated last year
- [ICASSP'24] Investigating Personalization Methods in Text to Music Generation☆47Mar 27, 2024Updated 2 years ago
- ☆19Mar 25, 2025Updated last year
- ☆15Jan 2, 2026Updated 7 months ago
- ☆16Apr 14, 2021Updated 5 years ago
- This repository represents a basic implementation of the paper "Riemannian Geometry of Deep Generative Models", along with the results on…☆13Oct 23, 2019Updated 6 years ago
- [ICLR 26] Context Tokens are Anchors: Understanding the Repeat Curse in dMLLMs from an Information Flow Perspective☆19Mar 6, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆96Apr 18, 2026Updated 4 months ago
- DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models (ICLR 2024)☆82Oct 3, 2024Updated last year
- Mental state inference from observable behavior☆15Dec 3, 2021Updated 4 years ago
- Code and data for paper "(How) do Language Models Track State?"☆27Mar 31, 2025Updated last year
- Pretrained model 1024x1024 trained on 1970s scifi art☆16Jul 5, 2023Updated 3 years ago
- Integrated gradients attribution method implemented in PyTorch☆27Nov 5, 2020Updated 5 years ago
- Code for reproducing our paper "Not All Language Model Features Are Linear"☆91Nov 27, 2024Updated last year