Interpretable Causal Diffusion Language Models
☆240Jul 18, 2026Updated last month
Alternatives and similar repositories for steerling
Users that are interested in steerling are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for "Training Language Models To Explain Their Own Computations"☆35Jul 7, 2026Updated last month
- Parameter Decomposition☆140Aug 24, 2026Updated last week
- ☆25Apr 23, 2024Updated 2 years ago
- A comprehensive AI & ML project portfolio from the University of Texas at Austin PG Program, demonstrating real-world data science and ma…☆18Jan 25, 2026Updated 7 months ago
- Find the samples, in the test data, on which your (generative) model makes mistakes.☆32Oct 16, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Feb 24, 2026Updated 6 months ago
- Attribute statements generated by LLMs to preceding tokens using attention weights.☆28Apr 22, 2025Updated last year
- Code for "On Measuring Faithfulness of Natural Language Explanations"☆23Jul 14, 2026Updated last month
- Official PyTorch implementation for "TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors" [ACL 2026]☆49Apr 14, 2026Updated 4 months ago
- Official Implementation of Knowledge Flow Prompting☆35Oct 20, 2025Updated 10 months ago
- ☆16Jul 7, 2026Updated last month
- The nnsight package enables interpreting and manipulating the internals of deep learned models.☆1,081Updated this week
- ☆13Oct 31, 2024Updated last year
- Code for Evaluating Explanations for Reading Comprehension with Realistic Counterfactuals.☆17Apr 25, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- XAI Experiments on an Annotated Dataset of Wild Bee Images☆20Sep 5, 2025Updated 11 months ago
- [NeurIPS 2025 MechInterp Workshop - Spotlight] Official implementation of the paper "RelP: Faithful and Efficient Circuit Discovery in La…☆29Nov 3, 2025Updated 9 months ago
- MishformerLens intends to be a drop-in replacement for TransformerLens that AST patches HuggingFace Transformers rather than implementing…☆10Oct 7, 2024Updated last year
- Fast Axiomatic Attribution for Neural Networks (NeurIPS*2021)☆15Feb 24, 2026Updated 6 months ago
- ☆18Oct 11, 2025Updated 10 months ago
- I replicated Ng's RYS method and found that duplicating 3 specific layers in Qwen2.5-32B boosts reasoning by 17% and duplicating layers 1…☆242Mar 20, 2026Updated 5 months ago
- Head Vis Public Release☆41May 4, 2026Updated 3 months ago
- ☆87Mar 12, 2026Updated 5 months ago
- ☆15Dec 12, 2024Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Implementation of a list sweeper: instead of the cartesian product, sweep over the zipped list☆28Mar 17, 2025Updated last year
- A toolkit for describing model features and intervening on those features to steer behavior.☆255Mar 16, 2026Updated 5 months ago
- ⚓️ Repository for the "Thought Anchors: Which LLM Reasoning Steps Matter?" paper.☆142Oct 27, 2025Updated 10 months ago
- Contextual Parameter Generation for Knowledge Graph Link Prediction☆22May 3, 2024Updated 2 years ago
- ☆63Jul 10, 2025Updated last year
- [ACL'26 Findings] Steering LLM Thinking with Budget Guidance☆34Feb 19, 2026Updated 6 months ago
- Find informative examples to efficiently (human)-evaluate NLG models.☆17Apr 22, 2026Updated 4 months ago
- ☆41May 7, 2026Updated 3 months ago
- This was designed for interp researchers who want to do research on or with interp agents to give quality of life improvements and fix …☆146Feb 8, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Oct 6, 2022Updated 3 years ago
- Model souping for LLMs☆75Nov 18, 2025Updated 9 months ago
- Sparse Autoencoders (SAE) vs CLIP fine-tuning fun.☆18Dec 19, 2024Updated last year
- Code for the paper: Discover-then-Name: Task-Agnostic Concept Bottlenecks via Automated Concept Discovery. ECCV 2024.☆60Nov 3, 2024Updated last year
- Do input gradients highlight discriminative features? [NeurIPS 2021] (https://arxiv.org/abs/2102.12781)☆12Jan 10, 2023Updated 3 years ago
- A quick way to get started with Transformer Lens☆15Dec 13, 2023Updated 2 years ago
- Delphi was the home of a temple to Phoebus Apollo, which famously had the inscription, 'Know Thyself.' This library lets language models …☆275Updated this week