☆15Jan 20, 2026Updated 6 months ago
Alternatives and similar repositories for Actionable-MI
Users that are interested in Actionable-MI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- code for EMNLP 2024 paper: Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis☆12Nov 17, 2024Updated last year
- The Github repo for our survey paper: "Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large…☆152Apr 15, 2026Updated 3 months ago
- code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models☆52Nov 17, 2024Updated last year
- Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures☆34Jan 29, 2026Updated 6 months ago
- Procedural Knowledge at Scale Improves ReasoningThis repository contains the minimal, end-to-end pipeline for reproducing the paper resul…☆15Apr 1, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Source Code for our ICLR'26 paper☆17Feb 22, 2026Updated 5 months ago
- code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for M…☆13Nov 17, 2024Updated last year
- 🚀 First survey on Attention Sink in Transformers — 200+ papers on utilization, interpretation, and mitigation.☆138Jun 5, 2026Updated 2 months ago
- PhyX: Does Your Model Have the "Wits" for Physical Reasoning?☆54Mar 16, 2026Updated 4 months ago
- FeedbackQA: Improving Question Answering Post-Deployment with Interactive Feedback☆12Jul 13, 2022Updated 4 years ago
- awesome SAE papers☆80May 24, 2025Updated last year
- Code for Research Project TLDR☆26Jul 28, 2025Updated last year
- my solution for UC Berkeley AI projects pacman☆11Jul 25, 2020Updated 6 years ago
- ☆12Jun 12, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆34Nov 16, 2025Updated 8 months ago
- A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs). This repository agg…☆217Mar 4, 2026Updated 5 months ago
- Code for "Automatic Circuit Finding and Faithfulness"☆19Jul 11, 2024Updated 2 years ago
- ✱ Understanding the underlying learning dynamics of simple tasks in Transformer networks☆19Aug 16, 2024Updated last year
- 🎓Automatically Update circult-eda-mlsys-tinyml Papers Daily using Github Actions (Update Every 8th hours)☆10Aug 3, 2026Updated last week
- A lightweight Inference Engine built for block diffusion models☆47Apr 12, 2026Updated 3 months ago
- Unofficial pytorch implementation of the paper "Learnable Fourier Features for Multi-Dimensional Spatial Positional Encoding", NeurIPS 20…☆13Apr 24, 2024Updated 2 years ago
- This repository collects all relevant resources about interpretability in LLMs☆404Nov 1, 2024Updated last year
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 2019 PyCon kr tutorial: "네이버 영화 평점 데이터로 자연어처리 논문 구현 시작하기"☆13Aug 21, 2019Updated 6 years ago
- ReadMe++: A Multi-domain Multilingual Dataset for Readability Assessment☆14Apr 15, 2025Updated last year
- A versatile toolkit for applying Logit Lens to modern large language models (LLMs). Currently supports Llama-3.1-8B and Qwen-2.5-7B, enab…☆174Aug 14, 2025Updated 11 months ago
- ☆22Mar 19, 2024Updated 2 years ago
- [AAAI 2025] Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks☆12Jun 19, 2025Updated last year
- Documentation at☆14Mar 27, 2025Updated last year
- This repo is for the safety topic, including attacks, defenses and studies related to reasoning and RL☆67Sep 5, 2025Updated 11 months ago
- Code for the paper "A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis"☆20Jun 12, 2025Updated last year
- A repository for awesome resources in mechanistic interpretability☆16Jan 18, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Source codes for the paper "Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning" (PDMER) which p…☆14Mar 24, 2025Updated last year
- ☆14Sep 17, 2025Updated 10 months ago
- ☆42Jun 11, 2025Updated last year
- awesome papers in LLM interpretability☆625Aug 20, 2025Updated 11 months ago
- Official implementation of the ACL 2022 paper "Learning Non-Autoregressive Models from Search for Unsupervised Sentence Summarization"☆14Dec 26, 2022Updated 3 years ago
- [ICLR 2025 Oral] Knowledge Entropy Decay during Language Model Pretraining Hinders New Knowledge Acquisition☆17Nov 25, 2024Updated last year
- Exploring the Limitations of Large Language Models on Multi-Hop Queries☆33Mar 2, 2025Updated last year