☆15Jan 20, 2026Updated 7 months ago
Alternatives and similar repositories for Actionable-MI
Users that are interested in Actionable-MI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Github repo for our survey paper: "Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large…☆154Apr 15, 2026Updated 4 months ago
- code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models☆52Nov 17, 2024Updated last year
- [ICLR2026🔥Oral] SwingArena: Competitive Programming Arena for Long-context GitHub Issue Solving☆15Feb 26, 2026Updated 6 months ago
- Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures☆35Jan 29, 2026Updated 7 months ago
- Procedural Knowledge at Scale Improves ReasoningThis repository contains the minimal, end-to-end pipeline for reproducing the paper resul…☆16Apr 1, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ReasoningLens: a user-friendly toolkit to visualize, understand, and debug model reasoning chains.☆26Jul 7, 2026Updated last month
- Source Code for our ICLR'26 paper☆17Feb 22, 2026Updated 6 months ago
- code for EMNLP 2024 paper: How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for M…☆13Nov 17, 2024Updated last year
- 🚀 First survey on Attention Sink in Transformers — 200+ papers on utilization, interpretation, and mitigation.☆141Jun 5, 2026Updated 2 months ago
- PhyX: Does Your Model Have the "Wits" for Physical Reasoning?☆55Mar 16, 2026Updated 5 months ago
- The latest progress of Personalized Large Language Models (LLMs).☆63Updated this week
- [ACL '26] source code for the paper: "Long-Chain Reasoning Distillation via Adaptive Prefix Alignment"☆17Jan 21, 2026Updated 7 months ago
- awesome SAE papers☆81May 24, 2025Updated last year
- Code for Research Project TLDR☆26Jul 28, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Jun 12, 2024Updated 2 years ago
- A curated collection of resources focused on the Mechanistic Interpretability (MI) of Large Multimodal Models (LMMs). This repository agg…☆219Mar 4, 2026Updated 5 months ago
- Code for "Automatic Circuit Finding and Faithfulness"☆19Jul 11, 2024Updated 2 years ago
- ✱ Understanding the underlying learning dynamics of simple tasks in Transformer networks☆19Aug 16, 2024Updated 2 years ago
- ☆29Aug 8, 2025Updated last year
- Self-Distribution BNN☆10Mar 8, 2022Updated 4 years ago
- MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following☆16Oct 31, 2024Updated last year
- This repository collects all relevant resources about interpretability in LLMs☆404Nov 1, 2024Updated last year
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated 2 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 2019 PyCon kr tutorial: "네이버 영화 평점 데이터로 자연어처리 논문 구현 시작하기"☆13Aug 21, 2019Updated 7 years ago
- ReadMe++: A Multi-domain Multilingual Dataset for Readability Assessment☆14Apr 15, 2025Updated last year
- A versatile toolkit for applying Logit Lens to modern large language models (LLMs). Currently supports Llama-3.1-8B and Qwen-2.5-7B, enab…☆175Aug 14, 2025Updated last year
- SMART introduces a novel test-time framework where Small Language Models (SLMs) reason step-by-step, and Large Language Models (LLMs) pro…☆12Jul 9, 2025Updated last year
- ☆22Mar 19, 2024Updated 2 years ago
- [AAAI 2025] Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks☆12Jun 19, 2025Updated last year
- Documentation at☆14Mar 27, 2025Updated last year
- This repo is for the safety topic, including attacks, defenses and studies related to reasoning and RL☆67Sep 5, 2025Updated 11 months ago
- ☆98Mar 28, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for the paper "A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis"☆20Jun 12, 2025Updated last year
- A repository for awesome resources in mechanistic interpretability☆16Jan 18, 2023Updated 3 years ago
- The source code for running LLMs on the AAAR-1.0 benchmark.☆20Apr 5, 2025Updated last year
- [CIKM-21] Pytorch implementation of LiteGT: Efficient and Lightweight Graph Transformers☆12Nov 16, 2021Updated 4 years ago
- ☆42Jun 11, 2025Updated last year
- awesome papers in LLM interpretability☆626Aug 20, 2025Updated last year
- Official implementation of the ACL 2022 paper "Learning Non-Autoregressive Models from Search for Unsupervised Sentence Summarization"☆14Dec 26, 2022Updated 3 years ago