A Mechanistic‑Interpretability study that finds the structural dynamics of Large Language Models under fine‑tuning.
☆17May 30, 2025Updated last year
Alternatives and similar repositories for FinetuneCircuits
Users that are interested in FinetuneCircuits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mixture of Lora Experts☆11Apr 7, 2024Updated 2 years ago
- Official repository for Activation-Informed Merging (AIM) of Large Language Models☆25Feb 10, 2025Updated last year
- Efficient multi-token attribution for reasoning language models — Python package, CLI, and HTML token traces☆35Updated this week
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated 2 months ago
- [NeurIPS 2024] Knowledge Circuits in Pretrained Transformers☆172Nov 14, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆24Feb 13, 2026Updated 6 months ago
- Code for the experiments and websites of the paper "Same Task, Different Circuits"☆37Jul 21, 2026Updated 3 weeks ago
- Bayesian scaling laws for in-context learning.☆16Mar 12, 2025Updated last year
- ☆18Mar 25, 2026Updated 4 months ago
- Official implementation of "Modeling Multi-Task Model Merging as Adaptive Projective Gradient Descent".☆23May 23, 2025Updated last year
- [CVPR Findings 2026] "Circuit Tracing in Vision-Language Models"☆29Jul 14, 2026Updated last month
- Collections of RLxLM experiments using minimal codes☆14Feb 17, 2025Updated last year
- ☆16Jul 31, 2025Updated last year
- ☆27Oct 12, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ECCV'24 Oral] PiTe: Pixel-Temporal Alignment for Large Video-Language Model☆17Feb 13, 2025Updated last year
- A library for efficient patching and automatic circuit discovery.☆99Dec 31, 2025Updated 7 months ago
- [ICML 2025] EffiCoder: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning☆15May 24, 2025Updated last year
- ☆32Jul 15, 2024Updated 2 years ago
- ☆17Aug 23, 2025Updated 11 months ago
- [ACL 2025] How Do LLMs Acquire New Knowledge? A Knowledge Circuits Perspective on Continual Pre-Training☆50Jul 18, 2025Updated last year
- 在criteo数据集上,用pytorch复现一些ctr模型☆13Jun 20, 2020Updated 6 years ago
- DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling☆38Jul 12, 2024Updated 2 years ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)☆15Jul 24, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Paper List for our ACL 2026 paper "Towards Intrinsic Interpretability of Large Language Models: A Survey of Design Principles and Archite…☆17Apr 23, 2026Updated 3 months ago
- [NeurIPS 2024] For paper Parameter Competition Balancing for Model Merging☆48Oct 11, 2024Updated last year
- ☆24May 23, 2025Updated last year
- UniEval: Unified Holistic Evaluation for Unified Multimodal Understanding and Generation☆25May 16, 2025Updated last year
- 面向对象与多线程课程 Database of Documents☆14Nov 27, 2020Updated 5 years ago
- Research on the Construction and Application of Paraphrase Parallel Corpus☆11Oct 26, 2020Updated 5 years ago
- Syphus: Automatic Instruction-Response Generation Pipeline☆14Dec 14, 2023Updated 2 years ago
- ☆14Apr 22, 2024Updated 2 years ago
- Codebase for Hyperdecoders https://arxiv.org/abs/2203.08304☆14Oct 11, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆10May 26, 2022Updated 4 years ago
- 多 语言降噪预训练模型MBart的中文生成任务☆11May 27, 2021Updated 5 years ago
- A benchmark for testing memorization abilities of LMs☆24Oct 15, 2024Updated last year
- Code for ACL 2024 paper: PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models.☆16Feb 5, 2025Updated last year
- kNN-TL: k-Nearest-Neighbor Transfer Learning for Low-Resource Neural Machine Translation (ACL2023)☆11Jul 26, 2023Updated 3 years ago
- ☆19Jan 17, 2024Updated 2 years ago
- The official repository of "Whoever Started the Interference Should End It: Guiding Data-Free Model Merging via Task Vectors""☆50Oct 1, 2025Updated 10 months ago