☆29Feb 27, 2025Updated last year
Alternatives and similar repositories for composable-interventions
Users that are interested in composable-interventions are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS'23] Aging with GRACE: Lifelong Model Editing with Discrete Key-Value Adaptors☆85Dec 21, 2024Updated last year
- [ICML 2023] Protecting Language Generation Models via Invisible Watermarking☆13Sep 8, 2023Updated 3 years ago
- A Model Agnostic function to directly remove specified layers from the LLM☆11May 23, 2024Updated 2 years ago
- ☆25Feb 18, 2025Updated last year
- Whispering Experts: Neural Interventions for Toxicity Mitigation in Language Models, ICML 2024☆26Sep 11, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [EMNLP 2025 Main] ConceptVectors Benchmark and Code for the paper "Intrinsic Evaluation of Unlearning Using Parametric Knowledge Traces"☆39Aug 20, 2025Updated last year
- Code repository for our paper, "Medical Large Language Models are Vulnerable to Data Poisoning Attacks" (Nature Medicine, 2024).☆13Jan 5, 2025Updated last year
- Probing for Labeled Dependency Trees (ACL 2022) + Sorting LMs by Structure (NAACL 2022)☆10Jun 11, 2024Updated 2 years ago
- Reproduction Code for Paper "Investigating Multi-Hop Factual Shortcuts in Knowledge Editing of Large Language Models"☆14Jun 1, 2024Updated 2 years ago
- distilled Self-Critique refines the outputs of a LLM with only synthetic data☆11Apr 11, 2024Updated 2 years ago
- Code associated with the paper: "Few-Shot Self-Rationalization with Natural Language Prompts"☆12Apr 27, 2022Updated 4 years ago
- ☆24Jun 12, 2026Updated 3 months ago
- The official implementation for "Mitigating Overthinking in Large Reasoning Models via Manifold Steering"