☆19Sep 1, 2025Updated last year
Alternatives and similar repositories for SADI
Users that are interested in SADI are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Steering Llama 2 with Contrastive Activation Addition☆250May 23, 2024Updated 2 years ago
- GASP: Efficient Black-Box Generation of Adversarial Suffixes for Jailbreaking LLMs☆17Nov 12, 2025Updated 9 months ago
- ☆13Aug 19, 2024Updated 2 years ago
- [ICLR 2025] General-purpose activation steering library☆188Sep 18, 2025Updated 11 months ago
- Official code for Steering Large Language Models using Conceptors, presented at the NeurIPS 2024 MINT Workshop.☆16Mar 13, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation Code for "LLM-based Medical Assistant Personalization with Short- and Long-Term Memory Coordination"☆14May 17, 2026Updated 3 months ago
- Official codebase for "Analyzing the Generalization and Reliability of Steering Vectors"☆22Dec 14, 2024Updated last year
- ☆21Aug 19, 2025Updated last year
- Materials for "Multi-property Steering of Large Language Models with Dynamic Activation Composition"☆14Nov 22, 2024Updated last year
- ☆15Apr 25, 2025Updated last year
- ☆13Mar 7, 2025Updated last year
- Extended Inductive Reasoning for Personalized Preference Inference from Behavioral Signals☆11Jan 8, 2026Updated 7 months ago
- [ICCV 2025] ONLY: One-Layer Intervention Sufficiently Mitigates Hallucinations in Large Vision-Language Models☆51Jul 7, 2025Updated last year
- The dataset consists of public social media url pairs and the corresponding entailment label for an external conference (ACL 2021). Each …☆14Aug 16, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [EMNLP 2025] Circuit-Aware Editing Enables Generalizable Knowledge Learners☆19Nov 17, 2025Updated 9 months ago
- Enhancing contextual understanding in large language models through contrastive decoding☆19May 3, 2024Updated 2 years ago
- ☆19Sep 3, 2024Updated 2 years ago
- Look, Compare, Decide: Alleviating Hallucination in Large Vision-Language Models via Multi-View Multi-Path Reasoning☆24Sep 9, 2024Updated last year
- Official implementation of ICLR 2026 paper "LUMINA: Detecting Hallucinations in RAG System with Context–Knowledge Signals"☆19Jan 31, 2026Updated 7 months ago
- An unofficial implementation of SOLAR-10.7B model and the newly proposed interlocked-DUS(iDUS) implementation and experiment details.☆14Mar 20, 2024Updated 2 years ago
- Implementaiton of "DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation" (accepted by NAACL2024 Findings)".☆28Feb 10, 2025Updated last year
- A curated list of resources for activation engineering☆139Oct 2, 2025Updated 11 months ago
- [WIP] [NeurIPS 2025 Spotlight] Angular Steering: Behavior Control via Rotation in Activation Space☆25May 25, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Inference-Time Intervention: Eliciting Truthful Answers from a Language Model☆583Jan 28, 2025Updated last year
- EMNLP 2021: A Label-Aware BERT Attention Network for Zero-Shot Multi-Intent Detection in Spoken Language Understanding☆10Apr 8, 2022Updated 4 years ago
- ☆18Jan 17, 2024Updated 2 years ago
- [ICLR 2025] Is Your Model Really A Good Math Reasoner? Evaluating Mathematical Reasoning with Checklist☆34Oct 23, 2024Updated last year
- The implement of paper:"ReDeEP: Detecting Hallucination in Retrieval-Augmented Generation via Mechanistic Interpretability"☆69Jun 3, 2025Updated last year
- ☆11Jun 20, 2023Updated 3 years ago
- Reproduction Code for Paper "Investigating Multi-Hop Factual Shortcuts in Knowledge Editing of Large Language Models"☆14Jun 1, 2024Updated 2 years ago
- Code for Reducing Hallucinations in Vision-Language Models via Latent Space Steering☆119Nov 23, 2024Updated last year
- ☆10Sep 10, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 【NeurIPS 2024】The official code of paper "Automated Multi-level Preference for MLLMs"☆22Sep 26, 2024Updated last year
- ☆16Apr 11, 2022Updated 4 years ago
- ACL 2024: LoRA-Flow Dynamic LoRA Fusion for Large Language Models in Generative Tasks☆25Oct 9, 2024Updated last year
- [NeurIPS 2025] Official Implementation for "Enhancing Vision-Language Model Reliability with Uncertainty-Guided Dropout Decoding"☆22Dec 8, 2024Updated last year
- ☆24Apr 20, 2024Updated 2 years ago
- ☆20Mar 11, 2026Updated 5 months ago
- Official code for the paper Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception. The code is based on t…☆21Aug 5, 2025Updated last year