☆18Jun 19, 2023Updated 3 years ago
Alternatives and similar repositories for logit-explanations
Users that are interested in logit-explanations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Measuring the Mixing of Contextual Information in the Transformer☆35May 27, 2023Updated 3 years ago
- ☆35Mar 2, 2023Updated 3 years ago
- Source code of ACL 2023 accepted paper "AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression"☆13Jun 14, 2023Updated 3 years ago
- ☆18Oct 6, 2022Updated 3 years ago
- RecAlpaca: A simple framework combing Alpaca and Recommendations.☆34Mar 30, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is the official repository for the "Towards Vision-Language Mechanistic Interpretability: A Causal Tracing Tool for BLIP" paper acce…☆25Feb 16, 2026Updated 7 months ago
- [EMNLP 2026] PISanitizer: Preventing Prompt Injection to Long-Context LLMs via Prompt Sanitization☆18Aug 21, 2026Updated last month
- ☆60Feb 28, 2023Updated 3 years ago
- ☆11Jun 14, 2024Updated 2 years ago
- ☆14Sep 11, 2026Updated last week
- Code for the paper "Towards Generalizable Neuro-Symbolic Systems for Commonsense Question Answering" (EMNLP-COIN 2019)☆11Feb 19, 2021Updated 5 years ago
- code for the paper 'Multi-Task Learning for Knowledge Graph Completion with Pre-trained Language Models'☆16Jan 15, 2021Updated 5 years ago
- Code for our paper "AMR-DA: Data augmentation by abstract meaning representation" in ACL 2022☆14May 17, 2022Updated 4 years ago
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Tutorials on training and testing retrieval-based models (DrQA & DPR)☆51Nov 30, 2020Updated 5 years ago
- ☆17Jun 17, 2025Updated last year
- For Blend API OpenAPI Specifications and Related Postman Collections☆11Mar 1, 2024Updated 2 years ago
- MTEB: Massive Text Embedding Benchmark☆11Jan 29, 2024Updated 2 years ago
- This is AlpaGasus2-QLoRA based on LLaMA2 with AlpaGasus mechanism using QLoRA!☆15Nov 22, 2023Updated 2 years ago
- Code for our EMNLP-2023 paper: "Active Instruction Tuning: Improving Cross-Task Generalization by Training on Prompt Sensitive Tasks"☆26Nov 16, 2023Updated 2 years ago
- [WWW 2022] Topic Discovery via Latent Space Clustering of Pretrained Language Model Representations☆92Feb 10, 2022Updated 4 years ago
- Very concise example of integrated gradients (a method to reveal areas of attention in input images)☆10Jun 17, 2019Updated 7 years ago
- This code accompanies the paper DisentQA: Disentangling Parametric and Contextual Knowledge with Counterfactual Question Answering.☆16Mar 20, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Training code for Sparse Autoencoders on Embedding models☆40Jul 11, 2026Updated 2 months ago
- "Visual Prompt Selection for In-Context Learning Segmentation Framework"☆14Dec 13, 2024Updated last year
- [AAAI 2023 Oral] Peeling the Onion: Hierarchical Reduction of Data Redundancy for Efficient Vision Transformer Training☆14Apr 19, 2023Updated 3 years ago
- ☆17Nov 10, 2021Updated 4 years ago
- Converting Chinese number string <=> int/float/str☆20Apr 29, 2025Updated last year
- docker for UTH-BERT: https://ai-health.m.u-tokyo.ac.jp/uth-bert☆14Mar 24, 2023Updated 3 years ago
- The repo for x-ray diffraction pattern crystallography via deep learning.☆16Mar 7, 2024Updated 2 years ago
- How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients☆21Jun 17, 2025Updated last year
- An implementation of "Subspace Representations for Soft Set Operations and Sentence Similarities" (NAACL 2024)☆10May 31, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code Repo for the ACL21 paper "Common Sense Beyond English: Evaluating and Improving Multilingual LMs for Commonsense Reasoning"☆23Oct 26, 2021Updated 4 years ago
- ☆14Feb 24, 2025Updated last year
- ☆15Nov 17, 2020Updated 5 years ago
- Wuxia style novel generation using T5-PEGASUS model. 中文武侠小说续写☆12Nov 22, 2022Updated 3 years ago
- I-SHEEP: Iterative Self-enHancEmEnt Paradigm of LLMs through Self-Instruct and Self-Assessment☆17Jan 16, 2025Updated last year
- Code for ACL2023 paper: Pre-Training to Learn in Context☆106Jul 26, 2024Updated 2 years ago
- Repo accompanying our paper "Do Llamas Work in English? On the Latent Language of Multilingual Transformers".☆88Mar 11, 2024Updated 2 years ago