Code and data for evaluating whether AI assistants know when they do not know the answer
☆86Feb 5, 2024Updated 2 years ago
Alternatives and similar repositories for Say-I-Dont-Know
Users that are interested in Say-I-Dont-Know are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2024 Outstanding Paper] Source code for the NAACL 2024 paper entitled "R-Tuning: Instructing Large Language Models to Say 'I Don't…☆140Jul 10, 2024Updated 2 years ago
- Repo for paper: Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge☆14Feb 20, 2024Updated 2 years ago
- ☆77May 22, 2024Updated 2 years ago
- Do Large Language Models Know What They Don’t Know?☆105Nov 8, 2024Updated last year
- Dataset and evaluation code for measuring hallucinations in Chinese large language models☆139Jun 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- code for Preprint paper at Arxiv: MoT: Pre-thinking and Recalling Enable ChatGPT to Self-Improve with Memory-of-Thoughts☆24Nov 29, 2023Updated 2 years ago
- [EMNLP 2022] RLET: A Reinforcement Learning Based Approach for Explainable QA with Entailment Trees☆11Jul 15, 2023Updated 3 years ago
- Repository containing the SPIN experiments on the DIBT 10k ranked prompts☆23Mar 12, 2024Updated 2 years ago
- ☆15Jan 14, 2026Updated 8 months ago
- Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"☆71Feb 27, 2024Updated 2 years ago
- A framework for training, analyzing, and visualizing sparse autoencoders and related interpretability methods☆228Sep 6, 2026Updated 2 weeks ago
- Grade-School Math with Irrelevant Context (GSM-IC) benchmark is an arithmetic reasoning dataset built upon GSM8K, by adding irrelevant se…☆67Feb 13, 2023Updated 3 years ago
- Pile Deduplication Code☆18May 15, 2023Updated 3 years ago
- ☆42Aug 21, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆112Jul 15, 2025Updated last year
- Self-Knowledge Guided Retrieval Augmentation for Large Language Models (EMNLP Findings 2023)☆27Dec 8, 2023Updated 2 years ago
- 记录Transformer升级的论文笔记☆19Jun 25, 2023Updated 3 years ago
- ☆43Sep 3, 2024Updated 2 years ago
- [NeurIPS'22 Spotlight] Data and code for our paper CoNT: Contrastive Neural Text Generation☆152May 10, 2023Updated 3 years ago
- [Findings of EMNLP'2024] Unified Active Retrieval for Retrieval Augmented Generation☆23Sep 30, 2024Updated last year
- ☆16Jul 9, 2025Updated last year
- Repo for ACL2023 paper "Won't Get Fooled Again: Answering Questions with False Premises"☆23Jun 11, 2023Updated 3 years ago
- This is the code repo for the paper <UTC-IE: A Unified Token-pair Classification Architecture for Information Extraction>☆16Aug 10, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 23] CodeIE: Large Code Generation Models are Better Few-Shot Information Extractors☆42Dec 14, 2025Updated 9 months ago
- Awesome papers on Language-Model-as-a-Service (LMaaS)☆544May 14, 2024Updated 2 years ago
- In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation (ICML 2024)☆64Mar 30, 2024Updated 2 years ago
- [Findings of ACL'2023] Improving Contrastive Learning of Sentence Embeddings from AI Feedback☆40Aug 14, 2023Updated 3 years ago
- An all-in-one framework for Ad-hoc Information Retrieval.☆18Apr 3, 2024Updated 2 years ago
- The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning (NeurIPS 2022)☆16Feb 11, 2023Updated 3 years ago
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- Multi-GPU supported kmeans clustering for cluser-clip☆15Jun 3, 2024Updated 2 years ago
- Self-Supervised Alignment with Mutual Information☆20May 24, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This is the repo for our work "Towards Persona-Based Empathetic Conversational Models" (EMNLP 2020)☆40Dec 28, 2020Updated 5 years ago
- [EMNLP'24] LongHeads: Multi-Head Attention is Secretly a Long Context Processor☆32Apr 8, 2024Updated 2 years ago
- Towards Systematic Measurement for Long Text Quality☆39Sep 5, 2024Updated 2 years ago
- ☆283Jan 6, 2025Updated last year
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago
- ☆90Nov 11, 2022Updated 3 years ago
- ACL'23: Unified Demonstration Retriever for In-Context Learning☆38Dec 2, 2023Updated 2 years ago