Code and data for evaluating whether AI assistants know when they do not know the answer
☆86Feb 5, 2024Updated 2 years ago
Alternatives and similar repositories for Say-I-Dont-Know
Users that are interested in Say-I-Dont-Know are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL 2024 Outstanding Paper] Source code for the NAACL 2024 paper entitled "R-Tuning: Instructing Large Language Models to Say 'I Don't…☆139Jul 10, 2024Updated 2 years ago
- Repo for paper: Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge☆14Feb 20, 2024Updated 2 years ago
- ☆77May 22, 2024Updated 2 years ago
- Do Large Language Models Know What They Don’t Know?☆105Nov 8, 2024Updated last year
- Dataset and evaluation code for measuring hallucinations in Chinese large language models☆139Jun 5, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2022] RLET: A Reinforcement Learning Based Approach for Explainable QA with Entailment Trees☆11Jul 15, 2023Updated 3 years ago
- Repository containing the SPIN experiments on the DIBT 10k ranked prompts☆23Mar 12, 2024Updated 2 years ago
- ☆15Jan 14, 2026Updated 7 months ago
- A survey of long-context language models covering architecture, infrastructure, training, and evaluation☆63Mar 31, 2025Updated last year
- Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"☆71Feb 27, 2024Updated 2 years ago
- A framework for training, analyzing, and visualizing sparse autoencoders and related interpretability methods☆227Updated this week
- [ACL 2024] Benchmarking Knowledge Boundary for Large Language Models: A Different Perspective on Model Evaluation☆10May 26, 2024Updated 2 years ago
- The codebase for "Learning from Easy to Complex: Adaptive Multi-curricula Learning for Neural Dialogue Generation" (Cai et al., AAAI 2020…☆20Jun 18, 2024Updated 2 years ago
- Pile Deduplication Code☆18May 15, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆42Aug 21, 2025Updated last year
- ☆113Jul 15, 2025Updated last year
- Self-Knowledge Guided Retrieval Augmentation for Large Language Models (EMNLP Findings 2023)☆27Dec 8, 2023Updated 2 years ago
- ☆43Sep 3, 2024Updated last year
- ☆12Apr 15, 2024Updated 2 years ago
- "Towards Improving Document Understanding: An Exploration on Text-Grounding via MLLMs" 2023☆16Nov 28, 2024Updated last year
- [NeurIPS'22 Spotlight] Data and code for our paper CoNT: Contrastive Neural Text Generation☆152May 10, 2023Updated 3 years ago
- ☆16Jul 9, 2025Updated last year
- Distributional Generalization in NLP. A roadmap.☆86Dec 12, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is the code repo for the paper <UTC-IE: A Unified Token-pair Classification Architecture for Information Extraction>☆16Aug 10, 2023Updated 3 years ago
- [ACL 23] CodeIE: Large Code Generation Models are Better Few-Shot Information Extractors☆42Dec 14, 2025Updated 8 months ago
- Awesome papers on Language-Model-as-a-Service (LMaaS)☆545May 14, 2024Updated 2 years ago
- In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation (ICML 2024)☆64Mar 30, 2024Updated 2 years ago
- [Findings of ACL'2023] Improving Contrastive Learning of Sentence Embeddings from AI Feedback☆40Aug 14, 2023Updated 3 years ago
- The Unreliability of Explanations in Few-shot Prompting for Textual Reasoning (NeurIPS 2022)☆16Feb 11, 2023Updated 3 years ago
- An all-in-one framework for Ad-hoc Information Retrieval.☆18Apr 3, 2024Updated 2 years ago
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- Multi-GPU supported kmeans clustering for cluser-clip☆15Jun 3, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Self-Supervised Alignment with Mutual Information☆20May 24, 2024Updated 2 years ago
- [EMNLP'24] LongHeads: Multi-Head Attention is Secretly a Long Context Processor☆32Apr 8, 2024Updated 2 years ago
- Towards Systematic Measurement for Long Text Quality☆39Sep 5, 2024Updated last year
- ☆283Jan 6, 2025Updated last year
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago
- ☆90Nov 11, 2022Updated 3 years ago
- ACL'23: Unified Demonstration Retriever for In-Context Learning☆38Dec 2, 2023Updated 2 years ago