Align your LM to express calibrated verbal statements of confidence in its long-form generations.
☆29Jun 4, 2024Updated 2 years ago
Alternatives and similar repositories for linguistic_calibration
Users that are interested in linguistic_calibration are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Confidence Regulation Neurons in Language Models (NeurIPS 2024)☆15Feb 1, 2025Updated last year
- ☆25Jun 10, 2025Updated last year
- Uncertainty quantification for in-context learning of large language models☆15Apr 1, 2024Updated 2 years ago
- ☆11Sep 11, 2022Updated 3 years ago
- This repository contains the code for the paper "Can Transformers Learn Full Bayesian Inference In Context?"☆16Apr 6, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for identifying natural backdoors in existing image datasets.☆15Aug 24, 2022Updated 3 years ago
- Teaching Models to Express Their Uncertainty in Words☆38May 26, 2022Updated 4 years ago
- Source code of "Calibrating Large Language Models Using Their Generations Only", ACL2024☆22Nov 20, 2024Updated last year
- Code repository for the WWW 2019 paper "Predicting ConceptNet Path Quality Using Crowdsourced Assessments of Naturalness"☆12Feb 1, 2019Updated 7 years ago
- Official repository for ALT (ALignment with Textual feedback).☆10Jul 25, 2024Updated last year
- Official repository for Beyond Binary Rewards: Training LMs to Reason about Their Uncertainty☆67Aug 20, 2025Updated 11 months ago
- Code and Hummingbird dataset for EMNLP 2021 paper "Does BERT Learn as Humans Perceive? Understanding Linguistic Styles through Lexica"☆14Apr 13, 2022Updated 4 years ago
- Code Release for the 2023 NeurIPS Paper How does GPT-2 compute greater-than?: Interpreting mathematical abilities in a pre-trained langua…☆17Dec 6, 2024Updated last year
- Public code repo for paper "SaySelf: Teaching LLMs to Express Confidence with Self-Reflective Rationales"☆113Sep 28, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models☆17Jun 24, 2024Updated 2 years ago
- code repo for ICLR 2024 paper "Can LLMs Express Their Uncertainty? An Empirical Evaluation of Confidence Elicitation in LLMs"☆148Mar 14, 2024Updated 2 years ago
- Model-agnostic posthoc calibration without distributional assumptions☆41Oct 20, 2023Updated 2 years ago
- Code for the paper "Distinguishing the Knowable from the Unknowable with Language Models"☆11Updated this week
- ☆43Feb 2, 2024Updated 2 years ago
- AbstainQA, ACL 2024☆29Feb 4, 2026Updated 5 months ago
- ☆11Oct 28, 2022Updated 3 years ago
- ☆10Nov 1, 2019Updated 6 years ago
- In-Context Sharpness as Alerts: An Inner Representation Perspective for Hallucination Mitigation (ICML 2024)☆63Mar 30, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official code for our paper "Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models"☆26Oct 31, 2025Updated 8 months ago
- Multiscale Score Matching Analysis☆11Jan 19, 2023Updated 3 years ago
- ☆14Jan 14, 2026Updated 6 months ago
- ☆21May 14, 2026Updated 2 months ago
- Lightweight Adapting for Black-Box Large Language Models☆25Feb 15, 2024Updated 2 years ago
- Benchmarking LLMs via Uncertainty Quantification☆262Jan 30, 2024Updated 2 years ago
- [ICLR 2022] Denoising Likelihood Score Matching for Conditional Score-based Data Generation☆11Jun 15, 2026Updated last month
- BeHonest: Benchmarking Honesty in Large Language Models☆35Aug 15, 2024Updated last year
- ☆22Nov 1, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15Aug 3, 2021Updated 4 years ago
- [ECCV 2024] Official implementation of "Uncertainty Calibration with Energy Based Instance-wise Scaling in the Wild Dataset"☆11Aug 13, 2024Updated last year
- ☆16Nov 30, 2022Updated 3 years ago
- Bias Benchmark for Natural Language Inference. Code repo for the Findings of NAACL 2022 paper "On Measuring Social Biases in Prompt-Based…☆15Apr 28, 2022Updated 4 years ago
- Code for ModularQA☆27Jun 8, 2021Updated 5 years ago
- Code for☆15Oct 16, 2020Updated 5 years ago
- Code for Unbiased Implicit Variational Inference (UIVI)☆15Jan 18, 2019Updated 7 years ago