Fact-Checking the Output of Generative Large Language Models in both Annotation and Evaluation.
☆118Jan 6, 2024Updated 2 years ago
Alternatives and similar repositories for Factcheck-GPT
Users that are interested in Factcheck-GPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- codes for "Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models"☆13Feb 10, 2025Updated last year
- ☆63Jun 7, 2024Updated 2 years ago
- A package to evaluate factuality of long-form generation. Original implementation of our EMNLP 2023 paper "FActScore: Fine-grained Atomic…☆457Apr 13, 2025Updated last year
- This repository contains the dataset and code for "WiCE: Real-World Entailment for Claims in Wikipedia" in EMNLP 2023.☆44Dec 15, 2023Updated 2 years ago
- About Data and Codes for EMNLP 2023 System Demo Paper "QACHECK: A Demonstration System for Question-Guided Multi-Hop Fact-Checking"☆20Dec 19, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆22Jun 1, 2023Updated 3 years ago
- Github repository for "FELM: Benchmarking Factuality Evaluation of Large Language Models" (NeurIPS 2023)☆65Dec 25, 2023Updated 2 years ago
- ☆78Nov 27, 2024Updated last year
- ☆78Feb 16, 2024Updated 2 years ago
- RARR: Researching and Revising What Language Models Say, Using Language Models☆53Jun 22, 2023Updated 3 years ago
- Official implementation of the ACL 2023 paper: "Zero-shot Faithful Factual Error Correction"☆17Aug 14, 2023Updated 3 years ago
- [Data + code] ExpertQA : Expert-Curated Questions and Attributed Answers☆139Mar 14, 2024Updated 2 years ago
- Benchmarking long-form factuality in large language models. Original code for our paper "Long-form factuality in large language models".☆692Jun 18, 2026Updated 3 months ago
- ☆50Jan 7, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- FacTool: Factuality Detection in Generative AI☆935Aug 19, 2024Updated 2 years ago
- An original implementation of the paper "CREPE: Open-Domain Question Answering with False Presuppositions"☆16Nov 5, 2024Updated last year
- MiniCheck: Efficient Fact-Checking of LLMs on Grounding Documents [EMNLP 2024]☆228Aug 27, 2025Updated last year
- ☆12Oct 17, 2024Updated last year
- Source code of our EMNLP 2024 paper "FactAlign: Long-form Factuality Alignment of Large Language Models"☆19Oct 3, 2024Updated 2 years ago
- A lightweight, agent-style framework for fact-checking atomic claims using iterative retrieval and verification. Reduces LLM and search c…☆25Jun 4, 2025Updated last year
- [ACL'24] WebCiteS: Attributed Query-Focused Summarization on Chinese Web Search Results with Citations☆13Sep 11, 2024Updated 2 years ago
- FactScoreLite is an implementation of the FactScore metric, designed for detailed accuracy assessment in text generation. This package bu…☆14Apr 25, 2024Updated 2 years ago
- FactCG: Enhancing Fact Checkers with Graph-Based Multi-Hop Data (NAACL 2025)☆17Jul 14, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The dataset and code for PeerSum at EMNLP'23.☆17Oct 20, 2025Updated 11 months ago
- FakeCovid- A Multilingual Cross-domain Fact Check News Dataset for COVID-19☆37Sep 8, 2022Updated 4 years ago
- ☆17Aug 27, 2018Updated 8 years ago
- The implementation of <Factual Consistency Evaluation for Text Summarization via Counterfactual Estimation> in PyTorch.☆17Nov 11, 2021Updated 4 years ago
- Interpretable unified language safety checking with large language models☆32Apr 15, 2023Updated 3 years ago
- Links to conference/journal publications in automated fact-checking (resources for the TACL22/EMNLP23 paper).☆580Feb 23, 2025Updated last year
- Do Large Language Models Know What They Don’t Know?☆104Nov 8, 2024Updated last year
- A dataset of fine-grained knowledge graphs of scientific claims☆18Sep 24, 2021Updated 5 years ago
- The code for HerO: a fact-checking pipeline based on open LLMs (the runner-up in AVeriTeC)☆15Mar 18, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆78May 3, 2024Updated 2 years ago
- ☆12May 29, 2025Updated last year
- Repository for "Attribute First, then Generate: Locally-attributable Grounded Text Generation", ACL 2024☆32Dec 19, 2024Updated last year
- [ICLR'24 Spotlight] "Adaptive Chameleon or Stubborn Sloth: Revealing the Behavior of Large Language Models in Knowledge Conflicts"☆88Apr 12, 2024Updated 2 years ago
- WikiWhy is a new benchmark for evaluating LLMs' ability to explain between cause-effect relationships. It is a QA dataset containing 9000…☆49Dec 7, 2023Updated 2 years ago
- A Flask application for analyzing activity on an online discussion forum, using scraping, indexing, analytics, relational graph and NLP.☆11Nov 24, 2020Updated 5 years ago
- This is the repository of HaluEval, a large-scale hallucination evaluation benchmark for Large Language Models.☆602Feb 12, 2024Updated 2 years ago