Code and data for the FACTOR paper
☆54Nov 15, 2023Updated 2 years ago
Alternatives and similar repositories for factor
Users that are interested in factor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [IJCAI 2024] FactCHD: Benchmarking Fact-Conflicting Hallucination Detection☆89Apr 28, 2024Updated 2 years ago
- Butler 是一个用于自动化服务管理和任务调度的工具项目。☆17Updated this week
- Github repository for "FELM: Benchmarking Factuality Evaluation of Large Language Models" (NeurIPS 2023)☆65Dec 25, 2023Updated 2 years ago
- Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"☆71Feb 27, 2024Updated 2 years ago
- ☆19Jun 21, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is the repository of HaluEval, a large-scale hallucination evaluation benchmark for Large Language Models.☆596Feb 12, 2024Updated 2 years ago
- [NeurIPS 2024 D&B] Evaluating Copyright Takedown Methods for Language Models☆17Jul 17, 2024Updated 2 years ago
- 🌏 UI component library for the future, based on WebComponent.☆23Nov 12, 2024Updated last year
- ☆43Sep 3, 2024Updated last year
- TruthfulQA: Measuring How Models Imitate Human Falsehoods☆941Jan 16, 2025Updated last year
- ☆28Oct 6, 2024Updated last year
- Code for "FactKB: Generalizable Factuality Evaluation using Language Models Enhanced with Factual Knowledge". EMNLP 2023.☆20Dec 25, 2023Updated 2 years ago
- ☆12Mar 7, 2024Updated 2 years ago
- ☆16Jul 11, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Aug 19, 2024Updated 2 years ago
- Official Implementation of ACL2023: Don't Parse, Choose Spans! Continuous and Discontinuous Constituency Parsing via Autoregressive Span …☆14Aug 25, 2023Updated 3 years ago
- Do Large Language Models Know What They Don’t Know?☆105Nov 8, 2024Updated last year
- Official repo for EMNLP'24 paper "SOUL: Unlocking the Power of Second-Order Optimization for LLM Unlearning"☆30Oct 1, 2024Updated last year
- ☆61Aug 22, 2024Updated 2 years ago
- ☆19Jul 20, 2022Updated 4 years ago
- Generated geosite.dat based on Antifilter Community List☆28Updated this week
- ☆33Aug 9, 2024Updated 2 years ago
- Dataset and evaluation script for "Evaluating Hallucinations in Chinese Large Language Models"☆139Jun 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Feb 7, 2023Updated 3 years ago
- ☆90Nov 11, 2022Updated 3 years ago
- Revisiting Cross-Lingual Summarization: A Corpus-based Study and A New Benchmark with Improved Annotation☆19Mar 23, 2024Updated 2 years ago
- ☆48Oct 1, 2024Updated last year
- Official repository of the "Transformer Fusion with Optimal Transport" paper, published as a conference paper at ICLR 2024.☆31Apr 19, 2024Updated 2 years ago
- Converter for EN16931 invoices from CII to UBL☆45Updated this week
- Benchmarking multimodal agents on realistic, ultra-challenging visual scenarios requiring long-horizon hybrid tool use.☆73Mar 10, 2026Updated 5 months ago
- This is the code repo for the paper <UTC-IE: A Unified Token-pair Classification Architecture for Information Extraction>☆16Aug 10, 2023Updated 3 years ago
- ☆45Mar 3, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Pack of LLMs: Model Fusion at Test-Time via Perplexity Optimization☆15Apr 25, 2024Updated 2 years ago
- [NeurIPS25] Official repo for "Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning"☆46Oct 3, 2025Updated 10 months ago
- BeHonest: Benchmarking Honesty in Large Language Models☆36Aug 15, 2024Updated 2 years ago
- Source code for Truth-Aware Context Selection: Mitigating the Hallucinations of Large Language Models Being Misled by Untruthful Contexts☆17Sep 2, 2024Updated last year
- ☆17Nov 10, 2021Updated 4 years ago
- [ICLR 2024]Data for "Multilingual Jailbreak Challenges in Large Language Models"☆108Mar 7, 2024Updated 2 years ago
- ☆50Jan 7, 2024Updated 2 years ago