This repository contains the resource introduced in the paper: "Truth or Mirage? Towards End-to-End Factuality Evaluation with LLM-Oasis". LLM-Oasis is a large-scale resource for end-to-end factuality evaluation obtained by extracting and falsifying information from Wikipedia.
β25Oct 15, 2025Updated 11 months ago
Alternatives and similar repositories for LLM-Oasis
Users that are interested in LLM-Oasis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Word Sense Linking model is designed to identify and disambiguate spans of text to their most suitable senses from a reference inventory.β13Aug 23, 2024Updated 2 years ago
- A Word Level Transformer layer based on PyTorch and π€ Transformers.β34Jan 31, 2024Updated 2 years ago
- β70Jun 10, 2025Updated last year
- β21Updated this week
- Code repository for the paper "The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Leβ¦β14Jan 16, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The official implementation for Collaborative Word-based Pre-trained Item Representation for Transferable Recommendation.β25Jan 30, 2024Updated 2 years ago
- The official implementation of Cross-Task Experience Sharing (COPS)β29Oct 23, 2024Updated last year
- Entity Disambiguation as text extraction (ACL 2022)β182Apr 17, 2022Updated 4 years ago
- β42Sep 11, 2026Updated last week
- This is the repository for NAACL'25 paper "TART: An Open-Source Tool-Augmented Framework for Explainable Table-based Reasoning"β59May 3, 2025Updated last year
- NLP Preprocessing Pipeline Wrappersβ11May 12, 2023Updated 3 years ago
- β15Apr 12, 2021Updated 5 years ago
- This repository hosts the dataset for the paper Computer Science Named Entity Recognition in the Open Research Knowledge Graphβ21Jan 8, 2024Updated 2 years ago
- This is the official code for the paper "Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation"β56Feb 2, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- β11Oct 12, 2023Updated 2 years ago
- (CVPR 2025) Official implementation to DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation which outperforms SOTAβ¦β28Aug 23, 2025Updated last year
- HelloBench: Evaluating Long Text Generation Capabilities of Large Language Modelsβ60Nov 26, 2024Updated last year
- Code associated with the EMNLP 2024 Main paper: "Image, tell me your story!" Predicting the original meta-context of visual misinformatioβ¦β45Dec 6, 2025Updated 9 months ago
- Official repo for EMNLP 2023 paper "Explain-then-Translate: An Analysis on Improving Program Translation with Self-generated Explanationsβ¦β29Dec 5, 2023Updated 2 years ago
- Simple replication of [ColBERT-v1](https://arxiv.org/abs/2004.12832).β83Mar 18, 2024Updated 2 years ago
- DALI Multi Agent System Frameworkβ43Mar 24, 2026Updated 5 months ago
- [NeurIPS 2024] GTBench: Uncovering the Strategic Reasoning Limitations of LLMs via Game-Theoretic Evaluationsβ73Sep 6, 2024Updated 2 years ago
- Problem-Oriented Segmentation and Retrieval EMNLP 2024 Findingsβ34Nov 12, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is the official repository for OVIR-3D: Open-Vocabulary 3D Instance Retrieval Without Training on 3D Data. (CoRL'23)β115Nov 10, 2023Updated 2 years ago
- [EMNLP 2023] TESTA: Temporal-Spatial Token Aggregation for Long-form Video-Language Understandingβ50Jan 9, 2024Updated 2 years ago
- Statewide Visual Geolocalization in the Wild (ECCV 2024)β75Dec 2, 2024Updated last year
- [ECCV'24 Workshops Oral] DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scalingβ33Feb 6, 2026Updated 7 months ago
- [ECAI 2023] MonoSKD: General Distillation Framework for Monocular 3D Object Detection via Spearman Correlation Coefficientβ32Dec 8, 2023Updated 2 years ago
- Official Implementation of "Pix2Cap-COCO: Advancing Visual Comprehension via Pixel-Level Captioning"β28Dec 16, 2025Updated 9 months ago
- A collection of Italian benchmarks for LLM evaluationβ37Jun 9, 2026Updated 3 months ago
- The multilingual language model for Switzerlandβ29Jan 19, 2024Updated 2 years ago
- EvoEval: Evolving Coding Benchmarks via LLMβ84Apr 6, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SFC: Shared Feature Calibration in Weakly Supervised Semantic Segmentation (AAAI24)β25Jul 2, 2024Updated 2 years ago
- Generic template to bootstrap your Python project.β22Updated this week
- Code and data releases for the paper -- DelTA: An Online Document-Level Translation Agent Based on Multi-Level Memoryβ64Feb 10, 2025Updated last year
- Forecasting.β37Aug 2, 2025Updated last year
- [IEEE OJSP'26, IEEE SLT'24] "Speaker-Disentangled Chunk-Wise Regression for Syllabic Tokenization"β46Updated this week
- β31Dec 6, 2024Updated last year
- [TMLR'24] This repository includes the official implementation our paper "FedConv: Enhancing Convolutional Neural Networks for Handling Dβ¦β25Apr 30, 2024Updated 2 years ago