Natural Questions (NQ) contains real user questions issued to Google search, and answers found from Wikipedia by annotators. NQ is designed for the training and evaluation of automatic question answering systems.
☆1,137Jul 30, 2021Updated 5 years ago
Alternatives and similar repositories for natural-questions
Users that are interested in natural-questions are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Shared repository for open-sourced projects from the Google AI Language team.☆1,814Jun 10, 2026Updated 4 months ago
- Dense Passage Retriever - is a set of tools and models for open domain Q&A task.☆1,871Apr 6, 2023Updated 3 years ago
- Resources for the MRQA 2019 Shared Task☆294Aug 5, 2021Updated 5 years ago
- TyDi QA contains 200k human-annotated question-answer pairs in 11 Typologically Diverse languages, written without seeing the answer and …☆319May 28, 2020Updated 6 years ago
- Code and data to support the paper "PAQ 65 Million Probably-Asked Questions andWhat You Can Do With Them"☆210Aug 31, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Datasets for Question Answering by Search and Reading☆69Jan 19, 2018Updated 8 years ago
- ☆435Feb 4, 2024Updated 2 years ago
- ACL2020 Tutorial: Open-Domain Question Answering☆837Jan 1, 2021Updated 5 years ago
- Library for Knowledge Intensive Language Tasks☆978Mar 31, 2022Updated 4 years ago
- XLNet: Generalized Autoregressive Pretraining for Language Understanding☆6,188May 28, 2023Updated 3 years ago
- Phrase-Indexed Question Answering (PIQA)☆93Apr 27, 2019Updated 7 years ago
- Code for the TriviaQA reading comprehension dataset☆341Apr 5, 2024Updated 2 years ago
- PyTorch original implementation of Cross-lingual Language Model Pretraining.☆2,921Feb 14, 2023Updated 3 years ago
- Code for the paper "Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer"☆6,556Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- New dataset☆311Aug 31, 2021Updated 5 years ago
- ☆603Apr 26, 2021Updated 5 years ago
- A toolkit for evaluating the linguistic knowledge and transferability of contextual representations. Code for "Linguistic Knowledge and T…☆212Oct 20, 2021Updated 4 years ago
- This repository contains the NarrativeQA dataset. It includes the list of documents with Wikipedia summaries, links to full stories, and …☆520Apr 15, 2020Updated 6 years ago
- Authors' implementation of EMNLP-IJCNLP 2019 paper "Answering Complex Open-domain Questions Through Iterative Query Generation"☆196Oct 29, 2019Updated 6 years ago
- An original implementation of ACL 2019, "Multi-hop Reading Comprehension through Question Decomposition and Rescoring"☆138Apr 23, 2022Updated 4 years ago
- An open-source NLP research library, built on PyTorch.☆11,883Nov 22, 2022Updated 3 years ago
- Bi-directional Attention Flow (BiDAF) network is a multi-stage hierarchical process that represents context at different levels of granul…☆1,545May 31, 2023Updated 3 years ago
- ☆32Jun 19, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Reading Wikipedia to Answer Open-Domain Questions☆4,467Oct 1, 2023Updated 3 years ago
- This dataset contains 108,463 human-labeled and 656k noisily labeled pairs that feature the importance of modeling structure, context, an…☆572Jan 4, 2022Updated 4 years ago
- MS MARCO(Microsoft Machine Reading Comprehension) is a large scale dataset focused on machine reading comprehension and question answerin…☆235Jun 12, 2023Updated 3 years ago
- Fusion-in-Decoder☆595Oct 4, 2023Updated 3 years ago
- Multi-Task Deep Neural Networks for Natural Language Understanding☆2,257Mar 7, 2024Updated 2 years ago
- Scripts and links to recreate the ELI5 dataset.☆325Aug 31, 2021Updated 5 years ago
- jiant is an nlp toolkit☆1,675Jul 6, 2023Updated 3 years ago
- A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.☆2,307Oct 16, 2025Updated 11 months ago
- An original implementation of EMNLP 2020, "AmbigQA: Answering Ambiguous Open-domain Questions"☆123Apr 23, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Source code and dataset for ACL 2019 paper "ERNIE: Enhanced Language Representation with Informative Entities"☆1,417Jan 10, 2024Updated 2 years ago
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than Generators☆2,365Mar 23, 2024Updated 2 years ago
- [ACL 2021] Learning Dense Representations of Phrases at Scale; EMNLP'2021: Phrase Retrieval Learns Passage Retrieval, Too https://arxiv.o…☆606Jun 15, 2022Updated 4 years ago
- A novel embedding training algorithm leveraging ANN search and achieved SOTA retrieval on Trec DL 2019 and OpenQA benchmarks☆390Jan 6, 2026Updated 9 months ago
- Pyserini is a Python toolkit for reproducible information retrieval research with sparse and dense representations.☆2,168Updated this week
- The Natural Language Decathlon: A Multitask Challenge for NLP☆2,335May 1, 2025Updated last year
- EMNLP 2021 - Pre-training architectures for dense retrieval☆256Mar 18, 2022Updated 4 years ago