Seahorse is a dataset for multilingual, multi-faceted summarization evaluation. It consists of 96K summaries with human ratings along 6 quality dimensions: comprehensibility, repetition, grammar, attribution, main idea(s), and conciseness, covering 6 languages, 9 systems and 4 datasets.
☆90Feb 27, 2024Updated 2 years ago
Alternatives and similar repositories for seahorse
Users that are interested in seahorse are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆41Jun 7, 2023Updated 3 years ago
- Russian coreference resolution competition☆11Mar 24, 2023Updated 3 years ago
- Crawling engine that crawls a set of top-level domains looking for documents in a list of languages☆11Feb 6, 2024Updated 2 years ago
- Converter for Rhetorical Structure Theory (RST) trees to dependency representation☆17Aug 21, 2025Updated last year
- QAmeleon introduces synthetic multilingual QA data using PaLM, a 540B large language model. This dataset was generated by prompt tuning P…☆34Aug 15, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Resources for the "SummEval: Re-evaluating Summarization Evaluation" paper☆414Jun 23, 2024Updated 2 years ago
- “Style Transfer as Data Augmentation: A Case Study on Named Entity Recognition” (EMNLP 2022)☆16Feb 2, 2023Updated 3 years ago
- Repository for DISRPT2023 shared task☆17Jul 26, 2024Updated 2 years ago
- This is a repo for DCQA QUD parsing implemenation☆12Aug 5, 2025Updated last year
- ☆25Nov 1, 2022Updated 3 years ago
- Repository with code for MaChAmp: https://aclanthology.org/2021.eacl-demos.22/☆92Jun 3, 2026Updated 3 months ago
- ☆18Jun 18, 2021Updated 5 years ago
- ☆31Apr 14, 2023Updated 3 years ago
- How (but not why) to do Twitter sociolinguistic analysis in the Unix Shell☆10Apr 19, 2016Updated 10 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Materials for "Quantifying the Plausibility of Context Reliance in Neural Machine Translation" at ICLR'24 🐑 🐑☆16Apr 18, 2024Updated 2 years ago
- Course for Interpreting ML Models☆51Feb 16, 2023Updated 3 years ago
- ☆25May 11, 2024Updated 2 years ago
- The codebase for our ACL2023 paper: Did You Read the Instructions? Rethinking the Effectiveness of Task Definitions in Instruction Learni…☆30Jul 16, 2023Updated 3 years ago
- Python framework for processing Universal Dependencies data☆58Jun 26, 2026Updated 2 months ago
- Enhancing Sentence Embedding with Generalized Pooling☆20Oct 4, 2022Updated 3 years ago
- [NeurIPS 2024] 🕸 GlotCC Dataset and Pipline☆21Apr 6, 2025Updated last year
- преобразования регулярных выражений и коне чных автоматов☆21Feb 26, 2025Updated last year
- Project OCELoT: an Open, Collaborative Evaluation Leaderboard of Translations☆23Jul 11, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Witwicky: An implementation of Transformer in PyTorch.☆22Aug 17, 2020Updated 6 years ago
- SWIM-IR is a Synthetic Wikipedia-based Multilingual Information Retrieval training set with 28 million query-passage pairs spanning 33 la…☆50Nov 13, 2023Updated 2 years ago
- ☆27Nov 29, 2022Updated 3 years ago
- ☆24Jun 12, 2023Updated 3 years ago
- A genetic algorithm to find optimal solutions for TSP (Travelling Salesman Problem) using the CUDA Architecture (GPU)☆18Aug 21, 2015Updated 11 years ago
- Code of Robust Lottery Tickets for Pre-trained Language Models (ACL2022)☆20Jul 18, 2022Updated 4 years ago
- One implementation of the paper "Coreference-Aware Dialogue Summarization".☆20Nov 9, 2023Updated 2 years ago
- https://arxiv.org/abs/2201.06499☆29Apr 9, 2024Updated 2 years ago
- The code implementation of the EMNLP2022 paper: DisCup: Discriminator Cooperative Unlikelihood Prompt-tuning for Controllable Text Gene…☆27Nov 13, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- HANNA, a large annotated dataset of Human-ANnotated NArratives for ASG evaluation.☆39Oct 15, 2024Updated last year
- The official code of EMNLP 2022, "How Far are We from Robust Long Abstractive Summarization?".☆18Sep 26, 2023Updated 2 years ago
- Defeasible Natural Language Inference☆14Dec 4, 2020Updated 5 years ago
- Automatically harvested multilingual contrastive word sense disambiguation test sets for machine translation☆18Jan 18, 2021Updated 5 years ago
- A High-Quality Multilingual Dataset for Structured Documentation Translation☆39May 1, 2025Updated last year
- Implementation of the paper "FactGraph: Evaluating Factuality in Summarization with Semantic Graph Representations (NAACL 2022)"☆52Jul 26, 2023Updated 3 years ago
- Materials for ACL-2022 tutorial: A Gentle Introduction to Deep Nets and Opportunities for the Future☆17May 24, 2022Updated 4 years ago