Repository for DEMETR: Diagnosing Evaluation Metrics for Translation
☆17Nov 29, 2022Updated 3 years ago
Alternatives and similar repositories for demetr
Users that are interested in demetr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official codebase accompanying our ACL 2022 paper "RELiC: Retrieving Evidence for Literary Claims" (https://relic.cs.umass.edu).☆20May 14, 2022Updated 4 years ago
- ☆29Dec 2, 2024Updated last year
- ☆24Apr 2, 2024Updated 2 years ago
- ☆12Sep 1, 2021Updated 5 years ago
- ☆40Dec 17, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Building and Using A Seed Corpus for the Human Language Project☆11Feb 9, 2018Updated 8 years ago
- Python package to augment multilingual data☆15Feb 15, 2023Updated 3 years ago
- Code and data for the paper "Disentangling Uncertainty in Machine Translation Evaluation", accepted at EMNLP 2022.☆23Jun 23, 2023Updated 3 years ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 10 months ago
- Code and data for the IWSLT 2022 shared task on Formality Control for SLT☆22May 24, 2023Updated 3 years ago
- ☆21Dec 8, 2022Updated 3 years ago
- Neural Fuzzy Repair (NFR) is a data augmentation pipeline, which integrates fuzzy matches (i.e. similar translations) into neural machine…☆12Aug 14, 2024Updated 2 years ago
- [ChatGPT4MTevaluation] ErrorAnalysis Prompt for MT Evaluation in ChatGPT☆91Oct 14, 2025Updated 11 months ago
- Official repository for our EACL 2023 paper "LongEval: Guidelines for Human Evaluation of Faithfulness in Long-form Summarization" (https…☆45Aug 10, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Official repository with code and data accompanying the NAACL 2021 paper "Hurdles to Progress in Long-form Question Answering" (https://a…☆46Jul 30, 2022Updated 4 years ago
- Official code and model checkpoints for our EMNLP 2022 paper "RankGen - Improving Text Generation with Large Ranking Models" (https://arx…☆140Aug 2, 2023Updated 3 years ago
- Code & data for EMNLP 2020 paper "MOCHA: A Dataset for Training and Evaluating Reading Comprehension Metrics".☆16May 3, 2022Updated 4 years ago
- A method for evaluating the high-level coherence of machine-generated texts. Identifies high-level coherence issues in transformer-based …☆12Mar 18, 2023Updated 3 years ago
- ☆54Oct 24, 2024Updated last year
- Suri: Multi-constraint instruction following for long-form text generation [EMNLP’24]☆27Oct 3, 2025Updated 11 months ago
- ☆13Jul 15, 2024Updated 2 years ago
- A repository with the code related to experiments around context-aware machine translation☆51Sep 22, 2025Updated last year
- Minimum Bayes Risk Decoding for Hugging Face Transformers☆61Jun 3, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆22Sep 19, 2023Updated 3 years ago
- Story understanding and plot analysis pilot.☆10Dec 27, 2022Updated 3 years ago
- A Neural Framework for MT Evaluation☆779Apr 21, 2026Updated 5 months ago
- Examples for the Spartan HPC cluster.☆10Sep 2, 2019Updated 7 years ago
- codes for "Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models"☆13Feb 10, 2025Updated last year
- explainable-machine-translation-metrics☆12Jul 15, 2022Updated 4 years ago
- GEMBA — GPT Estimation Metric Based Assessment☆156Dec 15, 2025Updated 9 months ago
- Multilingual Quality Estimation and Automatic Post-editing Dataset☆44Mar 24, 2022Updated 4 years ago
- Official repository for our NeurIPS 2023 paper "Paraphrasing evades detectors of AI-generated text, but retrieval is an effective defense…☆206Nov 9, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- EMNLP DiscoEval paper☆43Nov 12, 2019Updated 6 years ago
- UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation☆59Oct 13, 2020Updated 5 years ago
- Text Simplification System and Dataset☆15Jul 19, 2017Updated 9 years ago
- ☆33Nov 22, 2021Updated 4 years ago
- PyTorch code for Improving Commonsense in Vision-Language Models via Knowledge Graph Riddles (DANCE)☆22Nov 29, 2022Updated 3 years ago
- The Shmoop Corpus☆17Oct 27, 2020Updated 5 years ago
- Resources for paper "DialSummEval: Revisiting summarization evaluation for dialogues"☆14Jul 22, 2025Updated last year