This repository contains code used for the MCIF dataset and IWSLT 2025 Instruction Following shared task. This includes scripts used to create test sets and their references, as well as scripts used in the evaluation.
☆16Jul 17, 2026Updated last month
Alternatives and similar repositories for mcif
Users that are interested in mcif are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Evaluate the quality of SRT files using the multilingual multimodal SONAR model.☆15May 18, 2024Updated 2 years ago
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆30Jul 9, 2026Updated last month
- Repository containing the open source code of works published at the FBK MT unit.☆60Mar 19, 2026Updated 5 months ago
- Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSI…☆25Oct 8, 2025Updated 10 months ago
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆35Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code and data for the IWSLT 2022 shared task on Formality Control for SLT☆22May 24, 2023Updated 3 years ago
- Code for extracting parallel corpora from pmindia☆17Jan 28, 2020Updated 6 years ago
- Collection of Open Source Speech Data☆166Oct 3, 2025Updated 10 months ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- ☆22Sep 19, 2023Updated 2 years ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆44Feb 9, 2023Updated 3 years ago
- MAchine Translation Evaluation Online (MATEO)☆26May 8, 2026Updated 3 months ago
- ☆24Feb 4, 2020Updated 6 years ago
- SimulEval: A General Evaluation Toolkit for Simultaneous Translation☆126Sep 13, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆178Nov 10, 2021Updated 4 years ago
- Bicleaner fork that uses neural networks☆40Feb 23, 2026Updated 6 months ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 2 months ago
- ☆34Jan 26, 2026Updated 7 months ago
- ☆67Mar 25, 2022Updated 4 years ago
- Listen Attend and Spell (LAS) implement in pytorch☆60Sep 4, 2018Updated 7 years ago
- BLOOM+1: Adapting BLOOM model to support a new unseen language☆75Mar 2, 2024Updated 2 years ago
- Promting Whisper for Audio-Visual Speech Recognition, Code-Switched Speech Recognition, and Zero-Shot Speech Translation☆151Jan 16, 2024Updated 2 years ago
- CoVoST: A Large-Scale Multilingual Speech-To-Text Translation Corpus (CC0 Licensed)☆401Sep 14, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆93Feb 13, 2024Updated 2 years ago
- A torch implementation of a recursion which turns out to be useful for RNN-T.☆148Aug 25, 2023Updated 3 years ago
- The ParroT framework to enhance and regulate the Translation Abilities during Chat based on open-sourced LLMs (e.g., LLaMA-7b, Bloomz-7b1…☆177Dec 31, 2024Updated last year
- ☆166Updated this week
- This repository contains code and metadata of How2 dataset☆192Dec 30, 2024Updated last year
- Official implementation of our RAL'24 paper: Multi-Camera Unified Pre-training for Autonomous Driving☆237Feb 15, 2024Updated 2 years ago
- A Pytorch Implementation of Transducer Model for End-to-End Speech Recognition☆238May 12, 2020Updated 6 years ago
- Code, Dataset, and Pretrained Models for Audio and Speech Large Language Model "Listen, Think, and Understand".☆478Apr 24, 2024Updated 2 years ago
- ☆242Nov 27, 2023Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Tracking the progress in end-to-end speech translation☆260Oct 25, 2023Updated 2 years ago
- VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Design☆647Sep 11, 2023Updated 2 years ago
- Unofficial Parallel WaveGAN (+ MelGAN & Multi-band MelGAN & HiFi-GAN & StyleMelGAN) with Pytorch☆1,645Apr 22, 2024Updated 2 years ago
- A neural word aligner based on multilingual BERT☆384Mar 10, 2022Updated 4 years ago
- INTERSPEECH 2023-2024 Papers: A complete collection of influential and exciting research papers from the INTERSPEECH 2023-24 conference. …☆684Dec 25, 2024Updated last year
- The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.☆450Jul 11, 2025Updated last year
- State-of-the-art LLM-based translation models.☆590Apr 9, 2025Updated last year