Large scale (>200h) and publicly available read audio book corpus. This corpus is an augmentation of LibriSpeech ASR Corpus (1000h) and contains English utterances (from audiobooks) automatically aligned with French text. Our dataset offers ~236h of speech aligned to translated text.
☆44Jul 9, 2022Updated 4 years ago
Alternatives and similar repositories for Translation-Augmented-LibriSpeech-Corpus
Users that are interested in Translation-Augmented-LibriSpeech-Corpus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Fisher and CALLHOME Spanish–English Speech Translation Corpus☆41Feb 10, 2022Updated 4 years ago
- ☆10Dec 21, 2022Updated 3 years ago
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- An adaptation of Fairseq to (End-to-end) speech translation.☆22Jun 1, 2022Updated 4 years ago
- ☆12Aug 31, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆14Nov 16, 2022Updated 3 years ago
- eXtensible Neural Machine Translation☆189Sep 22, 2025Updated last year
- Joint Source-Target Self Attention with Locality Constraints☆20May 9, 2020Updated 6 years ago
- ☆34Updated this week
- Spoken Language Translation System☆20Jul 26, 2021Updated 5 years ago
- Zero -- A neural machine translation system☆152May 8, 2023Updated 3 years ago
- Some other AutoML algorithms as baselines.☆12Apr 2, 2019Updated 7 years ago
- ☆178Nov 10, 2021Updated 4 years ago
- Translation Error Rate (TER)☆44May 25, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆18Oct 16, 2018Updated 7 years ago
- ☆22Sep 24, 2018Updated 8 years ago
- Tracking the progress in end-to-end speech translation☆260Oct 25, 2023Updated 2 years ago
- An experiment with some brain-inspired ideas.☆14Apr 18, 2018Updated 8 years ago
- ☆24Sep 20, 2024Updated 2 years ago
- This repository contains code and metadata of How2 dataset☆193Dec 30, 2024Updated last year
- Modified version of fairseq, including new implementations for criterions using reinforcement learning methods.☆11Aug 14, 2019Updated 7 years ago
- Implementation of CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning☆48Nov 8, 2023Updated 2 years ago
- Neural end-to-end Speech Translation Toolkit☆305Jun 28, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation for "Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling"☆38Aug 8, 2017Updated 9 years ago
- Fast Gabor spectral transforms in Java. Using a JNI bridge with the gaborator C++ library.☆14Jan 20, 2023Updated 3 years ago
- Attention-based sequence to sequence learning☆387May 9, 2019Updated 7 years ago
- ☆11Jul 15, 2020Updated 6 years ago
- Toolkit for Reproducible Execution of Speech, Text and Language Experiments☆11Mar 24, 2026Updated 6 months ago
- Fast and differentiable hidden Markov model in C++☆19Jan 20, 2023Updated 3 years ago
- Python code of the paper Model order reduction of deep structured state-space models: A system-theoretic approach☆14Nov 22, 2024Updated last year
- Simultaneous NMT/MMT framework in PyTorch☆38Mar 22, 2025Updated last year
- The implementation for "Empowering Whisper as a Joint Multi-Talker and Target-Talker Speech Recognition System".☆34Aug 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A dataset of Hip Hop samples for Music Information Retrieval research☆11Jun 1, 2016Updated 10 years ago
- The implementation for "Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions"☆51Apr 7, 2025Updated last year
- Facebook AI Research Sequence-to-Sequence Toolkit written in Python.☆12Oct 10, 2020Updated 5 years ago
- Official implementation of SIGIR 2022 Paper "Task-Oriented Dialogue System as Natural Language Generation".☆14Apr 6, 2022Updated 4 years ago
- CoVoST: A Large-Scale Multilingual Speech-To-Text Translation Corpus (CC0 Licensed)☆401Sep 14, 2021Updated 5 years ago
- End-to-end Speech Translation☆35Apr 12, 2021Updated 5 years ago
- ☆13Mar 23, 2026Updated 6 months ago