Large scale (>200h) and publicly available read audio book corpus. This corpus is an augmentation of LibriSpeech ASR Corpus (1000h) and contains English utterances (from audiobooks) automatically aligned with French text. Our dataset offers ~236h of speech aligned to translated text.
☆44Jul 9, 2022Updated 4 years ago
Alternatives and similar repositories for Translation-Augmented-LibriSpeech-Corpus
Users that are interested in Translation-Augmented-LibriSpeech-Corpus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Fisher and CALLHOME Spanish–English Speech Translation Corpus☆41Feb 10, 2022Updated 4 years ago
- Cross-Speaker Encoding Network for Multi-talker Speech Recognition☆12Mar 14, 2025Updated last year
- An adaptation of Fairseq to (End-to-end) speech translation.☆22Jun 1, 2022Updated 4 years ago
- ☆14Nov 16, 2022Updated 3 years ago
- ☆34Jun 12, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Spoken Language Translation System☆20Jul 26, 2021Updated 5 years ago
- SWIG bindings for Kaldi I/O, built with Conda☆15Dec 15, 2024Updated last year
- Zero -- A neural machine translation system☆152May 8, 2023Updated 3 years ago
- A Python module to generate large scale Music datasets using both Spotify and MusixMatch API's.☆43Aug 28, 2020Updated 5 years ago
- Some other AutoML algorithms as baselines.☆12Apr 2, 2019Updated 7 years ago
- Weakly Supervised CRNN System for Sound Event Detection With Large-scale Unlabeled In-domain Data☆11Oct 31, 2018Updated 7 years ago
- ☆178Nov 10, 2021Updated 4 years ago
- Translation Error Rate (TER)☆45May 25, 2018Updated 8 years ago
- ☆18Oct 16, 2018Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Tracking the progress in end-to-end speech translation☆260Oct 25, 2023Updated 2 years ago
- Layer normalization implemented in Keras☆60Jan 22, 2022Updated 4 years ago
- ☆24Sep 20, 2024Updated last year
- ADAPTING SELF-SUPERVISED MODELS TO MULTI-TALKER SPEECH RECOGNITION USING SPEAKER EMBEDDINGS☆34Mar 16, 2023Updated 3 years ago
- This repository contains code and metadata of How2 dataset☆192Dec 30, 2024Updated last year
- Modified version of fairseq, including new implementations for criterions using reinforcement learning methods.☆11Aug 14, 2019Updated 7 years ago
- Implementation of CoBERT: Self-Supervised Speech Representation Learning Through Code Representation Learning☆48Nov 8, 2023Updated 2 years ago
- Implementation for "Tying Word Vectors and Word Classifiers: A Loss Framework for Language Modeling"☆38Aug 8, 2017Updated 9 years ago
- Fast Gabor spectral transforms in Java. Using a JNI bridge with the gaborator C++ library.☆14Jan 20, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Evaluating methods to improve model transfer for intensive care unit models☆16Jul 6, 2023Updated 3 years ago
- machine translation and quality estimation☆35Jan 13, 2019Updated 7 years ago
- ☆11Jul 15, 2020Updated 6 years ago
- Toolkit for Reproducible Execution of Speech, Text and Language Experiments☆10Mar 24, 2026Updated 4 months ago
- Fast and differentiable hidden Markov model in C++☆19Jan 20, 2023Updated 3 years ago
- Python code of the paper Model order reduction of deep structured state-space models: A system-theoretic approach☆14Nov 22, 2024Updated last year
- The implementation for "Empowering Whisper as a Joint Multi-Talker and Target-Talker Speech Recognition System".☆34Aug 2, 2025Updated last year
- A dataset of Hip Hop samples for Music Information Retrieval research☆11Jun 1, 2016Updated 10 years ago
- Facebook AI Research Sequence-to-Sequence Toolkit written in Python.☆12Oct 10, 2020Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Experiments from "The Description Length of Deep Learning Models"☆10Aug 1, 2018Updated 8 years ago
- Official implementation of SIGIR 2022 Paper "Task-Oriented Dialogue System as Natural Language Generation".☆14Apr 6, 2022Updated 4 years ago
- CoVoST: A Large-Scale Multilingual Speech-To-Text Translation Corpus (CC0 Licensed)☆401Sep 14, 2021Updated 4 years ago
- ☆12Nov 17, 2023Updated 2 years ago
- End-to-end Speech Translation☆35Apr 12, 2021Updated 5 years ago
- ☆13Mar 23, 2026Updated 4 months ago
- Speech Transmission Index (STI) from real speech waveforms☆15May 1, 2011Updated 15 years ago