SHAS: Approaching optimal Segmentation for End-to-End Speech Translation
☆44Feb 9, 2023Updated 3 years ago
Alternatives and similar repositories for SHAS
Users that are interested in SHAS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Systems submitted to IWSLT 2021 by the MT-UPC group.☆14Feb 23, 2023Updated 3 years ago
- ☆13Aug 23, 2024Updated last year
- ☆35Sep 1, 2022Updated 3 years ago
- Repository containing the open source code of works published at the FBK MT unit.☆60Mar 19, 2026Updated 4 months ago
- Pushing the Limits of Zero-shot End-to-End Speech Translation☆25Dec 12, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Unsupervised Voice Activity Detection by Modeling Source and System Information using Zero Frequency Filtering☆23Oct 19, 2023Updated 2 years ago
- A library for data streaming and augmentation☆22May 5, 2025Updated last year
- A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world spee…☆33Updated this week
- SubER - Subtitle Edit Rate☆26May 7, 2026Updated 3 months ago
- Prabhupadavani: A Code-mixed Speech Translation Data for 25 languages☆13Oct 12, 2022Updated 3 years ago
- Measuring the Mixing of Contextual Information in the Transformer☆35May 27, 2023Updated 3 years ago
- This is a repository for a paper accepted at the 2022 IEEE Spoken Language Technology Workshop (SLT 2022)☆17Dec 1, 2022Updated 3 years ago
- wake-up word emotion recognition [APSIPA 2022]☆17Nov 11, 2022Updated 3 years ago
- phoneme tokenizer and grapheme-to-phoneme model for 8k languages☆174Jun 9, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation for Fast-HuBERT: An Efficient Training Framework for Self-Supervised Speech Representation Learning☆100Nov 20, 2024Updated last year
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated 2 years ago
- StyleTTS2 + Vocos as a Decoder☆13Jul 31, 2026Updated last week
- Code and data for the IWSLT 2022 shared task on Formality Control for SLT☆22May 24, 2023Updated 3 years ago
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 6 months ago
- ☆16Updated this week
- This will hold the crowdsourcing platform to be used to store voice data from various speakers which will act as input dataset for speech…☆17Mar 6, 2023Updated 3 years ago
- Spoken Language Translation System☆20Jul 26, 2021Updated 5 years ago
- ☆16Jun 13, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Neural end-to-end Speech Translation Toolkit☆306Jun 28, 2022Updated 4 years ago
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆30Jul 9, 2026Updated last month
- An extension of thu-spmi/CAT which contains a full-fledged implementation of CTC-CRF for Tensorflow.☆12Jul 5, 2021Updated 5 years ago
- End-to-end Speech Translation☆35Apr 12, 2021Updated 5 years ago
- BurrMill core☆22Nov 2, 2021Updated 4 years ago
- AI based singing voice synthesis database generator☆13Aug 12, 2022Updated 4 years ago
- Redesign of the new version of the QB-Inventory☆13Feb 19, 2024Updated 2 years ago
- A Benchmark Corpus for Low-Resource Cantonese Punctuation Restoration from Speech Transcripts☆15Dec 3, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- text to speech☆10Mar 19, 2024Updated 2 years ago
- A program to generate microphone wind noise audio. Ideal for generating example data for designing noise removal algorithms.☆19Jun 4, 2018Updated 8 years ago
- Official repository of the work "Low-complexity Unsupervised Audio Anomaly Detection exploiting Separable Convolutions and Angular Loss" …☆11Nov 6, 2024Updated last year
- ☆17Jun 2, 2025Updated last year
- ☆16May 15, 2019Updated 7 years ago
- ☆81Aug 8, 2025Updated last year
- Automatic speech annotator processing speech with voice activaty detection, overlapping speech detection, speaker diarization and automat…☆33Jun 14, 2024Updated 2 years ago