A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world speech phenomena.
☆32Apr 25, 2026Updated 2 months ago
Alternatives and similar repositories for hearing2translate
Users that are interested in hearing2translate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆16Updated this week
- Platform for Evaluating and Reviewing of Multilingual Tasks☆32Updated this week
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆29Jul 9, 2026Updated last week
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Getting confidences from any end-to-end systems☆11May 24, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆18Mar 6, 2026Updated 4 months ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 8 months ago
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated last year
- Onnx compatible styletts2 code☆16Apr 4, 2026Updated 3 months ago
- Lightweight self-hosted span annotation tool☆44Apr 20, 2026Updated 3 months ago
- ☆21Feb 13, 2023Updated 3 years ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆44Feb 9, 2023Updated 3 years ago
- ☆51Jul 25, 2024Updated last year
- ☆23Apr 8, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆16Mar 31, 2025Updated last year
- A framework for evaluating Machine Translation models.☆13Apr 21, 2026Updated 2 months ago
- Pure-PyTorch Parakeet TDT inference☆48Mar 10, 2026Updated 4 months ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speech☆52Oct 8, 2021Updated 4 years ago
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated 11 months ago
- Audio Research in US. US-based professors who work on audio (music, speech, acoustics). For students who would like to apply for RA, PhD,…☆27Feb 27, 2026Updated 4 months ago
- ☆28Updated this week
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 10 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆27Apr 9, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- OutboundEval, a comprehensive benchmark for evaluating large language models (LLMs) in expert-level intelligent outbound calling scenario…☆16Oct 28, 2025Updated 8 months ago
- A system for live lecture translation (speech to text) where the audience can easily provide corrections.☆14Aug 4, 2024Updated last year
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 7 months ago
- Speed-optimized streaming neural speech enhancement network☆133Jul 3, 2026Updated 2 weeks ago
- As a Pangolin looks for bugs and catches them, the goal of this library is ot help developers finding bugs in their neural networks and n…☆14Apr 29, 2026Updated 2 months ago
- ☆12Apr 19, 2022Updated 4 years ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆40Apr 11, 2026Updated 3 months ago
- Python tool for simple Czech alphabetization☆14Jul 12, 2023Updated 3 years ago
- Evaluate the quality of SRT files using the multilingual multimodal SONAR model.☆15May 18, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Native End-to-End Full-Duplex Spoken Language Model☆54Updated this week
- Deepspeech ASR Model for the Catalan Language☆17Feb 15, 2021Updated 5 years ago
- The code for Generative Locally Linear Embedding (GLLE).☆14Nov 4, 2021Updated 4 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- ☆21Jun 4, 2026Updated last month
- Demo for AudioSAE paper☆15Apr 26, 2026Updated 2 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆34Sep 25, 2025Updated 9 months ago