A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world speech phenomena.
☆33Aug 9, 2026Updated last month
Alternatives and similar repositories for hearing2translate
Users that are interested in hearing2translate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Platform for Evaluating and Reviewing of Multilingual Tasks☆34Aug 21, 2026Updated 3 weeks ago
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆54Updated this week
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆30Jul 9, 2026Updated 2 months ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Getting confidences from any end-to-end systems☆11May 24, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆21Mar 6, 2026Updated 6 months ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 9 months ago
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated 2 years ago
- Onnx compatible styletts2 code☆16Apr 4, 2026Updated 5 months ago
- Lightweight self-hosted span annotation tool☆44Updated this week
- ☆21Feb 13, 2023Updated 3 years ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆45Feb 9, 2023Updated 3 years ago
- ☆53Jul 25, 2024Updated 2 years ago
- ☆23Apr 8, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆17Mar 31, 2025Updated last year
- A framework for evaluating Machine Translation models.☆14Aug 7, 2026Updated last month
- Pure-PyTorch Parakeet TDT inference☆54Mar 10, 2026Updated 6 months ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speech☆52Oct 8, 2021Updated 4 years ago
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated last year
- ☆32Aug 18, 2026Updated last month
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 11 months ago
- Audio Research in US. US-based professors who work on audio (music, speech, acoustics). For students who would like to apply for RA, PhD,…☆33Feb 27, 2026Updated 6 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆30Apr 9, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 9 months ago
- A system for live lecture translation (speech to text) where the audience can easily provide corrections.☆15Aug 4, 2024Updated 2 years ago
- Speed-optimized streaming neural speech enhancement network☆153Jul 3, 2026Updated 2 months ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆45Apr 11, 2026Updated 5 months ago
- As a Pangolin looks for bugs and catches them, the goal of this library is ot help developers finding bugs in their neural networks and n…☆14Apr 29, 2026Updated 4 months ago
- Python tool for simple Czech alphabetization☆14Jul 27, 2026Updated last month
- Deepspeech ASR Model for the Catalan Language☆17Feb 15, 2021Updated 5 years ago
- SubER - Subtitle Edit Rate☆28May 7, 2026Updated 4 months ago
- The code for Generative Locally Linear Embedding (GLLE).☆14Nov 4, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- Evaluate the quality of SRT files using the multilingual multimodal SONAR model.☆15May 18, 2024Updated 2 years ago
- OutboundEval, a comprehensive benchmark for evaluating large language models (LLMs) in expert-level intelligent outbound calling scenario…☆17Oct 28, 2025Updated 10 months ago
- Demo for AudioSAE paper☆16Apr 26, 2026Updated 4 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 11 months ago
- ☆21Apr 9, 2026Updated 5 months ago
- ☆37Jan 6, 2026Updated 8 months ago