A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world speech phenomena.
☆33Aug 9, 2026Updated 2 months ago
Alternatives and similar repositories for hearing2translate
Users that are interested in hearing2translate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Platform for Evaluating and Reviewing of Multilingual Tasks☆35Aug 21, 2026Updated last month
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆62Updated this week
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆31Sep 30, 2026Updated last week
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Getting confidences from any end-to-end systems☆11May 24, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Mar 6, 2026Updated 7 months ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 10 months ago
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated 2 years ago
- Onnx compatible styletts2 code☆16Apr 4, 2026Updated 6 months ago
- Lightweight self-hosted span annotation tool☆44Sep 18, 2026Updated 3 weeks ago
- ☆21Feb 13, 2023Updated 3 years ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆45Feb 9, 2023Updated 3 years ago
- ☆53Jul 25, 2024Updated 2 years ago
- ☆23Apr 8, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆17Mar 31, 2025Updated last year
- A framework for evaluating Machine Translation models.☆14Oct 1, 2026Updated last week
- Pure-PyTorch Parakeet TDT inference☆54Mar 10, 2026Updated 6 months ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speech☆52Oct 8, 2021Updated 5 years ago
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated last year
- ☆32Sep 24, 2026Updated 2 weeks ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated last year
- Audio Research in US. US-based professors who work on audio (music, speech, acoustics). For students who would like to apply for RA, PhD,…☆33Feb 27, 2026Updated 7 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆30Apr 9, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 10 months ago
- A system for live lecture translation (speech to text) where the audience can easily provide corrections.☆15Aug 4, 2024Updated 2 years ago
- Speed-optimized streaming neural speech enhancement network☆157Sep 28, 2026Updated last week
- ☆12Apr 19, 2022Updated 4 years ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆46Apr 11, 2026Updated 5 months ago
- As a Pangolin looks for bugs and catches them, the goal of this library is ot help developers finding bugs in their neural networks and n…☆14Apr 29, 2026Updated 5 months ago
- Python tool for simple Czech alphabetization☆14Jul 27, 2026Updated 2 months ago
- Deepspeech ASR Model for the Catalan Language☆17Feb 15, 2021Updated 5 years ago
- SubER - Subtitle Edit Rate☆28May 7, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The code for Generative Locally Linear Embedding (GLLE).☆14Nov 4, 2021Updated 4 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- Evaluate the quality of SRT files using the multilingual multimodal SONAR model.☆15May 18, 2024Updated 2 years ago
- OutboundEval, a comprehensive benchmark for evaluating large language models (LLMs) in expert-level intelligent outbound calling scenario…☆17Oct 28, 2025Updated 11 months ago
- Demo for AudioSAE paper☆16Apr 26, 2026Updated 5 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated last year
- ☆22Apr 9, 2026Updated 6 months ago