A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world speech phenomena.
☆32Aug 8, 2026Updated this week
Alternatives and similar repositories for hearing2translate
Users that are interested in hearing2translate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Platform for Evaluating and Reviewing of Multilingual Tasks☆32Updated this week
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆25Updated this week
- simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existi…☆30Jul 9, 2026Updated last month
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- Getting confidences from any end-to-end systems☆11May 24, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆20Mar 6, 2026Updated 5 months ago
- [EMNLP2025] Remedy: Learning Machine Translation Evaluation from Human Preferences with Reward Modeling☆16Nov 20, 2025Updated 8 months ago
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated last year
- Onnx compatible styletts2 code☆16Apr 4, 2026Updated 4 months ago
- Lightweight self-hosted span annotation tool☆44Apr 20, 2026Updated 3 months ago
- ☆21Feb 13, 2023Updated 3 years ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆44Feb 9, 2023Updated 3 years ago
- ☆51Jul 25, 2024Updated 2 years ago
- ☆23Apr 8, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆16Mar 31, 2025Updated last year
- A framework for evaluating Machine Translation models.☆13Updated this week
- Pure-PyTorch Parakeet TDT inference☆52Mar 10, 2026Updated 4 months ago
- This repository describes our reproducible framework for assessing self-supervised representation learning from speech☆52Oct 8, 2021Updated 4 years ago
- This repository implement a novel zero-shot TTS framework, named Flamed-TTS, focusing on the efficient generation and dynamic pacing in …☆57Aug 9, 2025Updated last year
- Audio Research in US. US-based professors who work on audio (music, speech, acoustics). For students who would like to apply for RA, PhD,…☆28Feb 27, 2026Updated 5 months ago
- ☆30Jul 17, 2026Updated 3 weeks ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 10 months ago
- Ultra-Sortformer for Scalable Speaker Diarization☆27Apr 9, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- OutboundEval, a comprehensive benchmark for evaluating large language models (LLMs) in expert-level intelligent outbound calling scenario…☆17Oct 28, 2025Updated 9 months ago
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 8 months ago
- Speed-optimized streaming neural speech enhancement network☆144Jul 3, 2026Updated last month
- A system for live lecture translation (speech to text) where the audience can easily provide corrections.☆15Aug 4, 2024Updated 2 years ago
- ☆12Apr 19, 2022Updated 4 years ago
- Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasonin…☆42Apr 11, 2026Updated 3 months ago
- As a Pangolin looks for bugs and catches them, the goal of this library is ot help developers finding bugs in their neural networks and n…☆14Apr 29, 2026Updated 3 months ago
- SubER - Subtitle Edit Rate☆26May 7, 2026Updated 3 months ago
- Python tool for simple Czech alphabetization☆14Jul 27, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Deepspeech ASR Model for the Catalan Language☆17Feb 15, 2021Updated 5 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- The code for Generative Locally Linear Embedding (GLLE).☆14Nov 4, 2021Updated 4 years ago
- Evaluate the quality of SRT files using the multilingual multimodal SONAR model.☆15May 18, 2024Updated 2 years ago
- Demo for AudioSAE paper☆15Apr 26, 2026Updated 3 months ago
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 10 months ago
- ☆19Apr 9, 2026Updated 4 months ago