simulstream is a Python library for simultaneous/streaming speech recognition and translation. It enables both the simulation with existing files to score systems, like in the SimulEval project, and the possibility to run demos on a browser.
☆30Jul 9, 2026Updated last month
Alternatives and similar repositories for simulstream
Users that are interested in simulstream are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains code used for the MCIF dataset and IWSLT 2025 Instruction Following shared task. This includes scripts used to c…☆16Jul 17, 2026Updated last month
- ☆25May 27, 2026Updated 3 months ago
- ☆13Aug 23, 2024Updated 2 years ago
- A unified evaluation suite for speech-to-text translation, covering SpeechLLMs, SFMs, and cascaded systems across diverse real-world spee…☆33Aug 9, 2026Updated 3 weeks ago
- Crowdsourcing of hard to translate inputs (texts, images, audios) at scale.☆35Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The project for speech translation☆12Sep 28, 2023Updated 2 years ago
- RealSI: Open Benchmark for Simultaneous Interpretation in Real-world Scenarios☆86Jul 4, 2025Updated last year
- [ACL 2024] An easily extensible framework for simultaneous, text-to-text neural machine translation (SimulMT) for LLMs.☆18Apr 21, 2025Updated last year
- This project uses gpt-4 to build agents to play one night werewolf.☆10Jul 14, 2023Updated 3 years ago
- SimulEval: A General Evaluation Toolkit for Simultaneous Translation☆126Sep 13, 2024Updated last year
- ☆39Jul 21, 2026Updated last month
- SubER - Subtitle Edit Rate☆26May 7, 2026Updated 3 months ago
- A framework for evaluating Machine Translation models.☆14Aug 7, 2026Updated 3 weeks ago
- SHAS: Approaching optimal Segmentation for End-to-End Speech Translation☆44Feb 9, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSI…☆25Oct 8, 2025Updated 10 months ago
- Multi-talker ASR based on DiCoW with Serialized Output Training☆21Sep 18, 2025Updated 11 months ago
- A sound device mixer made with rofi☆22Aug 20, 2025Updated last year
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- LLaST: Improved End-to-end Speech Translation System Leveraged by Large Language Models☆26Aug 11, 2024Updated 2 years ago
- Code for Q-learning with parametrized quantum circuits in OpenAI Gym environments.☆14Nov 12, 2021Updated 4 years ago
- ☆31Aug 18, 2026Updated last week
- A python implementation of Maximum Variance Unfolding using CVXPY, Numpy, Scipy, and SK-Learn.☆16Nov 15, 2018Updated 7 years ago
- ☆23Sep 29, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Collection of Open Source Speech Data☆166Oct 3, 2025Updated 10 months ago
- The code used, and a docker image to run it, of the paper `Exploiting locality and physical invariants to design effective Deep Reinforce…☆13Dec 10, 2019Updated 6 years ago
- open-source Mandarian biased word dataset☆14Sep 21, 2023Updated 2 years ago
- ☆37Jan 6, 2026Updated 7 months ago
- ☆32Apr 24, 2024Updated 2 years ago
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 2 months ago
- Pure-PyTorch inference for CohereLabs/cohere-transcribe-03-2026 (2B Conformer + Transformer ASR, 14 languages).☆44Apr 29, 2026Updated 4 months ago
- ☆14Apr 4, 2025Updated last year
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The geometry of multilingual language model representations (EMNLP 2022).☆22Oct 21, 2022Updated 3 years ago
- Example of easily implementing custom tensor operations in C and CUDA.☆15Apr 30, 2020Updated 6 years ago
- Implemented Laplacian Eigenmaps☆21Oct 7, 2021Updated 4 years ago
- ☆26May 19, 2022Updated 4 years ago
- The Full-Duplex Interaction Track of the ICASSP 2026 Human-like Spoken Dialogue Systems Challenge aims to advance the evaluation of full-…☆38Apr 27, 2026Updated 4 months ago
- Implementation of "Look, Listen and Recognise:character-aware audio-visual subtitling"☆21Nov 3, 2025Updated 9 months ago
- MAGIC-TTS: Fine-Grained Controllable Speech Synthesis with Explicit Local Duration and Pause Control☆53Apr 28, 2026Updated 4 months ago