ictnlp/StreamSpeech

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/ictnlp/StreamSpeech)

ictnlp / StreamSpeech

StreamSpeech is an “All in One” seamless model for offline and simultaneous speech recognition, speech translation and speech synthesis.

☆1,279

Alternatives and similar repositories for StreamSpeech

Users that are interested in StreamSpeech are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

ictnlp / NAST-S2x
View on GitHub
A fast speech-to-speech & speech-to-text translation model that supports simultaneous decoding and offers 28× speedup.
☆78Oct 22, 2024Updated last year
superjcd / gocrawler
View on GitHub
gocrawler, go分布式爬虫框架
☆109Jun 4, 2024Updated 2 years ago
flutter-youni / flutter_youni_gromore
View on GitHub
Flutter的Gromore广告插件
☆115Apr 17, 2024Updated 2 years ago
Mactarvish / ocr-sample-generator
View on GitHub
OCR训练样本生成器，自动生成用于训练OCR检测和识别模型的图片样本和标注
☆133Aug 27, 2024Updated last year
ajwlforever / go-ratelimit-manager
View on GitHub
使用go实现单机式/分布式限流方案
☆110Mar 18, 2024Updated 2 years ago
Simple, predictable pricing with DigitalOcean hosting • Ad
Always know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
ictnlp / ComSpeech
View on GitHub
Code for ACL 2024 main conference paper "Can We Achieve High-quality Direct Speech-to-Speech Translation Without Parallel Speech Data?".
☆27Jul 2, 2024Updated 2 years ago
seraphembera / rust-snake
View on GitHub
Snake game with rust
☆32May 29, 2024Updated 2 years ago
Oldsquaw / Web
View on GitHub
☆47Jun 15, 2024Updated 2 years ago
joeljhou / RabbitMQ
View on GitHub
高并发实战-RabbitMQ消息队列入门指南
☆43Jun 15, 2024Updated 2 years ago
ictnlp / DASpeech
View on GitHub
Code for NeurIPS 2023 paper "DASpeech: Directed Acyclic Transformer for Fast and High-quality Speech-to-Speech Translation".
☆63Jul 22, 2024Updated 2 years ago
Foleyzhao / lacerate
View on GitHub
简单的静态博客生成器
☆159Mar 25, 2024Updated 2 years ago
X-LANCE / VoiceFlow-TTS
View on GitHub
[ICASSP 2024] This is the official code for "VoiceFlow: Efficient Text-to-Speech with Rectified Flow Matching"
☆376Sep 3, 2024Updated last year
KdaiP / StableTTS
View on GitHub
Next-generation TTS model using flow-matching and DiT, inspired by Stable Diffusion 3
☆437Sep 13, 2024Updated last year
WUHU-G / RCC_Transformer
View on GitHub
☆109Jun 13, 2024Updated 2 years ago
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
ictnlp / LSG
View on GitHub
The code for AAAI 2025 “Large Language Models Are Read/Write Policy-Makers for Simultaneous Generation”
☆15Jan 3, 2025Updated last year
line / LibriTTS-P
View on GitHub
LibriTTS-P: A Corpus with Speaking Style and Speaker Identity Prompts for Text-to-Speech and Style Captioning
☆161Jun 13, 2024Updated 2 years ago
WolfLink-DevTeam / lite-etl
View on GitHub
☆39Jun 11, 2024Updated 2 years ago
X-LANCE / SLAM-LLM
View on GitHub
A Framework for Speech, Language, Audio, Music Processing with Large Language Model
☆1,049Jan 15, 2026Updated 6 months ago
facebookresearch / AudioDec
View on GitHub
An Open-source Streaming High-fidelity Neural Audio Codec
☆510Mar 4, 2025Updated last year
ZhangXInFD / SpeechTokenizer
View on GitHub
This is the code for the SpeechTokenizer presented in the SpeechTokenizer: Unified Speech Tokenizer for Speech Language Models. Samples a…
☆658Jun 9, 2024Updated 2 years ago
jishengpeng / WavChat
View on GitHub
A Survey of Spoken Dialogue Models (60 pages)
☆316Nov 28, 2024Updated last year
shivammehta25 / Matcha-TTS
View on GitHub
[ICASSP 2024] 🍵 Matcha-TTS: A fast TTS architecture with conditional flow matching
☆1,335Jul 13, 2026Updated last week
skygazer42 / ForeSight
View on GitHub
机器学习与深度学习模型时间序列预测包
☆120May 3, 2026Updated 2 months ago
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
yangdongchao / SimpleSpeech
View on GitHub
The open source code for SimpleSpeech series
☆147Oct 8, 2024Updated last year
lmliheng / FastWebServer
View on GitHub
😎Focus on forwarding static resource web servers
☆69Jan 3, 2026Updated 6 months ago
JittorRepos / JDiffusion
View on GitHub
JDiffusion is a diffusion model library for generating images or videos based on Diffusers and Jittor.
☆226Nov 17, 2025Updated 8 months ago
haoxiangxu23 / stado
View on GitHub
Spatio-Temporal Action Detection with Occlusion
☆196Jun 19, 2024Updated 2 years ago
noctisynth / oblivion-rs
View on GitHub
A fast, lightweight, and full duplex secure end-to-end encryption protocol based on ECDHE
☆93Oct 28, 2025Updated 8 months ago
huutuongtu / Lightvoc
View on GitHub
LIGHTVOC AN UPSAMPLING-FREE GAN VOCODER BASED ON CONFORMER AND INVERSE SHORT-TIME FOURIER TRANSFORM
☆18May 17, 2024Updated 2 years ago
yangdongchao / UniAudio
View on GitHub
The Open Source Code of UniAudio
☆605Jul 22, 2024Updated 2 years ago
scutcsq / Neural-Transducers-for-Two-Stage-Text-to-Speech-via-Semantic-Token-Prediction
View on GitHub
Unofficial pytorch reproduction for the paper "Utilizing Neural Transducers for Two-Stage Text-to-Speech via Semantic Token Prediction" (…
☆60Apr 4, 2024Updated 2 years ago
DigitalPhonetics / IMS-Toucan
View on GitHub
Controllable and fast Text-to-Speech for over 7000 languages!
☆2,207Jan 25, 2026Updated 6 months ago
Wordpress hosting with auto-scaling - Free Trial Offer • Ad
Fully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
X-LANCE / StoryTTS
View on GitHub
[ICASSP 2024] StoryTTS: A Highly Expressive Text-to-Speech Dataset with Rich Textual Expressiveness Annotations
☆141Apr 27, 2024Updated 2 years ago
X-E-Speech / X-E-Speech-code
View on GitHub
X-E-Speech: Joint Training Framework of Non-Autoregressive Cross-lingual Emotional Text-to-Speech and Voice Conversion
☆112Apr 1, 2024Updated 2 years ago
842549829 / Panda
View on GitHub
Abp.vNext + EF Core The microservices Open source framework project supports the implementation of message push workflow certification ce…
☆177Dec 31, 2025Updated 6 months ago
0nutation / SpeechGPT
View on GitHub
SpeechGPT Series: Speech Large Language Models
☆1,402Jul 22, 2024Updated 2 years ago
QwenLM / Qwen2-Audio
View on GitHub
The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.
☆2,093Apr 21, 2025Updated last year
kyutai-labs / moshi
View on GitHub
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audi…
☆10,686May 16, 2026Updated 2 months ago
WolfLink-DevTeam / Sharine
View on GitHub
Competition work of qiniu 1024 code marathon
☆248Mar 28, 2024Updated 2 years ago