A library for making PyTorch models streamable
☆67Jan 23, 2026Updated 7 months ago
Alternatives and similar repositories for TorchStream
Users that are interested in TorchStream are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenFLAM: Framewise Language Audio Model☆115Jun 4, 2026Updated 3 months ago
- Code and data recipes for the paper: Optimal Condition Training for Target Source Separation by Efthymios Tzinis, Gordon Wichern, Paris S…☆14Feb 15, 2023Updated 3 years ago
- ☆30Jul 3, 2026Updated 2 months ago
- poorman's ar-dit tts☆45Dec 31, 2025Updated 8 months ago
- Properly handle position-dependent phones in a subword lexicon FST☆31Oct 26, 2020Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Text to Speech Synthesis based on controllable latent representation☆14Aug 30, 2019Updated 7 years ago
- ☆23Sep 14, 2021Updated 4 years ago
- Unicode Standard tokenization routines and orthography profile segmentation☆41Mar 7, 2026Updated 5 months ago
- Unified automatic quality assessment for speech, music, and sound.☆756Jun 5, 2025Updated last year
- Text Normalization utilities for normalizing text for TTS☆27Mar 4, 2026Updated 6 months ago
- Fast audio super resolution from 16khz to 48khz.☆218Jan 3, 2026Updated 8 months ago
- Encode and decode audio samples to/from continuous and discrete compressed representations!☆121Nov 25, 2025Updated 9 months ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- ☆52Apr 20, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Inference for the STFT-VAE continuous audio codec (24kHz, 3.125Hz latent)☆47Jul 12, 2026Updated last month
- The open source code for SimpleSpeech series☆146Oct 8, 2024Updated last year
- Codebase for the paper 'EncodecMAE: Leveraging neural codecs for universal audio representation learning'☆101Jul 24, 2024Updated 2 years ago
- ☆22Aug 24, 2026Updated last week
- A simple library for Fréchet Audio Distance (FAD) calculation☆266Aug 22, 2025Updated last year
- Generation of musical phrases that receive maximum score according to configurable evaluational criteria.☆12Oct 17, 2023Updated 2 years ago
- Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis☆1,158Updated this week
- Code for the paper: Unified Gradient Reweighting for Model Biasing with Applications to Source Separation☆14Nov 16, 2020Updated 5 years ago
- Simple and lightweight Zero-shot Text-to-Speech (TTS) synthesis model☆37Apr 29, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A notebooks based (soft) intro to modern TTS☆18Jun 8, 2025Updated last year
- A demo using Soundfonts to play music within JUCE.☆12Nov 1, 2023Updated 2 years ago
- Libzita-resampler is a C++ library for resampling audio signals. It is designed to be used within a real-time processing context, to be f…☆35Mar 27, 2021Updated 5 years ago
- Segment an audio file and obtain utterance alignments. (Python package)☆348May 15, 2024Updated 2 years ago
- KVAE-Audio: a continuous full-band audio waveform autoencoder☆122Aug 10, 2026Updated 3 weeks ago
- Native full-duplex speech dialogue inference for BayLing-Duplex.☆83Jun 22, 2026Updated 2 months ago
- Training code and dataset cleasing with Sidon☆187Apr 24, 2026Updated 4 months ago
- The TTSDS benchmark evaluates synthetic speech quality by considering prosody, speaker identity, and intelligibility, comparing these fac…☆101Jul 7, 2026Updated last month
- Ultra-Sortformer for Scalable Speaker Diarization☆28Apr 9, 2026Updated 4 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆87Apr 2, 2024Updated 2 years ago
- Unofficial fairseq-free PyTorch implementation of UTMOS (v1, 2022), matching the original system.☆35Jun 6, 2026Updated 2 months ago
- Reverse Engineering of Supervised Semantic Speech Tokenizer (S3Tokenizer) proposed in CosyVoice☆530Dec 22, 2025Updated 8 months ago
- An open-source application for 3D acoustic simulation☆84Aug 23, 2026Updated last week
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- a python library for midi to wav, generation, visualization, which is design for machine learning☆11Mar 25, 2019Updated 7 years ago
- A desktop app to fine-tune, generate and perform live music with diffusion models. Runs on MacOS, Windows and Linux.☆60Aug 20, 2026Updated 2 weeks ago