A variable-frame-rate 16 kHz speech codec based on FocalCodec
☆20Feb 11, 2026Updated 5 months ago
Alternatives and similar repositories for dycast
Users that are interested in dycast are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A low-bitrate single-codebook 16 / 24 kHz speech codec based on focal modulation☆173Nov 30, 2025Updated 8 months ago
- A collections of audio codecs with a standardized API☆43Apr 15, 2026Updated 3 months ago
- [ICLR2026] FlexiCodec: A Dynamic Neural Audio Codec for Low Frame Rates☆51Jul 1, 2026Updated last month
- Codebase for 'ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining'☆24Jun 20, 2026Updated last month
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- CodecHub: A Unified Library for Codec Models☆25Dec 24, 2025Updated 7 months ago
- Official code for paper:"Speaking Clearly: A Simplified Whisper-Based Codec for Low-Bitrate Speech Coding"☆37Jan 28, 2026Updated 6 months ago
- Voxtral Codec : Combining Semantic VQ and Acoustic FSQ for Ultra-Low Bitrate Speech Generation (Voxtral TTS Backbone)☆17Mar 27, 2026Updated 4 months ago
- A method that directly addresses the modality gap by aligning speech token with the corresponding text transcription during the tokenizat…☆119Sep 3, 2025Updated 11 months ago
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 6 months ago
- ☆28Feb 14, 2026Updated 5 months ago
- Extract a target speaker’s clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.☆21Jun 14, 2026Updated last month
- ICASSP 2024 - Generative De-Quantization for Neural Speech Codec via Latent Diffusion.☆56Updated this week
- ☆14Aug 1, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Bayesian deep learning for remaining useful life estimation via Stein variational gradient descent☆30Feb 5, 2024Updated 2 years ago
- Kanade is a single-layer disentangled speech tokenizer that extracts compact tokens suitable for both generative and discriminative model…☆109Jul 18, 2026Updated 2 weeks ago
- Neural Speech Codec☆26Jan 25, 2021Updated 5 years ago
- Official implementation of "Wave-Trainer-Fit: Neural Vocoder with Trainable Prior and Fixed-Point Iteration towards High-Quality Speech G…☆16Feb 6, 2026Updated 5 months ago
- Detect individual instruments activity in an audio file. 🎤🎹🎸🥁☆17Jun 29, 2021Updated 5 years ago
- The source code for Input-Adaptive Spectral Feature Compression by Sequence Modeling for Source Separation published in IEEE TASLPRO.☆19Jun 3, 2026Updated 2 months ago
- For IEEE ASRU(2025)☆15Jun 21, 2025Updated last year
- ☆34Mar 29, 2025Updated last year
- ☆22Mar 1, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Open-weights voice acting pipeline combining zero-shot voice cloning with natural-language direction. Provide a reference voice (or gener…☆17May 25, 2026Updated 2 months ago
- FastWave is a lightweight diffusion model for general audio super-resolution (any -> 48 kHz). SOTA quality reconstruction metrics with ju…☆18May 16, 2026Updated 2 months ago
- Variations of L1 SNR Loss function for training audio source separation machine learning models☆45May 1, 2026Updated 3 months ago
- A family of efficient speech models for multilingual phone recognition☆70Jul 18, 2026Updated 2 weeks ago
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one☆26Aug 5, 2024Updated last year
- libvits-ncnn is an ncnn implementation of the VITS library that enables cross-platform GPU-accelerated speech synthesis.🎙️💻☆62May 6, 2023Updated 3 years ago
- A neural speech codec based on discrete WavLM representations☆26Aug 28, 2024Updated last year
- Multi-band Frequency Reconstruction for Neural Psychoacoustic Coding☆19May 5, 2025Updated last year
- ☆14Oct 3, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of the paper "BigCodec: Pushing the Limits of Low-Bitrate Neural Speech Codec"☆218Sep 19, 2024Updated last year
- TensorFlow,DCGAN,VAE,LSTM,CNN,Acoustic Scene Classification☆11Jun 5, 2019Updated 7 years ago
- fd-sds☆21Apr 8, 2026Updated 3 months ago
- PyTorch code implementation of EfficientSpeech - to be presented at ICASSP2023.☆182Mar 18, 2024Updated 2 years ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆118Nov 24, 2025Updated 8 months ago
- ☆19Jul 23, 2025Updated last year
- High fidelity neural audio codec for TTS models☆36Dec 22, 2025Updated 7 months ago