Voca - Your local voice cloning assistant. Powered by VoxCPM
☆45Aug 5, 2026Updated last month
Alternatives and similar repositories for Voca
Users that are interested in Voca are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- Efficient voice activity detection algorithm using long-term spectral flatness measurement☆15Feb 21, 2017Updated 9 years ago
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 3 years ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆18Jun 27, 2026Updated 3 months ago
- PyTorch implementation of Tacotron and Tacotron2☆34Jul 19, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Demo for AudioSAE paper☆16Apr 26, 2026Updated 5 months ago
- ☆22Apr 9, 2026Updated 5 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆90Sep 15, 2026Updated 2 weeks ago
- PyTorch implementation of USR 2.0 (ICLR 2026)☆16Sep 9, 2026Updated 3 weeks ago
- ☆24Jul 16, 2025Updated last year
- Official repository for the paper "Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enh…☆23May 5, 2026Updated 4 months ago
- This challenge focuses on evaluating speech recognition and semantic understanding capabilities of AI glasses in complex real-world envir…☆18Aug 22, 2026Updated last month
- Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡☆11Jan 23, 2025Updated last year
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Real-Time Streamable Generative Speech Restoration with Flow Matching☆66Jun 5, 2026Updated 3 months ago
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- ☆32Jan 9, 2024Updated 2 years ago
- A pytorch implementation of the paper "ANSD-MA-MSE: Adaptive Neural Speaker Diarization Using Memory-Aware Multi-Speaker Embedding"☆62Sep 19, 2024Updated 2 years ago
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆23Jul 30, 2026Updated 2 months ago
- ☆27Aug 29, 2025Updated last year
- ☆27May 26, 2026Updated 4 months ago
- ☆315Sep 24, 2026Updated last week
- ☆27Mar 31, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 🤫A Lightweight One-Shot Whisper to Normal Voice Conversion Model Using Distillation of Self-Supervised Features☆25Dec 10, 2025Updated 9 months ago
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆33Jul 20, 2026Updated 2 months ago
- [ICLR 2026] StableToken: A state-of-the-art noise-robust semantic speech tokenizer featuring Voting-LFQ for resilient SpeechLLMs.☆34Feb 27, 2026Updated 7 months ago
- ☆29Jul 19, 2026Updated 2 months ago
- ☆24Jul 10, 2025Updated last year
- Your faithful, impartial partner for audio evaluation — know yourself, know your rivals. 真实评测,知己知彼。A unified benchmark framework for ASR/…☆326Sep 16, 2026Updated 2 weeks ago
- ☆12Oct 21, 2019Updated 6 years ago
- ICASSP2022 TTS&VC Summary☆13Jun 9, 2022Updated 4 years ago
- A repository for code used to produce the results the ICASSP 2024 paper: "SELF-SUPERVISED PRETRAINING FOR ROBUST PERSONALIZED VOICE ACTIV…☆27Nov 25, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆35Oct 23, 2025Updated 11 months ago
- ☆21Feb 27, 2024Updated 2 years ago
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago
- ☆16May 25, 2022Updated 4 years ago
- ☆25Mar 12, 2022Updated 4 years ago
- ☆47Aug 31, 2024Updated 2 years ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆47Jul 24, 2026Updated 2 months ago