Voca - Your local voice cloning assistant. Powered by VoxCPM
☆45Aug 5, 2026Updated last month
Alternatives and similar repositories for Voca
Users that are interested in Voca are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 这是一个面向家庭场景的离线监控视频分析系统。项目以 NAS/目录中的摄像头录像为输入,自动完成视频扫描、会话合并、AI 事件识别、家庭日报生成,并提供 Web 管理后台、自然语言问答、Webhook 与 MCP 对外能力。☆22Updated this week
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- Efficient voice activity detection algorithm using long-term spectral flatness measurement☆15Feb 21, 2017Updated 9 years ago
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 3 years ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆17Jun 27, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch implementation of Tacotron and Tacotron2☆34Jul 19, 2022Updated 4 years ago
- Demo for AudioSAE paper☆16Apr 26, 2026Updated 4 months ago
- On-device VAD / streaming STT / TTS / diarization in C++17 (ONNX + LiteRT) with a voice-agent pipeline. Linux, Windows, Android.☆83Sep 2, 2026Updated last week
- ☆21Apr 9, 2026Updated 5 months ago
- PyTorch implementation of USR 2.0 (ICLR 2026)☆16Updated this week
- ☆24Jul 16, 2025Updated last year
- Official repository for the paper "Exploring Resolution-Wise Shared Attention in Hybrid Mamba-U-Nets for Improved Cross-Corpus Speech Enh…☆22May 5, 2026Updated 4 months ago
- This challenge focuses on evaluating speech recognition and semantic understanding capabilities of AI glasses in complex real-world envir…☆18Aug 22, 2026Updated 3 weeks ago
- Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡☆11Jan 23, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 7 months ago
- Real-Time Streamable Generative Speech Restoration with Flow Matching☆63Jun 5, 2026Updated 3 months ago
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- 🌕营销号生成器☆11Apr 20, 2020Updated 6 years ago
- ☆32Jan 9, 2024Updated 2 years ago
- A pytorch implementation of the paper "ANSD-MA-MSE: Adaptive Neural Speaker Diarization Using Memory-Aware Multi-Speaker Embedding"☆62Sep 19, 2024Updated last year
- PyTorch-based room impulse response (RIR) simulation toolkit with dynamic scenes, GPU acceleration.☆23Jul 30, 2026Updated last month
- ☆27Aug 29, 2025Updated last year
- ☆27May 26, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆60Updated this week
- ☆303Sep 2, 2026Updated last week
- ☆27Mar 31, 2026Updated 5 months ago
- 🤫A Lightweight One-Shot Whisper to Normal Voice Conversion Model Using Distillation of Self-Supervised Features☆25Dec 10, 2025Updated 9 months ago
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆31Jul 20, 2026Updated last month
- [ICLR 2026] StableToken: A state-of-the-art noise-robust semantic speech tokenizer featuring Voting-LFQ for resilient SpeechLLMs.☆34Feb 27, 2026Updated 6 months ago
- LLM inference in C/C++☆22Updated this week
- ☆28Jul 19, 2026Updated last month
- ☆24Jul 10, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Your faithful, impartial partner for audio evaluation — know yourself, know your rivals. 真实评测,知己知彼。A unified benchmark framework for ASR/…☆324Aug 25, 2026Updated 2 weeks ago
- ☆12Oct 21, 2019Updated 6 years ago
- A repository for code used to produce the results the ICASSP 2024 paper: "SELF-SUPERVISED PRETRAINING FOR ROBUST PERSONALIZED VOICE ACTIV…☆27Nov 25, 2024Updated last year
- ICASSP2022 TTS&VC Summary☆13Jun 9, 2022Updated 4 years ago
- ☆35Oct 23, 2025Updated 10 months ago
- ☆21Feb 27, 2024Updated 2 years ago
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago