On-device streaming text-to-speech engine powered by deep learning
☆143Aug 14, 2026Updated this week
Alternatives and similar repositories for orca
Users that are interested in orca are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 内容审核及速率限制服务☆26May 18, 2025Updated last year
- On-device LLM Inference Powered by X-Bit Quantization☆316Updated this week
- On-device speaker recognition engine powered by deep learning☆57Updated this week
- On-device Speech-to-Index engine powered by deep learning☆36Apr 16, 2025Updated last year
- On-device speaker diarization powered by deep learning☆77Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- benchmark for Speech-to-Intent engines☆18Jul 30, 2026Updated 2 weeks ago
- Voice activity engine benchmark framework☆23Jan 14, 2026Updated 7 months ago
- On-device streaming speech-to-text engine powered by deep learning☆670Updated this week
- On-device noise suppression powered by deep learning☆92Updated this week
- Picovoice Browser Extension☆17Jun 24, 2026Updated last month
- On-device voice activity detection (VAD) powered by deep learning☆269Updated this week
- On-device speech-to-text engine powered by deep learning☆483Updated this week
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- LLm Collaboration☆12Aug 23, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Text-to-Speech Benchmark☆29Apr 2, 2026Updated 4 months ago
- This repository is maintained by the Speech Team at Alibaba’s Tongyi Lab, serving as an open-source platform for our cutting-edge researc…☆47Aug 3, 2026Updated last week
- On-device AI recipes for building private, real-time applications with production-grade voice, language, and vision SDKs. Ready-to-use, o…☆114Updated this week
- Diffusion-based Speech Enhancement: Demonstration of Performance and Generalization☆14Dec 21, 2024Updated last year
- ☆35Jun 3, 2026Updated 2 months ago
- Automated end-to-end data preprocessing, model training, and evaluation pipeline☆17Jun 3, 2026Updated 2 months ago
- Parse lightning network payment requests (invoices) in Kotlin.☆22May 27, 2026Updated 2 months ago
- ☆21Jun 4, 2026Updated 2 months ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speake…☆57Aug 7, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- On-device Speech-to-Intent engine powered by deep learning☆706Updated this week
- Tapestry CSRF Protection☆11Sep 23, 2025Updated 10 months ago
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- The code for AAAI 2025 “Large Language Models Are Read/Write Policy-Makers for Simultaneous Generation”☆15Jan 3, 2025Updated last year
- ☆203Sep 24, 2024Updated last year
- ☆29Jun 25, 2026Updated last month
- Terminal Voice Assistant is a powerful and flexible tool designed to help users interact with their terminal using natural language comma…☆18Jun 9, 2024Updated 2 years ago
- ☆81Aug 11, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆18Aug 28, 2025Updated 11 months ago
- Please visit https://thuhcsi.github.io/SnakeGAN/☆37Apr 25, 2023Updated 3 years ago
- Emotion Rendering for Conversational Speech Synthesis with Heterogeneous Graph-Based Context Modeling (Accepted by AAAI'2024)☆59Jun 20, 2024Updated 2 years ago
- Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions☆270Jan 13, 2025Updated last year
- Speaker diarization benchmark framework☆45Jul 17, 2026Updated 3 weeks ago
- Code for ACL 2024 main conference paper "Can We Achieve High-quality Direct Speech-to-Speech Translation Without Parallel Speech Data?".☆27Jul 2, 2024Updated 2 years ago
- ☆25Jan 16, 2025Updated last year