☆53Feb 19, 2026Updated 5 months ago
Alternatives and similar repositories for kani-tts-2-pretrain
Users that are interested in kani-tts-2-pretrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyPi package for KaniTTS-2 model☆63Jun 24, 2026Updated last month
- ☆24Aug 1, 2026Updated last week
- ☆27Nov 3, 2025Updated 9 months ago
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 4 months ago
- ☆46Oct 28, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆19Oct 13, 2025Updated 10 months ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆35Apr 22, 2026Updated 3 months ago
- ☆43Oct 9, 2025Updated 10 months ago
- ☆18Dec 1, 2025Updated 8 months ago
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆25Apr 1, 2025Updated last year
- A MCP stdio toolpack for local LLMs☆34Apr 6, 2026Updated 4 months ago
- Code for the blog "Neural audio codecs: how to get audio into LLMs"☆175Oct 20, 2025Updated 9 months ago
- A local-first web search agent☆29Jun 20, 2026Updated last month
- A collection of all our phonemeizers for dataset construction and inference☆30Feb 21, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Lightweight API Specification for Intelligent Systems☆15Feb 16, 2026Updated 5 months ago
- Drax: Speech Recognition with Discrete Flow Matching☆75Oct 15, 2025Updated 9 months ago
- Speaker embedding for anime speech domain based on ECAPA_TDNN☆21Jun 22, 2025Updated last year
- 3.34× faster inference on Apple Silicon — native MLX port of DFlash speculative decoding☆19Apr 11, 2026Updated 4 months ago
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- Agentic BYOK Browser-Based Website Builder☆52Updated this week
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated 11 months ago
- DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors☆39Feb 11, 2025Updated last year
- DocFinder is a local-first indexing and searching documents using semantic embeddings stored in SQLite. Everything runs on your machine, …☆27Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Prometheus metrics exporter for NVIDIA DGX Spark clusters.☆19Feb 16, 2026Updated 5 months ago
- ☆459Nov 2, 2025Updated 9 months ago
- Local banking voice assistant focused on banking☆70Apr 10, 2026Updated 4 months ago
- Fast audio super resolution from 16khz to 48khz.☆217Jan 3, 2026Updated 7 months ago
- A Chrome extension that enables virtual fashion try-on and model swap using FASHN AI. Hover over fashion images on any website to: (1) tr…☆22Aug 14, 2025Updated 11 months ago
- Whisper Speaker Identification (WSI), a cutting-edge model for multilingual speaker identification.☆27Jun 29, 2026Updated last month
- In this repository I have made FastApi Crud endpoint using pydantic and SQLite☆10Feb 13, 2024Updated 2 years ago
- Is strawberry a fruit or a vegetable?☆54Jun 10, 2026Updated 2 months ago
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆24Jan 24, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 4-bit Shampoo for Memory-Efficient Network Training (NeurIPS 2024)☆13Feb 13, 2025Updated last year
- A Voice Assistant in your Browser.☆23May 12, 2025Updated last year
- ProsodyLM: Uncovering the Emerging Prosody Processing Capabilities in Speech Language Models☆46Nov 18, 2025Updated 8 months ago
- finetune llm part for spark-tts model☆126Mar 25, 2025Updated last year
- A robust Python toolkit for converting video/audio content into accurate, multilingual subtitles using WhisperX for transcription and Goo…☆27Dec 2, 2025Updated 8 months ago
- My guide to create an italian TTS with Coqui☆14Feb 2, 2022Updated 4 years ago
- Jax/Flax implementation of Denoising Diffusion Implicit Models☆20Jul 18, 2022Updated 4 years ago