☆53Feb 19, 2026Updated 5 months ago
Alternatives and similar repositories for kani-tts-2-pretrain
Users that are interested in kani-tts-2-pretrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyPi package for KaniTTS-2 model☆63Jun 24, 2026Updated 3 weeks ago
- ☆23Jul 12, 2026Updated last week
- ☆27Nov 3, 2025Updated 8 months ago
- LLM training & inference in python/C++ with web UI☆37Updated this week
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆46Oct 28, 2025Updated 8 months ago
- ☆19Oct 13, 2025Updated 9 months ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆34Apr 22, 2026Updated 2 months ago
- ☆43Oct 9, 2025Updated 9 months ago
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆20Mar 23, 2026Updated 3 months ago
- ☆18Dec 1, 2025Updated 7 months ago
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆24Apr 1, 2025Updated last year
- A MCP stdio toolpack for local LLMs☆33Apr 6, 2026Updated 3 months ago
- Code for the blog "Neural audio codecs: how to get audio into LLMs"☆173Oct 20, 2025Updated 9 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆16Mar 15, 2025Updated last year
- Drax: Speech Recognition with Discrete Flow Matching☆75Oct 15, 2025Updated 9 months ago
- Speaker embedding for anime speech domain based on ECAPA_TDNN☆21Jun 22, 2025Updated last year
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- A TTS Trained on Universal Audio.☆41Jun 6, 2025Updated last year
- Agentic BYOK Browser-Based Website Builder☆51Updated this week
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated 10 months ago
- DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors☆39Feb 11, 2025Updated last year
- DocFinder is a local-first indexing and searching documents using semantic embeddings stored in SQLite. Everything runs on your machine, …☆26Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A collection of all our phonemeizers for dataset construction and inference☆30Feb 21, 2025Updated last year
- ☆35Jul 13, 2026Updated last week
- ☆461Nov 2, 2025Updated 8 months ago
- A rework of the gradio WebUI for the open-source unified multimodal model by ByteDance☆21Jun 3, 2025Updated last year
- Local banking voice assistant focused on banking☆69Apr 10, 2026Updated 3 months ago
- Fast audio super resolution from 16khz to 48khz.☆215Jan 3, 2026Updated 6 months ago
- A Chrome extension that enables virtual fashion try-on and model swap using FASHN AI. Hover over fashion images on any website to: (1) tr…☆22Aug 14, 2025Updated 11 months ago
- In this repository I have made FastApi Crud endpoint using pydantic and SQLite☆10Feb 13, 2024Updated 2 years ago
- Pashto Natural Language Processing Toolkit☆12May 21, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆24Jan 24, 2026Updated 5 months ago
- ☆19Mar 3, 2025Updated last year
- 4-bit Shampoo for Memory-Efficient Network Training (NeurIPS 2024)☆13Feb 13, 2025Updated last year
- ProsodyLM: Uncovering the Emerging Prosody Processing Capabilities in Speech Language Models☆46Nov 18, 2025Updated 8 months ago
- finetune llm part for spark-tts model☆125Mar 25, 2025Updated last year
- A robust Python toolkit for converting video/audio content into accurate, multilingual subtitles using WhisperX for transcription and Goo…☆28Dec 2, 2025Updated 7 months ago
- My guide to create an italian TTS with Coqui☆14Feb 2, 2022Updated 4 years ago