☆53Feb 19, 2026Updated 6 months ago
Alternatives and similar repositories for kani-tts-2-pretrain
Users that are interested in kani-tts-2-pretrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆24Aug 1, 2026Updated last month
- ☆26Nov 3, 2025Updated 9 months ago
- Multi-agent orchestration framework for AI applications - build, deploy, and manage AI agents across the full lifecycle with Forge, Conve…☆33Mar 28, 2026Updated 5 months ago
- ☆46Oct 28, 2025Updated 10 months ago
- This repository contains all the code necessary for running the multilingual distilwhisper from Ferraz et al. 2024 IEEE ICASSP paper.☆34Apr 22, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆20Jan 3, 2026Updated 7 months ago
- ☆44Oct 9, 2025Updated 10 months ago
- Inspect LLM's logprobs and perplexity over a piece of text, or compare two LLMs (like a git diff)☆18Aug 3, 2026Updated 3 weeks ago
- ☆18Dec 1, 2025Updated 9 months ago
- Run Orpheus 3B Locally with Gradio UI, Standalone App☆25Apr 1, 2025Updated last year
- A MCP stdio toolpack for local LLMs☆34Apr 6, 2026Updated 4 months ago
- Code for the blog "Neural audio codecs: how to get audio into LLMs"☆175Oct 20, 2025Updated 10 months ago
- A local-first web search agent☆30Jun 20, 2026Updated 2 months ago
- A collection of all our phonemeizers for dataset construction and inference☆32Feb 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Lightweight API Specification for Intelligent Systems☆16Feb 16, 2026Updated 6 months ago
- Arabic Grapheme-to-Phoneme (G2P) Conversion☆15Mar 15, 2025Updated last year
- Drax: Speech Recognition with Discrete Flow Matching☆75Oct 15, 2025Updated 10 months ago
- Adding a multi-text multi-speaker script (diffe) that is based on a script from asiff00 on issue 61 for Sesame: A Conversational Speech G…☆26Mar 28, 2025Updated last year
- A TTS Trained on Universal Audio.☆42Jun 6, 2025Updated last year
- Agentic BYOK Browser-Based Website Builder☆56Updated this week
- Meanflow and multilingual for F5-TTS model☆16Aug 23, 2025Updated last year
- DiTTo-TTS: Diffusion Transformers for Scalable Text-to-Speech without Domain-Specific Factors☆39Feb 11, 2025Updated last year
- A Prometheus metrics exporter for NVIDIA DGX Spark clusters.☆22Feb 16, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆42Jul 13, 2026Updated last month
- ☆460Nov 2, 2025Updated 10 months ago
- A rework of the gradio WebUI for the open-source unified multimodal model by ByteDance☆21Jun 3, 2025Updated last year
- Fast audio super resolution from 16khz to 48khz.☆218Jan 3, 2026Updated 7 months ago
- Whisper Speaker Identification (WSI), a cutting-edge model for multilingual speaker identification.☆27Jun 29, 2026Updated 2 months ago
- Is strawberry a fruit or a vegetable?☆54Jun 10, 2026Updated 2 months ago
- Pashto Natural Language Processing Toolkit☆12May 21, 2025Updated last year
- Protocol for Augmented Memory of Project Artifacts (MCP compatible) - extended☆24Jan 24, 2026Updated 7 months ago
- ProsodyLM: Uncovering the Emerging Prosody Processing Capabilities in Speech Language Models☆46Nov 18, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- finetune llm part for spark-tts model☆126Mar 25, 2025Updated last year
- A robust Python toolkit for converting video/audio content into accurate, multilingual subtitles using WhisperX for transcription and Goo…☆28Aug 23, 2026Updated last week
- My guide to create an italian TTS with Coqui☆14Feb 2, 2022Updated 4 years ago
- Code for Latent Speech-Text Transformer (LST)☆35Mar 12, 2026Updated 5 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆257Jan 13, 2026Updated 7 months ago
- Soprano: Instant, Ultra-Realistic Text-to-Speech☆1,547Jan 15, 2026Updated 7 months ago
- ☆31May 15, 2024Updated 2 years ago