☆111Sep 24, 2025Updated 9 months ago
Alternatives and similar repositories for IndicF5
Users that are interested in IndicF5 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple lightweight library for text normalization for Indian Languages☆18Sep 30, 2025Updated 9 months ago
- Indic TTS for Indian Languages: This is a project on developing text-to-speech (TTS) synthesis systems for Indian languages, improving qu…☆57Feb 5, 2026Updated 5 months ago
- ☆79Jun 20, 2025Updated last year
- Text-to-Speech for languages of India☆378Nov 8, 2024Updated last year
- ☆19Feb 22, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Make any person bald!! Component of the paper: Learning to regulate 3D head shape by removing occluding hair from in-the-wild images.☆12Jun 6, 2022Updated 4 years ago
- A notebooks based (soft) intro to modern TTS☆18Jun 8, 2025Updated last year
- ☆22Apr 6, 2025Updated last year
- Dataset release for Emotional TTS in Indian Accent☆41Mar 25, 2026Updated 3 months ago
- Vistaar: Diverse Benchmarks and Training Sets for Indian Language ASR☆87Jun 8, 2025Updated last year
- Echo-TTS OpenAI Compatible Speech Endpoint w/ Streaming☆29Apr 5, 2026Updated 3 months ago
- Train and finutune text-to-speech models for Bengali and many other languages!☆18Apr 2, 2025Updated last year
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion☆43Sep 9, 2025Updated 10 months ago
- Text Normalization utilities for normalizing text for TTS☆26Mar 4, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- This repository Aims to quantize a hindi speech recognition model from https://huggingface.co/ai4bharat/indicconformer_stt_hi_hybrid_ctc_…☆15Jun 19, 2025Updated last year
- Zero-shot voice cloning text-to-speech (TTS) with explicit emotion class conditioning built on F5-TTS☆40Mar 3, 2026Updated 4 months ago
- ☆16May 14, 2025Updated last year
- [ICML 2024] Official Repository for the paper "Transformers Get Stable: An End-to-End Signal Propagation Theory for Language Models"☆11Jul 19, 2024Updated 2 years ago
- Fork of "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"☆17Nov 27, 2024Updated last year
- UMETTS: A Unified Framework for Emotional Text-to-Speech Synthesis with Multimodal Prompts☆41Jun 12, 2025Updated last year
- A TTS model capable of generating ultra-realistic dialogue in one pass.☆16Jun 28, 2025Updated last year
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 3 years ago
- ☆12Oct 24, 2017Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- VoXtream is a Full-Stream Zero-shot TTS model with Extremely Low Latency and Speaking rate Control☆244May 30, 2026Updated last month
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matching☆1,015Dec 2, 2025Updated 7 months ago
- ☆15Aug 4, 2025Updated 11 months ago
- Using AI based approach to detect illegal parking of vehicles (Cars) from an image. The model will receive an image of parked car through…☆11Jun 2, 2020Updated 6 years ago
- The Gaming Zone is a web application that provides you with a collection of classic retro games, including puzzle games, trivia games, bo…☆10Feb 11, 2020Updated 6 years ago
- ☆21Apr 8, 2026Updated 3 months ago
- Modify-Anything is based on yolov5,yolov8 for video and image detection. Segment-anything,lama_cleaner is applied to segment, modify, era…☆17May 3, 2023Updated 3 years ago
- Repository having the code and models from the paper: data2vec-aqc: Search for the right Teaching Assistant in the Teacher-Student traini…☆13Mar 18, 2024Updated 2 years ago
- Codec for paper: LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis☆360Jun 25, 2026Updated 3 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- a Frontier Japanese Speech Generation net☆65May 15, 2025Updated last year
- ☆24May 6, 2025Updated last year
- Pytorch implementation of CS-Tacotron, a code-switching speech synthesis end-to-end generative TTS model.☆23Mar 14, 2019Updated 7 years ago
- A Massive Multilingual Multi-speaker Speech Corpus for Scaling Indian TTS☆64Dec 11, 2024Updated last year
- Image captioning with a locally stored Large Language Model (LLM)☆15Updated this week
- ☆275Jun 8, 2024Updated 2 years ago
- Generate animated ASCII art SVGs from text. CLI, Rust library, WASM, and web editor.☆16Apr 11, 2026Updated 3 months ago