Automated speech dataset creator
β224Jun 12, 2025Updated last year
Alternatives and similar repositories for Voice_Extractor
Users that are interested in Voice_Extractor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ποΈ Automatically transcribe audio/video into high-quality, speaker-specific Text-To-Speech datasetsβ142Aug 10, 2025Updated 11 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.β32May 1, 2025Updated last year
- Open source tool for transcirption and subtitling, alternative to happyscribe.β36Feb 12, 2025Updated last year
- Unlimited text-to-speech in the Browser using Kokoro-JS, 100% local, 100% open sourceβ346Jun 12, 2025Updated last year
- A web application that converts speech to speech 100% privateβ86Jun 3, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β27Jun 11, 2025Updated last year
- Perspt: **Per**sonal **S**pectrum **P**ertaining **T**houghts β the human lens through which we explore the enigma of AI and its implicatβ¦β33Jul 13, 2026Updated last week
- Audiobook Creator is an app that converts books (EPUB, PDF, TXT etc.) into fully voiced audiobooks with intelligent character voice attriβ¦β516Nov 17, 2025Updated 8 months ago
- Create Unmute voice embeddingsβ26Nov 15, 2025Updated 8 months ago
- β13Mar 10, 2025Updated last year
- β15Mar 18, 2026Updated 4 months ago
- β19Aug 19, 2025Updated 11 months ago
- Realtime tts reading of large textfiles by your favourite voice. +Translation via LLM (Python script)β51Oct 18, 2024Updated last year
- A random walk voice style cloning application for Kokoro text to speechβ264Apr 6, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Evaluating practical performance of local multi-turn conversational LLMs.β19Aug 1, 2025Updated 11 months ago
- A functioning Sesame CSM project with a desktop GUI - Real-time factor: 0.6x with 4070 Ti Super - Requires only 8GB VRAMβ81May 19, 2025Updated last year
- β54May 28, 2025Updated last year
- Analyze Reddit postsβ32Jun 5, 2026Updated last month
- Efforts toward giving Qwen 3 Coder 30B A3B proper agentic tool calling capabilities at or near 100% reliability.β63Aug 10, 2025Updated 11 months ago
- A GTK4-based text-to-speech and AI assistant app in Rust, featuring PDF reading and LLM chat powered by Kokoro TTSβ19Jul 10, 2026Updated last week
- Unofficial WIP LoRa Finetuning repository for VibeVoiceβ367Sep 24, 2025Updated 9 months ago
- β16Dec 16, 2024Updated last year
- Deploy Apollo HF space locallyβ40Dec 16, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Streaming and Fine-tuning for Chatterbox TTSβ291Jun 15, 2025Updated last year
- An MCP-enabled Qwen3 0.6B demo with adjustable thinking budget, all in your browser!β28Jun 2, 2025Updated last year
- Controllable Language Model Interactions in TypeScriptβ10May 17, 2024Updated 2 years ago
- My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into oneβ26Aug 5, 2024Updated last year
- Extract2MD is a powerful and versatile AI-enabled client-side JavaScript library for extracting text from PDF files and converting it intβ¦β109Apr 7, 2026Updated 3 months ago
- Realtime demo, Streaming and Finetuning code for CSMβ455Sep 17, 2025Updated 10 months ago
- FlexAudioPrint is a Python-based app for transcribing audio to text using OpenAI's Whisper model. It offers a Gradio web interface and a β¦β10Apr 22, 2026Updated 3 months ago
- SoTA open-source TTSβ165Dec 16, 2025Updated 7 months ago
- Glyphs, acting as collaboratively defined symbols linking related concepts, add a layer of multidimensional semantic richness to user-AI β¦β57Feb 10, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- AI-powered video frame extraction tool that automatically identifies and extracts high-quality frames containing people, with intelligentβ¦β167Mar 25, 2026Updated 3 months ago
- A utility that uses Whisper to transcribe videos and various translation APIs to translate the transcribed text and save them as SRT (subβ¦β76Aug 30, 2024Updated last year
- Hector RAG is a modular RAG framework built on PostgreSQL, offering advanced retrieval methods and fusion techniques for AI-driven applicβ¦β60Feb 24, 2025Updated last year
- The most feature-complete local AI workstation. Multi-GPU inference, integrated Stable Diffusion + ADetailer, voice cloning, research-graβ¦β63Updated this week
- Create text chunks which end at natural stopping points without using a tokenizerβ26Nov 26, 2025Updated 7 months ago
- Your personal and private AIβ54Apr 3, 2025Updated last year
- Speech-to-speech AI assistant with natural conversation flow, mid-speech interruption, vision capabilities and AI-initiated follow-ups. Fβ¦β308Apr 14, 2025Updated last year