A modern desktop application built with Tauri 2.0 for creating professional audiobooks using advanced text-to-speech and voice cloning technology (XTTS, Chatterbox, VibeVoice). Features drag & drop organization, multi-language support (17+ languages), smart text segmentation with NLP, and export to MP3/M4A/WAV formats.
β88Jan 4, 2026Updated 6 months ago
Alternatives and similar repositories for AudioBook-Maker
Users that are interested in AudioBook-Maker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π³ MCTS-inspired parallel beam search for conversation optimization. Explore multiple dialogue strategies simultaneously, stress-test aβ¦β36Jan 18, 2026Updated 6 months ago
- Audiobook Creator is an app that converts books (EPUB, PDF, TXT etc.) into fully voiced audiobooks with intelligent character voice attriβ¦β517Nov 17, 2025Updated 8 months ago
- β27Sep 25, 2024Updated last year
- Generate audiobooks from pdf or epub using Next-gen AI Chatterbox-tts from Resemble-AIβ20Apr 26, 2026Updated 2 months ago
- β18Jul 1, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ComfyUI Discord Bot - Share workflows with your friends through Discord!β15Feb 25, 2026Updated 5 months ago
- β116Dec 17, 2025Updated 7 months ago
- Audiobook creation app supporting too many TTS models (Qwen3-TTS, OmniVoice, VibeVoice, etc), focused on high-quality output. Plus audio-β¦β174Updated this week
- A realtime speech to text diarization system to gather and interleave speech from multiple speaker audio.β54Jan 29, 2026Updated 5 months ago
- FastAPI wrapper around original Vibevoice 1.5B and 7B models, with support for AWQ4 quantβ33Jun 22, 2026Updated last month
- klmbr - a prompt pre-processing technique to break through the barrier of entropy while generating text with LLMsβ90Sep 22, 2024Updated last year
- Open API and Wyoming wrapper around Chatterboxβ28Jan 2, 2026Updated 6 months ago
- SoTA open-source TTS for Audiobook and Podcast Generationβ206Jun 19, 2025Updated last year
- A proxy that hosts multiple single-model runners such as LLama.cpp and vLLMβ12May 30, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The most feature-complete local AI workstation. Multi-GPU inference, integrated Stable Diffusion + ADetailer, voice cloning, research-graβ¦β64Updated this week
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with β¦β20Feb 15, 2026Updated 5 months ago
- Glyphs, acting as collaboratively defined symbols linking related concepts, add a layer of multidimensional semantic richness to user-AI β¦β57Feb 10, 2025Updated last year
- Local-first desktop AI workbench for roleplay, multi-character chat, long-form writing, RAG, MCP tools, plugins, and local models.β105Updated this week
- Extract a target speakerβs clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.β21Jun 14, 2026Updated last month
- A terminal interface for your AI terminal assistant.β18May 30, 2025Updated last year
- Cognito: Supercharge your Chrome browser with AI. Guide, query, and control everything using natural language.β57Jun 19, 2026Updated last month
- A text-grid web renderer for AI agents β see the web without screenshotsβ102Mar 10, 2026Updated 4 months ago
- Vortex is a self-hosted RAG (Retrieval-Augmented Generation) application that lets you chat with your documents using any LLM provider. Uβ¦β17Jul 11, 2026Updated 2 weeks ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- β23Sep 20, 2025Updated 10 months ago
- Hyperaudio Converter - converts from JSON/SRT to HTML Based Interactive Transcriptβ14Dec 16, 2020Updated 5 years ago
- YATSEE - Yet Another Tool for Speech Extraction & Enrichmentβ31Updated this week
- A professional-grade interface for Qwen3-TTS, designed to unlock the model's full potential with fine-grained control and intuitive workfβ¦β285Mar 30, 2026Updated 3 months ago
- Echo-TTS OpenAI Compatible Speech Endpoint w/ Streamingβ29Apr 5, 2026Updated 3 months ago
- Open source tool for transcirption and subtitling, alternative to happyscribe.β36Feb 12, 2025Updated last year
- An experimental autonomous research system that conducts comprehensive, multi-hour research sessions and produces book-length reports witβ¦β47Dec 7, 2025Updated 7 months ago
- Chatbot-to-speech using Orpheus TTS model. Interactive console app.β21May 1, 2025Updated last year
- β224May 7, 2025Updated last year
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- β35Jul 4, 2026Updated 3 weeks ago
- Speaker diarization for Python β "who spoke when?" CPU-only, no API keys, Apache 2.0. ~10.8% DER on VoxConverse, 8x faster than real-timeβ¦β97May 6, 2026Updated 2 months ago
- Your personal and private AIβ54Apr 3, 2025Updated last year
- π₯ π₯ Alternative to Ollama π₯ π₯ multi-model <1ms LLM switchingβ38Updated this week
- Testbench for llama.cpp llama-serverβ15Aug 20, 2025Updated 11 months ago
- Spec-driven iterative development companion CLI for OpenCode.β18Jun 7, 2026Updated last month
- Frontend for retrogaming PC and cabinetsβ13Oct 28, 2020Updated 5 years ago