A modern desktop application built with Tauri 2.0 for creating professional audiobooks using advanced text-to-speech and voice cloning technology (XTTS, Chatterbox, VibeVoice). Features drag & drop organization, multi-language support (17+ languages), smart text segmentation with NLP, and export to MP3/M4A/WAV formats.
β90Jan 4, 2026Updated 8 months ago
Alternatives and similar repositories for AudioBook-Maker
Users that are interested in AudioBook-Maker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- π³ MCTS-inspired parallel beam search for conversation optimization. Explore multiple dialogue strategies simultaneously, stress-test aβ¦β36Jan 18, 2026Updated 7 months ago
- Audiobook Creator is an app that converts books (EPUB, PDF, TXT etc.) into fully voiced audiobooks with intelligent character voice attriβ¦β525Nov 17, 2025Updated 9 months ago
- Generate audiobooks from pdf or epub using Next-gen AI Chatterbox-tts from Resemble-AIβ20Apr 26, 2026Updated 4 months ago
- β18Jul 1, 2025Updated last year
- ComfyUI Discord Bot - Share workflows with your friends through Discord!β15Feb 25, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β117Dec 17, 2025Updated 8 months ago
- Audiobook creation app supporting too many TTS models (Qwen3-TTS, OmniVoice, VibeVoice, etc), focused on high-quality output. Plus audio-β¦β203Updated this week
- A realtime speech to text diarization system to gather and interleave speech from multiple speaker audio.β57Jan 29, 2026Updated 7 months ago
- FastAPI wrapper around original Vibevoice 1.5B and 7B models, with support for AWQ4 quantβ33Jun 22, 2026Updated 2 months ago
- klmbr - a prompt pre-processing technique to break through the barrier of entropy while generating text with LLMsβ90Sep 22, 2024Updated last year
- SoTA open-source TTS for Audiobook and Podcast Generationβ205Jun 19, 2025Updated last year
- A proxy that hosts multiple single-model runners such as LLama.cpp and vLLMβ12May 30, 2025Updated last year
- VellumForge2 is a Golang CLI for generating high-quality Direct Preference Optimization datasets via a hierarchical prompt pipeline with β¦β21Feb 15, 2026Updated 6 months ago
- Glyphs, acting as collaboratively defined symbols linking related concepts, add a layer of multidimensional semantic richness to user-AI β¦β57Feb 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Local-first desktop AI workbench for roleplay, multi-character chat, long-form writing, RAG, MCP tools, plugins, and local models.β125Updated this week
- Extract a target speakerβs clean, non-overlapped speech from multi-speaker audio and export word-safe LJSpeech-style TTS datasets.β23Jun 14, 2026Updated 2 months ago
- Cognito: Supercharge your Chrome browser with AI. Guide, query, and control everything using natural language.β57Updated this week
- A text-grid web renderer for AI agents β see the web without screenshotsβ104Mar 10, 2026Updated 5 months ago
- Vortex is a self-hosted RAG (Retrieval-Augmented Generation) application that lets you chat with your documents using any LLM provider. Uβ¦β17Updated this week
- β24Sep 20, 2025Updated 11 months ago
- A professional-grade interface for Qwen3-TTS, designed to unlock the model's full potential with fine-grained control and intuitive workfβ¦β293Mar 30, 2026Updated 5 months ago
- Echo-TTS OpenAI Compatible Speech Endpoint w/ Streamingβ29Apr 5, 2026Updated 5 months ago
- Open source tool for transcirption and subtitling, alternative to happyscribe.β36Feb 12, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits β’ AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Chatbot-to-speech using Orpheus TTS model. Interactive console app.β21May 1, 2025Updated last year
- An experimental autonomous research system that conducts comprehensive, multi-hour research sessions and produces book-length reports witβ¦β50Dec 7, 2025Updated 9 months ago
- β226May 7, 2025Updated last year
- β36Jul 4, 2026Updated 2 months ago
- Your personal and private AIβ54Apr 3, 2025Updated last year
- Speaker diarization for Python β "who spoke when?" CPU-only, no API keys, Apache 2.0. ~10.8% DER on VoxConverse, 8x faster than real-timeβ¦β119May 6, 2026Updated 4 months ago
- An extension for oobabooga/text-generation-webui that automatically unloads and reloads your model.β17Apr 22, 2024Updated 2 years ago
- π₯ Alternative to Ollama β multi-model serving with sub-ms model switching Β· CPU-only 20B inference for Edge AI Β· llama.cpp + stablediffuβ¦β41Aug 22, 2026Updated 2 weeks ago
- Testbench for llama.cpp llama-serverβ15Aug 20, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MCP server for video/audio processing via FFmpeg - convert, compress, trim, extract audio, add subtitlesβ26May 8, 2026Updated 3 months ago
- A bare-bones GUI application for the local inference engine, llama.cpp. Built-in TPE optimiser to find the best flags for your systemβ18Updated this week
- Personal voice assistant, with voice interruption and Twilio supportβ18Feb 24, 2025Updated last year
- A simple no-install web UI for Ollama and OAI-Compatible APIs!β31Jan 30, 2025Updated last year
- This repository is a CUA (computer use agent) system that, using the Qwen3-VL model on Ubuntu computers, aims to perform tasks on your beβ¦β22Mar 3, 2026Updated 6 months ago
- A lightweight API that returns Nvidia GPU utilisation information.β16Jun 27, 2026Updated 2 months ago
- Soprano: Instant, Ultra-Realistic Text-to-Speechβ1,564Jan 15, 2026Updated 7 months ago