Convert your PDFs and EPUBs into audiobooks effortlessly. Features intelligent text extraction, customizable text-to-speech settings, and efficient processing for low-resource systems.
☆202Feb 26, 2026Updated 6 months ago
Alternatives and similar repositories for pdf-narrator
Users that are interested in pdf-narrator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Transform your PDFs into captivating audio podcasts with this PDF-to-Podcast pipeline! Combining advanced language models and high-qualit…☆17Nov 11, 2024Updated last year
- Generate audiobooks from EPUBs, PDFs and text with synchronized captions.☆5,919Updated this week
- An open-source read-along document reader server with high-quality TTS options, synchronized highlighting, and audiobook export for EPUB,…☆504Updated this week
- Automatically convert epubs to audiobooks☆265Mar 8, 2025Updated last year
- Official code for "F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching"☆54Apr 13, 2026Updated 4 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A Socratic dialogue engine for AI agents.☆16Nov 30, 2025Updated 9 months ago
- A CLI text-to-speech tool using the Kokoro model, supporting multiple languages, voices (with blending), and various input formats includ…☆1,842Aug 22, 2026Updated 2 weeks ago
- StyleTTS-ZS: Efficient High-Quality Zero-Shot Text-to-Speech Synthesis with Distilled Time-Varying Style Diffusion☆10Sep 22, 2024Updated last year
- 🔊 Kokoro Web: Free AI text-to-speech, online or self-hosted, OpenAI compatible!☆733Mar 16, 2025Updated last year
- A powerful MCP tool for parsing and manipulating MIDI files based on Tone.js. This library leverages the Model Context Protocol (MCP) to …☆11May 9, 2025Updated last year
- Open API and Wyoming wrapper around Chatterbox☆32Aug 18, 2026Updated 3 weeks ago
- A self-hosted version of WaterCrawl, a powerful web crawling and data extraction platform.☆13Jul 27, 2025Updated last year
- find it hard to understand long github repos and pdfs? struggle no more, just enter your mindpalace. mindpalace helps you understand the …☆15Aug 27, 2025Updated last year
- 🔥🔥 Kokoro in Rust. https://huggingface.co/hexgrad/Kokoro-82M Insanely fast, realtime TTS with high quality you ever have.☆817Aug 4, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- CLASP: Contrastive Language-Speech Pretraining for Multilingual Multimodal Information Retrieval☆13Jun 27, 2025Updated last year
- Transform unstructured documents into actionable, structured data with enterprise-grade precision and reliability, ready for large-scale …☆21Oct 13, 2025Updated 10 months ago
- Never write commit messages again. Auto-checkpoint every change, auto-squash into professional git history. For Claude Code.☆17Jul 19, 2025Updated last year
- ☆99Apr 27, 2024Updated 2 years ago
- chatterbox TTS + Voice Clone using onnx☆28Aug 2, 2026Updated last month
- Generate audiobooks from e-books☆8,538Feb 27, 2026Updated 6 months ago
- Unofficial implementation of ConvNeXt-TTS powered by lightning☆18Oct 20, 2024Updated last year
- ☆44Oct 9, 2025Updated 10 months ago
- Forced alignment decoder for Whisper.☆16Mar 13, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Audiobook Creator is an app that converts books (EPUB, PDF, TXT etc.) into fully voiced audiobooks with intelligent character voice attri…☆525Nov 17, 2025Updated 9 months ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- https://hf.co/hexgrad/Kokoro-82M☆8,726Aug 6, 2025Updated last year
- Inference codebase for "Cacophony: An Improved Contrastive Audio-Text Model". Preprint: https://arxiv.org/abs/2402.06986☆49Jan 19, 2026Updated 7 months ago
- Scaled Uniform Noise for Ancestral & Stochastic samplers and Noisy latent image☆17Mar 30, 2025Updated last year
- Generate audiobooks from pdf or epub using Next-gen AI Chatterbox-tts from Resemble-AI☆20Apr 26, 2026Updated 4 months ago
- OmegaViT (ΩViT) is a cutting-edge vision transformer architecture that combines multi-query attention, rotary embeddings, state space mod…☆15Aug 28, 2026Updated last week
- ☆15Feb 2, 2026Updated 7 months ago
- ⛔ ARCHIVED — migrated to mcpcentral-io/mcpcentral apps/deep-researcher (ADR-043, 2026-07-23)☆16Jul 23, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Apr 16, 2026Updated 4 months ago
- TTS with kokoro and onnx runtime☆2,708Updated this week
- Software for a person-sized, full body DIY Raspberry Pi based 3D Scanner☆12Oct 7, 2023Updated 2 years ago
- A Dnn 8/9 responsive theme using Bootstrap 4 http://www.dnncontra.com☆12Apr 15, 2018Updated 8 years ago
- Text-to-Speech conversor for Basque and Spanish. It includes linguistic processing and built voices for the languages aforementioned. Its…☆18Jan 15, 2026Updated 7 months ago
- Expressive Speech Retrieval using Natural Language Descriptions of Speaking Style☆15Aug 18, 2025Updated last year
- Generate audiobooks from e-books, voice cloning & 1158+ languages!☆20,134Updated this week