High-performance ASR tool using Faster Whisper, supporting custom models, multi-language transcription, and real-time processing feedback.
☆10Sep 17, 2025Updated 10 months ago
Alternatives and similar repositories for SpeedScribe
Users that are interested in SpeedScribe are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open Source Text-to-Speech GUI Tool running on TalkNet☆11Dec 24, 2022Updated 3 years ago
- 100+ billion unique prompt. Create random prompts to help with learning better prompting techniques. Works with any AI platform!☆16Oct 10, 2024Updated last year
- ☆15Jun 23, 2024Updated 2 years ago
- Turn any common eBook file into an HQ Audiobook with F5-TTS (Easy Install)☆39Apr 6, 2026Updated 4 months ago
- Listen, transcribe, reply - Voice Assistant using OpenAI & ElevenLabs API's☆14Jun 24, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- EbSynth in Python, version 2☆36Jun 10, 2026Updated 2 months ago
- An AR+AR TTS attempt.☆18Jan 13, 2025Updated last year
- An official implementation of Style-Talker for Spoken Dialogue Generation☆23Jan 12, 2025Updated last year
- The implementation of the paper *SinGS: Animatable Single-Image Human Gaussian Splats with Kinematic Priors* [CVPR 2025]☆22Nov 11, 2025Updated 9 months ago
- Next-generation, fully open-source refacer. Images. GIFs. TIFFs. Full-length videos. Bulk refacing☆43May 16, 2025Updated last year
- ☆17Sep 4, 2025Updated 11 months ago
- StyleTTS 2 Optimized Training Fork☆32Feb 2, 2025Updated last year
- The AI-native CAD platform☆30Updated this week
- Vid Driven Portrait Animation 🤢😷☆18Jul 7, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Generate images from an initial frame and text☆37Jul 28, 2023Updated 3 years ago
- Personal GPEN scripts within the GPEN-Windows stand-alone package.☆20Jun 5, 2022Updated 4 years ago
- ☆25Sep 27, 2022Updated 3 years ago
- ☆16Aug 24, 2025Updated 11 months ago
- Musculoskeletal Analysis extension for 3D Slicer. Currently has cortical, cancellous, and bone density analysis.☆13May 2, 2024Updated 2 years ago
- ☆46Dec 3, 2024Updated last year
- ☆19Jan 15, 2024Updated 2 years ago
- A fork of Rope with webcam support☆13Mar 13, 2024Updated 2 years ago
- NeMo: a toolkit for conversational AI☆12Dec 23, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis☆27Mar 21, 2025Updated last year
- Higgs Audio v2 WebUI + One click installer WIN x64☆21Jul 25, 2025Updated last year
- Video dubbing with tts or other audio☆17Apr 23, 2025Updated last year
- StyleFlow: Attribute-conditioned Exploration of StyleGAN-generated Images using Conditional Continuous Normalizing Flows☆51Apr 29, 2021Updated 5 years ago
- A front-end GUI for interacting with AI Horde's distributed cluster of Stable Diffusion workers☆26Jul 4, 2025Updated last year
- A diffusion-based cross-lingual voice conversion model, as my bachelor's thesis☆45Jul 24, 2023Updated 3 years ago
- Post-processing OCR errors with seq2seq models☆28Jul 30, 2020Updated 6 years ago
- Adversarial Training of Denoising Diffusion Model Using Dual Discriminators for High-Fidelity Multi-Speaker TTS☆40Aug 4, 2023Updated 3 years ago
- TRAIL: Simulating the Impact of Human Locomotion on Natural Landscapes - 2024 - Computer Graphics International (CGI)☆13Apr 7, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Generative voice cloning model using TTS synthesis with state-of-the-art Zero-Shot Multi-Speaker functionality. An web api built with the…☆47Jan 4, 2023Updated 3 years ago
- Supervoice diffusion enhance☆28Jul 15, 2024Updated 2 years ago
- ☆10Oct 23, 2024Updated last year
- ☆20May 3, 2024Updated 2 years ago
- Orchestrating AI for stunning lip-synced videos. Effortless workflow, exceptional results, all in one place.☆79Jun 19, 2025Updated last year
- Centralized multi-channel notification management component for streamlined communication across email, SMS, WhatsApp, and push notificat…☆13Aug 5, 2026Updated last week
- ☆21Dec 8, 2023Updated 2 years ago