This is the combined forks of two repos to enable OpenAI Whisper large image with VAD for low VRAM GPUs.
☆33Mar 25, 2023Updated 3 years ago
Alternatives and similar repositories for whisper-webui-vad
Users that are interested in whisper-webui-vad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Charisma SDK for Unreal Engine 4.☆20Sep 8, 2025Updated last year
- Simple Android SDK for Publitio☆10Jan 16, 2021Updated 5 years ago
- This app is intended to automatically create a corpus for ASR systems using pseudo-labeling.☆27Feb 15, 2024Updated 2 years ago
- Robust Speech Recognition via Large-Scale Weak Supervision☆28Dec 16, 2023Updated 2 years ago
- ☆10Nov 6, 2021Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Image inpainting based on LAMA☆15Jul 4, 2022Updated 4 years ago
- Open Source Routing Machine for OpenStreetMap API Lib and App for Nim☆10Jun 6, 2019Updated 7 years ago
- A curated list of awesome open source and commercial platforms for serving models in production 🚀☆50Apr 20, 2022Updated 4 years ago
- Classified ads is internet messaging system done right☆11Feb 16, 2025Updated last year
- ☆13Dec 29, 2023Updated 2 years ago
- Experiment with github actions on a Nim project☆10Apr 23, 2020Updated 6 years ago
- 🎙️ Fast, installable, in-browser audio spectrum visualizer. Support for both realtime and audio files!☆18Mar 9, 2025Updated last year
- Fast Audio/Video transcribe using Openai's Whisper and Modal, an hour audio/video file can be transcribed in ~1 minute☆79Jul 30, 2023Updated 3 years ago
- ☆12Oct 25, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Typesafe IIIF presentation v3 parsing without external dependencies☆12Jun 29, 2026Updated 3 months ago
- a wordpress plugin to display interactive transcripts☆13Feb 20, 2026Updated 7 months ago
- UI frontend for lucidrains/big-sleep☆15Jan 26, 2021Updated 5 years ago
- Examples of hosting models with fal☆20Updated this week
- Open source Structure-from-Motion pipeline☆18May 6, 2025Updated last year
- 夏目悠李/男声歌声データベースの最新ラベルデータ☆12Sep 2, 2020Updated 6 years ago
- uvx is now uvenv☆16Dec 4, 2024Updated last year
- AI companions with memory: a lightweight stack to create and host your own AI companions☆16Jul 13, 2023Updated 3 years ago
- ☆12Nov 26, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jun 23, 2026Updated 3 months ago
- ☆16Oct 31, 2023Updated 2 years ago
- Visualization for hidden Markov model computations☆14Dec 19, 2014Updated 11 years ago
- Seamlessly integrate AI agents with Chargebee using AgentKit for smarter billing and subscription workflows.☆16Nov 18, 2025Updated 10 months ago
- Expected edit distance implementation using OpenFst tools☆11May 13, 2015Updated 11 years ago
- Create embeddings for LLM using the Nomic API☆23Nov 21, 2024Updated last year
- ☆23Aug 31, 2024Updated 2 years ago
- An introduction to global assessment techniques using Python☆12Apr 24, 2023Updated 3 years ago
- Wenet speech to text for react native☆10Nov 1, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12May 20, 2025Updated last year
- Collection of official scripts created by the Dione Team.☆15Feb 21, 2026Updated 7 months ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Powered by SideGuide and GPT-3☆13Jan 31, 2023Updated 3 years ago
- A simple streamlit app that performs Retrieval-Augmented Generation over a corpus of presidential speeches☆15Apr 24, 2024Updated 2 years ago
- Multiobjective Optimization Training of PLDA for Speaker Verification☆10Jun 14, 2018Updated 8 years ago
- Open source image editor for windows 10. Can be controlled by voice commands and Cortana.☆17Nov 14, 2017Updated 8 years ago