whisper.cpp bindings for python
☆112Aug 24, 2023Updated 3 years ago
Alternatives and similar repositories for whisper-cpp-python
Users that are interested in whisper-cpp-python are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python bindings for whisper.cpp☆347Aug 22, 2026Updated 2 weeks ago
- Pybind11 bindings for Whisper.cpp☆344Dec 8, 2024Updated last year
- Python bindings for whisper.cpp☆249Jun 1, 2024Updated 2 years ago
- Offline srt producer gui with whisper.cpp☆25Dec 31, 2023Updated 2 years ago
- ☆13Aug 7, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Minimal user-friendly demo of OpenAI's CLIP for semantic image search☆20Sep 28, 2024Updated last year
- Jointly encoding word confusion networks (WCNs) and dialogue context with BERT for spoken language understanding (SLU).☆12Jun 12, 2023Updated 3 years ago
- proof of concept conversation orchestrator with a speech-language model☆20Oct 19, 2024Updated last year
- Mixture of Expert (MoE) techniques for enhancing LLM performance through expert-driven prompt mapping and adapter combinations.☆11Feb 11, 2024Updated 2 years ago
- Python bindings for llama.cpp☆10,600Aug 17, 2026Updated 3 weeks ago
- Docker for building an environment for Dutch online and offline ASR.☆12Feb 2, 2021Updated 5 years ago
- Open-source object detection for Python developers. Frictionless installation. Free for commercial use.☆22Feb 18, 2026Updated 6 months ago
- Minimal extension of OpenAI's Whisper adding speaker diarization with special tokens☆553Nov 6, 2023Updated 2 years ago
- Transcribe Offline by openresearchtools.com is an open source desktop application that allows you to transcribe audio and video fully off…☆17Aug 18, 2026Updated 2 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- wake word spotting with kaldi☆19Dec 3, 2020Updated 5 years ago
- Simple Python library, distributed via binary wheels with few direct dependencies, for easily using wav2vec 2.0 models for speech recogni…☆23Aug 16, 2021Updated 5 years ago
- Emotion Recognition from Brazilian Portuguese Informal Spontaneous Speech☆22Mar 21, 2022Updated 4 years ago
- Baseline convolutional ASR system in PyTorch☆21Nov 16, 2023Updated 2 years ago
- Code repository for the paper "Improving End-to-End SLU performance with Prosodic Attention and Distillation" accepted at Interspeech 202…☆27May 17, 2023Updated 3 years ago
- Domain Adaptation and Adapters☆16Feb 28, 2023Updated 3 years ago
- Python package of MP-SENet from Explicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech Enhancement.☆22Nov 1, 2024Updated last year
- Simple diarization model☆53Jun 13, 2025Updated last year
- Speech-MASSIVE is a multilingual Spoken Language Understanding (SLU) dataset comprising the speech counterpart for a portion of the MASSI…☆25Oct 8, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- pico w powered led matrix with mvg departure information☆11Oct 23, 2024Updated last year
- The llama-cpp-agent framework is a tool designed for easy interaction with Large Language Models (LLMs). Allowing users to chat with LLM …☆660Mar 9, 2026Updated 5 months ago
- Standalone implementation of the CUDA-accelerated WFST Decoder available in Riva☆91Feb 18, 2025Updated last year
- Social previews generator as a microservice.☆12Apr 9, 2022Updated 4 years ago
- The YouTube Text-To-Speech dataset is comprised of waveform audio extracted from YouTube videos alongside their English transcriptions☆53Apr 1, 2021Updated 5 years ago
- ☆27Aug 31, 2022Updated 4 years ago
- Learning Pytorch☆13Jun 12, 2018Updated 8 years ago
- A python package to build AI-powered real-time audio applications☆2,024Jun 19, 2026Updated 2 months ago
- Use quantized versions of Whisper to speed up inference☆12Oct 16, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Suno AI's Bark model in C/C++ for fast text-to-speech generation☆867Nov 16, 2024Updated last year
- Python bindings for llama.cpp☆199Apr 22, 2023Updated 3 years ago
- Pytorch Lightning Template for Sematic Segmentation☆11Jan 17, 2023Updated 3 years ago
- RUSSE: Russian Semantic Evaluation.☆15Mar 1, 2022Updated 4 years ago
- SeeSR: Towards Semantics-Aware Real-World Image Super-Resolution☆14Jan 12, 2024Updated 2 years ago
- Falcon LLM ggml framework with CPU and GPU support☆250Jul 2, 2026Updated 2 months ago
- Audio Recorder for python that let's you record WAV and MP3 Files from any input source☆14Feb 11, 2025Updated last year