Generate transcriptions and subtitles using OpenAI whisper as a base model, stable-ts/whisperx as a timestamp stabilizer using ASR models and pyannote/nemo models in order to identify different speakers.
☆19Mar 10, 2023Updated 3 years ago
Alternatives and similar repositories for whisper_subtitler
Users that are interested in whisper_subtitler are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆15Aug 27, 2026Updated last week
- this master thesis project is based on OpenAI Whisper with the goal to transcibe interviews☆47Aug 6, 2024Updated 2 years ago
- A python script COMMAND LINE utility to AUTO GENERATE SUBTITLE FILE (using faster_whisper module which is a reimplementation of OpenAI Wh…☆31Updated this week
- ☆12Mar 25, 2024Updated 2 years ago
- node.js 敏感词/违禁词 检测 替换 过滤 ,超高效率,极小内存(8万个违禁词仅需要30MB)☆15Apr 15, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A simple offline *.srt translator using transformers models that helps you to translate subtitles.简单的离线字幕翻译器,使用transformers模型,助你翻译字幕。☆10Jun 13, 2023Updated 3 years ago
- Real-time multi-person pose estimation☆22Oct 19, 2018Updated 7 years ago
- Speech Recognition and Simple AI Summary:可用于本地语音转文字、说话人分割及简易的AI总结,搭配web端操作界面。☆11Jul 22, 2024Updated 2 years ago
- gradio bbox labeling tools☆11May 12, 2023Updated 3 years ago
- a version tools. face detector,face landmark detector,face parsing and so on☆12Jul 30, 2022Updated 4 years ago
- Merge and clean up multi-line and multi-language subtitle files. Updated with language-based subtitle split. 将带有多行英文的SRT字幕合并成单行,同时合并中文翻译。…☆14Mar 30, 2015Updated 11 years ago
- PyQt(+PySide) Stable Diffusion GUI☆15Aug 1, 2023Updated 3 years ago
- A simple os x desktop app built in electron, using Aeneas under the hood to generate captions files from media(audio or video) and plain …☆29Jan 4, 2023Updated 3 years ago
- # Vue 项目开源 凡客网站重构项目,具有完整的业务流程,以及后台数据api☆12Jan 15, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Connected Data London 2025 Masterclass: Combining Data from Structured and Unstructured Sources to create High-Quality Knowledge Graphs☆16Jan 30, 2026Updated 7 months ago
- 个人用的FFmpeg命令行、avs、vs代码备份☆12Oct 25, 2023Updated 2 years ago
- ANDROID APP that can RECOGNIZE VLC LIVE AUDIO/VIDEO STREAMING (using free Android Developers Speech Recognition API) then TRANSLATE (usin…☆21May 5, 2024Updated 2 years ago
- Offline Speaker Diarization with SenseVoice by Sherpa ONNX.☆15Dec 23, 2024Updated last year
- 科大讯飞线下销量挑战赛top7方案☆13Aug 21, 2021Updated 5 years ago
- Exploration of World Languages☆18Apr 5, 2024Updated 2 years ago
- Aegisub - Assdraw - AI2ASS☆23Jan 7, 2017Updated 9 years ago
- This Machine Learning project deals with Coupon Recommendations based on Revenue Uplift☆11May 4, 2021Updated 5 years ago
- Download videos from iq.com☆20Oct 22, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 创建软/硬链接的右键菜单(Windows)☆16Nov 26, 2023Updated 2 years ago
- Allow a UITabBar to work in kiosk mode.☆15May 17, 2012Updated 14 years ago
- GDPnet: "Geometry-guided Dense Perspective Network for Speech-Driven Facial Animation." (TVCG 2021)☆11Nov 21, 2021Updated 4 years ago
- This is the proposal network for MultiPerson Pose Estimation.☆14Oct 21, 2017Updated 8 years ago
- Value stepper control for iOS.☆21Sep 25, 2015Updated 10 years ago
- Deep Learning in Audio processing☆17Feb 2, 2023Updated 3 years ago
- A Colab Notebook for OpenAI Whisper and DeepL API, aiming to create human-comparable results of translation and transcription.☆33Feb 4, 2024Updated 2 years ago
- Uses ONNX Runtime for character role speaker identification.☆17Dec 28, 2025Updated 8 months ago
- [CVPR 2021] FMO Deblurring Benchmark☆15Jan 12, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Reads one or more audio files and creates a spectrogram visualization, with optional particle effects.☆14Mar 2, 2023Updated 3 years ago
- 企业web项目认证licences (在软件或产品交付时,我们往往会授权给第三方或者防止程序乱部署而对部署的服务器进行认证,此时License就排上用途了。授权的方便在于如果证书过期,我们可以重新生成一个认证文件而不用修改程序。)☆22Nov 19, 2018Updated 7 years ago
- The official code for [ECCV2020] "HALO: Hardware-aware Learning to Optimize"☆10Mar 22, 2023Updated 3 years ago
- TaiYiXLCheckpointLoader: An unoffical node support Taiyi-Diffusion-XL(Taiyi-XL) Chinese-English bilingual language model☆10Sep 1, 2024Updated 2 years ago
- GMOT-40: A Benchmark for Generic Multiple Object Tracking (CVPR 2021)☆40Apr 3, 2025Updated last year
- Web Service for Human Pose Annotation☆10Oct 21, 2018Updated 7 years ago
- 主流APP基本骨架的快速搭建☆15Sep 21, 2015Updated 10 years ago