这是一个基于 Python 开发的实时语音字幕显示程序,可以将用户的语音实时转换为屏幕上的字幕文本。支持中文和英文识别,适用于 macOS 和 Windows 系统
☆28Dec 25, 2024Updated last year
Alternatives and similar repositories for VoiceSubtitle
Users that are interested in VoiceSubtitle are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A real-time caption translation tool based on VOSK speech recognition and machine translation, which supports transcribing audio into tar…☆10Mar 12, 2025Updated last year
- 本项目使用python对影响共享单车使用量的因素进行可视化分析,并使用lightGBM算法对已知条件下的共享单车使用量进行预测。其中为了选择最优模型,使用了k折交叉验证和网格搜索选择最优参数。☆10Jul 15, 2020Updated 6 years ago
- 使用Spring Boot作为后端,Vue作为前端,采用SSE技术实现类似GPT官网的打字机效果和语言交互。☆14Jan 23, 2024Updated 2 years ago
- 基于TSN网络模型的手语识别系统☆16Oct 21, 2020Updated 5 years ago
- 放置一些半斤化八两个人整理的数据分析项目☆17Jan 3, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 基于MiniMax API的实时语音翻译系统,支持多语言同声传译、语音识别、文本翻译和语音合成☆20Nov 17, 2025Updated 10 months ago
- an ffmpeg plugin for Dify☆20Dec 1, 2025Updated 9 months ago
- A study about the Generalized Cross-Correlation with Phase Transform algorithm.☆14Nov 23, 2021Updated 4 years ago
- This node is base on VisualCloze method, A Universal Image Generation Framework via Visual In-Context Learning☆11May 21, 2025Updated last year
- speech enhancement using DNN: [1] Xu, Y., Du, J., Dai, L.R. and Lee, C.H., 2015. A regression approach to speech enhancement based on dee…☆14Sep 17, 2019Updated 7 years ago
- 自适应的小波阈值降噪☆14Aug 11, 2023Updated 3 years ago
- Whisper realtime streaming for long speech-to-text transcription and translation☆62Apr 9, 2024Updated 2 years ago
- This repository consists of application of Speech Denoising using DNN, CNN (1D and 2D) and RNN (LSTM) in tensorflow.☆16Jun 15, 2019Updated 7 years ago
- 录制麦克风或者系统扬声器的声音,并实时翻译,自动纠错☆16Jan 22, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 基于pynq-z2的声源定位系统☆14Nov 15, 2020Updated 5 years ago
- Modified version of OpenVINO noise_suppression_demo. This version can handle real-time audio stream from microphone and output to headpho…☆16Aug 5, 2021Updated 5 years ago
- 全卷积网络进行语音降噪☆18Dec 8, 2021Updated 4 years ago
- The single-microphone noise reduction algorithm based on IMCRA-OMLSA has been implemented and improved. Besides, a transient noise suppre…☆18Oct 25, 2023Updated 2 years ago
- Scaled Uniform Noise for Ancestral & Stochastic samplers and Noisy latent image☆17Mar 30, 2025Updated last year
- Platform for Audio Filtering (Digital Filters) in Real-Time using Convolution Theorem and Fast Fourier Transform.☆13Aug 16, 2021Updated 5 years ago
- 基于matlab的数字信号降噪系统以及GUI界面☆15Dec 19, 2021Updated 4 years ago
- 小红书的flux版本的透明图生成(layerdiffuse),支持文生图和图生图☆18Mar 17, 2025Updated last year
- FPGA based, Real-time processing of audio, including voiceprint recognition, adaptive noise suppression, et al.☆18May 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- https://github.com/xie-lab-ml/Golden-Noise-for-Diffusion-Models for ComfyUI☆18Dec 10, 2024Updated last year
- 本项目包含一个 Python 脚本,用于分离双人(或多人)对话播客音频文件中的不同说话人语音。它利用 `pyannote.audio` 库进行说话人日志分析(Speaker Diarization),找出“谁在什么时候说话”,并将每个说话人的语音 片段提取到单独的音轨中。☆17Apr 30, 2025Updated last year
- An Experimental Study on Speech Enhancement based on DNN.☆14Aug 11, 2018Updated 8 years ago
- This repo contains the scripts, models and required files for the Interspeech 2020 Deep Noise Suppression (DNS) Challenge. We are open so…☆15May 15, 2020Updated 6 years ago
- Python package of MP-SENet from Explicit Estimation of Magnitude and Phase Spectra in Parallel for High-Quality Speech Enhancement.☆22Nov 1, 2024Updated last year
- Inference of resemble denoiser☆30Mar 11, 2024Updated 2 years ago
- ☆15May 27, 2025Updated last year
- ComfyUI ShadowR Wrapper☆18Feb 21, 2025Updated last year
- This is an unofficial Pytorch implementation of the DTLN model repository, which contains denoising and inference code for the DTLN model…☆24Jun 18, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Pytorch Models for Speech Enhancement☆24Mar 31, 2023Updated 3 years ago
- Node to tryoff clothes☆23Apr 14, 2025Updated last year
- 语音情感识别代码,结合1D-CNN与GRU在语音增强的CASIA数据集实现语音情感识别,并利用注意力机制进行模型优化☆19Jan 16, 2022Updated 4 years ago
- 通过单层圆形麦克风阵列采集音频,实现MUSIC算法的声源定位。☆23Mar 16, 2023Updated 3 years ago
- ☆20May 13, 2026Updated 4 months ago
- Noise reduction for speech enhancement using matlab☆23Jun 14, 2015Updated 11 years ago
- MCP MySQL Server 是一个基于 @modelcontextprotocol/sdk 的 MySQL 工具服务,支持 SQL 查询、表结构获取、连接检测等功能,适用于 AI 代理、自动化工具等场景。☆32Aug 27, 2025Updated last year