豆包输入法语音识别的非官方 Python 客户端
☆29Feb 4, 2026Updated 7 months ago
Alternatives and similar repositories for doubaoime-asr
Users that are interested in doubaoime-asr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- semantic tokenizer for speech and music☆20Jul 6, 2025Updated last year
- ☆35Sep 6, 2025Updated last year
- trying to reproduce suno v3☆36Jan 29, 2025Updated last year
- 处理VIVO手机安装验证问题☆12Jun 25, 2021Updated 5 years ago
- A streaming audio reader, processor, and writer built on top of soundfile, and PyAV (bindings for FFmpeg)☆39Aug 27, 2026Updated 3 weeks ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Sep 12, 2024Updated 2 years ago
- Evaluation tool used in the BigVSAN paper☆14Mar 22, 2024Updated 2 years ago
- Tensorflow implementation of DeepMind's Tacotron-2 (without wavenet)☆11Jul 12, 2019Updated 7 years ago
- A tool that utilizes actions to generate Generic System Images (GSIs).☆11Mar 19, 2024Updated 2 years ago
- 基于 Cloudflare Workers 的 Microsoft 365 用户自助开通与轻量级管理面板。这是一个无服务器(Serverless)的解决方案,用于快速部署 M365 账号分发系统。它包含一个面向用户的自助注册页面,以及一个面向管理员的简易后台,支持用户管理、…☆17Jan 22, 2026Updated 7 months ago
- ☆11Updated this week
- Implementation of "DurIAN: Duration Informed Attention Network For Multimodal Synthesis".☆15Jul 6, 2020Updated 6 years ago
- ☆11Sep 11, 2026Updated last week
- ✨✨VITA: Towards Open-Source Interactive Omni Multimodal LLM☆11Jun 16, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Auto Read 是一款专为 Linux.do 社区设计的油猴脚本,旨在自动化阅读流程,帮助用户高效浏览和管理未读帖子。脚本通过模拟真实用户行为,自动滚动阅读内容,并支持自动点赞功能,大幅提升社区浏览体验。☆16Jun 8, 2025Updated last year
- Fast Streaming TTS with MTP Acceleration and X-pred Mean Flow Distillation☆75Jun 16, 2026Updated 3 months ago
- 基于手淘 flexible.js 的 Vue 组件☆16Dec 26, 2017Updated 8 years ago
- FREECODEC: A DISENTANGLED NEURAL SPEECH CODEC WITH FEWER TOKENS☆24Sep 9, 2024Updated 2 years ago
- 使用 PHP 编写的云湖机器人 SDK☆11Oct 19, 2024Updated last year
- Training code for kokoro tts model☆45Nov 15, 2025Updated 10 months ago
- 这是一个由第三方提供的轻巧的智学网PC客户端(具有Windows/Linux/Mac版),方便电脑端查分☆18Apr 3, 2024Updated 2 years ago
- ☆17Dec 26, 2025Updated 8 months ago
- Torch Audio Forced Aligner for Mixed Chinese (Mandarin or Cantonese) and English.☆63Sep 5, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Models and codes for INTERSPEECH 2023 paper DistilXLSR: A Light Weight Cross-Lingual Speech Representation Model☆13Mar 30, 2025Updated last year
- Finally, a native GUI application for NoneBot2 written in pure Python.☆14Sep 4, 2023Updated 3 years ago
- Some microbenchmarks and design docs before commencement☆11Feb 1, 2021Updated 5 years ago
- WhisperMesh is an advanced chatbot that integrates voice and text interactions, delivering personalized responses through LLM models and …☆17Apr 23, 2025Updated last year
- A pitch detection model trained to be robust against noise and reverberation environments.☆27Jan 21, 2025Updated last year
- A simple command line tool to calculate WER for ASR.☆15Jul 28, 2026Updated last month
- differentiable top-k operator☆23Dec 30, 2024Updated last year
- Code for CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignment, Interspeech 2026.☆21Aug 14, 2026Updated last month
- Complex-number-aware Variational Autoencoder for audio tasks☆16Jun 2, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆25Mar 29, 2026Updated 5 months ago
- OpenLuaX+开源项目☆13Nov 23, 2024Updated last year
- ☆19Mar 18, 2026Updated 6 months ago
- ☆17Apr 27, 2026Updated 4 months ago
- Android Utils Collection.☆17Sep 9, 2026Updated last week
- ☆12Jul 23, 2024Updated 2 years ago
- Indonesian speech/phoneme recognizer powered by Kaldi 2.0 (lhotse, icefall, sherpa).☆17Jun 30, 2023Updated 3 years ago