Whisper系列のPEFTと、PEFT済のモデルを使ったストリーミング書き起こしを実装するためのリポジトリです。
☆15Oct 16, 2025Updated 11 months ago
Alternatives and similar repositories for WhisperLive-PEFT
Users that are interested in WhisperLive-PEFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10May 16, 2024Updated 2 years ago
- This project uses llama.cpp as an LLM server to perform inference and generate speech using Synthetic voice library☆22Mar 5, 2024Updated 2 years ago
- ☆15Nov 10, 2025Updated 11 months ago
- Aivis Voice Model File (.aivm/.aivmx) Generator / Editor☆15Feb 5, 2026Updated 8 months ago
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Dec 12, 2019Updated 6 years ago
- speaker-disentangled speech linguistic content quantizer☆26Mar 19, 2025Updated last year
- AI based singing voice synthesis☆37Jun 10, 2024Updated 2 years ago
- 日本語の短いテーマから、画像生成プロンプト&和訳とアップスケールした絵とセリフと感情付き音声を、雑然と生成する EasyZatuGen です。☆38Jan 24, 2024Updated 2 years ago
- ☆41Oct 21, 2025Updated 11 months ago
- ☆26Oct 22, 2025Updated 11 months ago
- Real-time Turn-taking, Backchannel, and Head-nodding Prediction for Full-duplex Interaction☆154Oct 1, 2026Updated last week
- ☆49Jul 22, 2024Updated 2 years ago
- Unofficial entropix impl for Gemma2 and Llama and Qwen2 and Mistral☆17Jan 12, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AIキャラクター・フィラー・笑い声・感情表現に特化した日本語TTSコーパス(CC0)☆22May 10, 2026Updated 5 months ago
- ☆35Oct 23, 2025Updated 11 months ago
- superfast text to speech in any voice☆63Feb 16, 2026Updated 7 months ago
- A debugging tool for VMC protocol☆14Sep 9, 2022Updated 4 years ago
- ☆23Dec 23, 2025Updated 9 months ago
- nosdav spec☆11Apr 24, 2023Updated 3 years ago
- ☆16Apr 2, 2025Updated last year
- flutter WebRTC sample ios app with flutter_ios_voip_kit https://github.com/masashi-sutou/flutter_ios_voip_kit☆13Jul 23, 2020Updated 6 years ago
- Latitude and longitude conversion to UTM || Arduino IDE || Wiring☆13Dec 5, 2019Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆22Aug 28, 2024Updated 2 years ago
- Googleの音声復元モデルMiipher-2の再現実装の学習および推論コード。学習済みモデルも公開しています。☆33Feb 7, 2026Updated 8 months ago
- ☆26Feb 16, 2026Updated 7 months ago
- 🤖🛠🗺 TRIDENT - AI Assistant for Interactive Smart Maps☆26Updated this week
- 英単語から読みを推測するライブラリ。☆31May 4, 2026Updated 5 months ago
- Unofficial PyTorch implementation of "Autoregressive Speech Synthesis without Vector Quantization (MELLE)"☆41Jun 28, 2025Updated last year
- codeinterpreter-api with Streamlit☆39Jul 21, 2023Updated 3 years ago
- Run OpenAI Codex Desktop on Linux - automated installer☆23Aug 8, 2026Updated 2 months ago
- alpacaデータセットを日本語化したものです☆86Jun 3, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- sampling frequency independent convolution for MOS prediction☆16Jul 22, 2025Updated last year
- Awesome Dev Containers template for developing a trading bot.☆15Dec 3, 2022Updated 3 years ago
- ☆27Oct 5, 2025Updated last year
- web enhancer extension based on nostr☆22Jul 20, 2022Updated 4 years ago
- JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit☆44Mar 13, 2026Updated 6 months ago
- ☆13Jan 1, 2025Updated last year
- [ICASSP'26] Real-time streaming voice anonymization & voice conversion☆96Jun 23, 2026Updated 3 months ago