Whisper系列のPEFTと、PEFT済のモデルを使ったストリーミング書き起こしを実装するためのリポジトリです。
☆15Oct 16, 2025Updated 10 months ago
Alternatives and similar repositories for WhisperLive-PEFT
Users that are interested in WhisperLive-PEFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10May 16, 2024Updated 2 years ago
- This project uses llama.cpp as an LLM server to perform inference and generate speech using Synthetic voice library☆22Mar 5, 2024Updated 2 years ago
- ☆15Nov 10, 2025Updated 9 months ago
- Aivis Voice Model File (.aivm/.aivmx) Generator / Editor☆15Feb 5, 2026Updated 6 months ago
- [ACL 2025] OZSpeech: One-step Zero-shot Speech Synthesis with Learned-Prior-Conditioned Flow Matching☆45Feb 9, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Dec 12, 2019Updated 6 years ago
- speaker-disentangled speech linguistic content quantizer☆26Mar 19, 2025Updated last year
- AI based singing voice synthesis☆37Jun 10, 2024Updated 2 years ago
- 日本語の短いテーマから、画像生成プロンプト&和訳とアップスケールした絵とセリフと感情付き音声を、雑然と生成する EasyZatuGen です。☆38Jan 24, 2024Updated 2 years ago
- ☆40Oct 21, 2025Updated 10 months ago
- ☆25Oct 22, 2025Updated 10 months ago
- A real-time software for turn-taking, backchannel, and head-nodding prediction☆127Updated this week
- ☆49Jul 22, 2024Updated 2 years ago
- Unofficial entropix impl for Gemma2 and Llama and Qwen2 and Mistral☆17Jan 12, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆25Feb 3, 2026Updated 6 months ago
- AIキャラクター・フィラー・笑い声・感情表現に特化した日本語TTSコーパス(CC0)☆22May 10, 2026Updated 3 months ago
- ☆35Oct 23, 2025Updated 10 months ago
- pytorch model for contexless-phoneme prediction from speech audio☆32Oct 30, 2025Updated 10 months ago
- superfast text to speech in any voice☆63Feb 16, 2026Updated 6 months ago
- Fine-tuning Moshi/J-Moshi on your own spoken dialogue data☆105Jan 5, 2026Updated 7 months ago
- ☆23Dec 23, 2025Updated 8 months ago
- nosdav spec☆11Apr 24, 2023Updated 3 years ago
- flutter WebRTC sample ios app with flutter_ios_voip_kit https://github.com/masashi-sutou/flutter_ios_voip_kit☆13Jul 23, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Latitude and longitude conversion to UTM || Arduino IDE || Wiring☆13Dec 5, 2019Updated 6 years ago
- ☆14Jul 10, 2021Updated 5 years ago
- NeRF implementation with minimal code and maximal readability using PyTorch☆11Aug 27, 2022Updated 4 years ago
- ☆22Aug 28, 2024Updated 2 years ago
- Googleの音声復元モデルMiipher-2の再現実装の学習および推論コード。学習済みモデルも公開しています。☆33Feb 7, 2026Updated 6 months ago
- ☆26Feb 16, 2026Updated 6 months ago
- 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.☆18Jul 21, 2023Updated 3 years ago
- 🤖🛠🗺 TRIDENT - AI Assistant for Interactive Smart Maps☆26Updated this week
- 英単語から読みを推測するライブラリ。☆31May 4, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Unofficial PyTorch implementation of "Autoregressive Speech Synthesis without Vector Quantization (MELLE)"☆41Jun 28, 2025Updated last year
- 🚀 Simple demo of Streaming React Server Component☆16Apr 11, 2022Updated 4 years ago
- codeinterpreter-api with Streamlit☆39Jul 21, 2023Updated 3 years ago
- Run OpenAI Codex Desktop on Linux - automated installer☆22Aug 8, 2026Updated 3 weeks ago
- alpacaデータセットを日本語化したものです☆86Jun 3, 2023Updated 3 years ago
- sampling frequency independent convolution for MOS prediction☆16Jul 22, 2025Updated last year
- ☆27Oct 5, 2025Updated 10 months ago