Taiwan Tongues ASR CE 是一個開源語音辨識(Automatic Speech Recognition, ASR)模型專案,專為台灣多元語言環境設計。 本模型支援 國語、台語、客語與英語,提供本地多語混合語音辨識,讓開發者與資訊服務業者可運用此開源模型進行 ASR 模型訓練、微調與發展在地化應用,以低成本、高效率進行 ASR 語音應用落地與智慧服務創新。 本專案為數位發展部數位產業署「114年數位產業跨域軟體基盤系統建置案」之實證成果之一,旨在推動台灣語音技術開源生態,協助資訊服務業者強化智慧應用能量,落實在地 AI 技術自主發展,由台灣大哥大執行與維護。
☆74Jun 24, 2026Updated 2 months ago
Alternatives and similar repositories for Taiwan-Tongues-ASR-CE
Users that are interested in Taiwan-Tongues-ASR-CE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 台灣閩南語大型語言模型 (Taiwanese Hokkien LLMs)☆69Aug 15, 2024Updated 2 years ago
- Breeze ASR 25 是一款先進的自動語音辨識(ASR)模型,基於 Whisper-large-v2 微調而成,特別針對台灣華語以及華語與英語混用的情境進行優化。Breeze ASR 25 is an advanced ASR model fine-tuned fro…☆203Jul 1, 2025Updated last year
- A Python tool that uses Google Gemini API to transcribe video or audio files into SRT subtitle files.☆21Jan 2, 2026Updated 7 months ago
- This repository focuses on leveraging OpenAI's Whisper model for speech recognition in Chinese (Mandarin) and Taiwanese Hokkien languages…☆73Mar 1, 2025Updated last year
- PhahTaigi - Taigi Input Method for iOS☆17Mar 3, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Evaluation code for benchmarking VLMs in traditional chinese understanding☆14Dec 22, 2025Updated 8 months ago
- Opus 4.6 Telegram AI assistant via Claude Max + OpenClaw. One-click setup.☆53Mar 21, 2026Updated 5 months ago
- GitHub 熱門專案追蹤工具,基於 React 19 與 Vite 6 構建,內建 AI 分析功能☆19Mar 28, 2026Updated 5 months ago
- 台灣媠聲標記網站☆15Oct 1, 2025Updated 10 months ago
- Toward Multi Modality Language Model - implementation of GPT-4o/Project Astra☆16Dec 10, 2024Updated last year
- Standardized agent data feeds and automation instruction sets for the OpenClaw framework. 專為 OpenClaw 框架設計的標準化代理資料流與自動化指令集。☆99May 16, 2026Updated 3 months ago
- Taiwanese Speech Synthesis with Tacotron2☆26Oct 2, 2022Updated 3 years ago
- Self-hosted, OpenAI-compatible AI gateway for organizations — multi-provider (Azure/OpenAI/Anthropic/Gemini), revocable per-allocation cr…☆15Updated this week
- High-performance LLM evaluation framework with parallel API calls — up to 17× faster than sequential tools. Supports box, math, and logit…☆109Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆15Aug 4, 2025Updated last year
- Generative Fusion Decoding (GFD) is a novel framework for integrating Large Language Models (LLMs) into multi-modal text recognition syst…☆87Jul 31, 2025Updated last year
- ☆335Jun 21, 2025Updated last year
- This is the repository for Learning to Generate Piano Music With Sustain Pedals☆12Nov 23, 2023Updated 2 years ago
- This repository is the implementation of the HiPAMA architecture, introduced in the paper, Hierarchical Pronunciation Assessment with Mul…☆40Apr 29, 2024Updated 2 years ago
- ☆19Updated this week
- The official repository for "Piano score rearrangement into multiple difficulty levels via notation-to-notation approach" incl. ST+ token…☆13Feb 26, 2024Updated 2 years ago
- ☆15Jul 29, 2022Updated 4 years ago
- Progressive update about OpenClaw☆118May 4, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for paper "Unsupervised Noise adaptation using Data Simulation"☆14May 16, 2024Updated 2 years ago
- ☆10Jun 11, 2024Updated 2 years ago
- ☆12Apr 18, 2025Updated last year
- ☆19Jul 12, 2020Updated 6 years ago
- Python script to transform the Mobile Detect JSON database into an UA-based mobile detection VCL subroutine easily integrable in any Varn…☆14Nov 13, 2023Updated 2 years ago
- CosyVoice3 text-to-speech for Unity using ONNX inference. Supports zero-shot voice cloning☆15Jan 15, 2026Updated 7 months ago
- The dataset repo of "CLCIFAR: CIFAR-Derived Benchmark Datasets with Human Annotated Complementary Labels" paper☆17May 11, 2026Updated 3 months ago
- Local-first Traditional Chinese dictation for macOS☆63Aug 1, 2026Updated 3 weeks ago
- Official implementation of "PhonMatchNet: Phoneme-Guided Zero-Shot Keyword Spotting for User-Defined Keywords" (INTERSPEECH 2023)☆63Jun 3, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A survey of spoken dialogue models (SDMs) with speech input and speech output. Focus on their Intermediate Representation and Generation …☆32Mar 24, 2026Updated 5 months ago
- Unofficial implementation of SCP-GAN☆18Jul 4, 2023Updated 3 years ago
- Real-time bilingual subtitles for any browser video. Offline-first speech recognition (Whisper + SenseVoice, 100+ languages) with local O…☆35Jul 31, 2026Updated last month
- The complete video-production skill behind the 蝦說 AI channel — lets an AI agent autonomously produce narrated educational videos (slides …☆101Jul 3, 2026Updated last month
- Codes for "Benchmarking the Generation of Fact Checking Explanations"☆10Aug 16, 2024Updated 2 years ago
- A truth inference tool in crowdsourcing☆13May 19, 2020Updated 6 years ago
- [ICME 2024 oral] Official Repository for The Paper, PianoBART: Symbolic Piano Music Understanding and Generating with Large-Scale Pre-Tra…☆23Aug 17, 2025Updated last year