Qwen3-ASR speech recognition on Apple Silicon via MLX
☆199Sep 7, 2026Updated this week
Alternatives and similar repositories for mlx-qwen3-asr
Users that are interested in mlx-qwen3-asr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MLX implementation of OminiX for LLM, image generataion, ASR and TTS☆59Aug 23, 2026Updated 2 weeks ago
- AI speech toolkit for Apple Silicon — ASR, TTS, speech-to-speech, VAD, and diarization powered by MLX and CoreML☆1,168Updated this week
- Rust implementation of Qwen3-ASR automatic speech recognition☆252Mar 28, 2026Updated 5 months ago
- MLX Local Serving (MLS) - Unified ASR, TTS, and Translation on Apple Silicon☆17Jun 20, 2026Updated 2 months ago
- Compare different Whisper implementations optimized for Apple Silicon☆50Aug 4, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music…☆3,484Jun 26, 2026Updated 2 months ago
- ☆20Feb 13, 2026Updated 6 months ago
- A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speec…☆7,848Updated this week
- An unofficial implementation of Lite-RTSE, a cost-effective lite model for real-time speech enhancement☆14Nov 19, 2023Updated 2 years ago
- NNSE (Neural Network Speech Enhancement) is a speech-denoiser optimized to run on Ambiq's low power platform☆44Nov 13, 2025Updated 9 months ago
- Run Qwen3-TTS text-to-speech locally on Mac (M1/M2/M3/M4). Voice cloning, voice design, custom voices. 100% offline using MLX.☆561May 2, 2026Updated 4 months ago
- ☆13Jun 24, 2021Updated 5 years ago
- Vocello: a local, private voice studio for Apple Silicon. Write a script, pick or describe a voice, and generate speech on-device, faster…☆362Updated this week
- Audio samples for the paper 'Phase-aware music super-resolution using generative adversarial networks'☆14May 15, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A modular Swift SDK for audio processing with MLX on Apple Silicon☆772Updated this week
- [ACL 2026]☆28Dec 6, 2025Updated 9 months ago
- MLX-Video is the best package for inference and finetuning of Image-Video-Audio generation models on your Mac using MLX.☆295May 13, 2026Updated 3 months ago
- An unofficial code reproduction of Channel Attention Dense U-Net for Multichannel Speech Enhancement☆13Jul 17, 2023Updated 3 years ago
- Implementation of CGMM-MVDR beamforming used for Clarity challenge☆15Jan 14, 2022Updated 4 years ago
- Frontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, …☆2,739Updated this week
- A high-performance API server that provides OpenAI-compatible endpoints for MLX models. Developed using Python and powered by the FastAPI…☆361Aug 31, 2026Updated last week
- Implementation of Sheffield entry for Clarity enhancement challenge.☆18Apr 19, 2022Updated 4 years ago
- 叮当同学 D1X 热敏打印机 HTTP -> BLE 桥☆59Jul 25, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆30May 21, 2026Updated 3 months ago
- This repository contains an unofficial pytorch implementation of BSRNN for music separation, attempting to reproduce the results of the o…☆12Jul 23, 2026Updated last month
- LattifAI skills☆30May 15, 2026Updated 3 months ago
- Efficient Personalized Speech Enhancement through Self-Supervised Learning☆24Mar 12, 2023Updated 3 years ago
- Streaming Text to Speech Web UI☆22May 6, 2024Updated 2 years ago
- Simple Docker container that serves OrcaSlicer via noVNC in your web browser.☆15Nov 30, 2023Updated 2 years ago
- NVV-SuperBench: Beyond Words, Beyond Quality—Benchmarking Nonverbal Vocalizations in Speech Generation (Interspeech 2026 long paper)☆18Jun 21, 2026Updated 2 months ago
- Burn benchmarks☆33Aug 21, 2026Updated 2 weeks ago
- An end-to-end ASR model, transcribing spoken Chinese to formal text.☆22Jun 26, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Modify headers by new declarativeNetRequest API☆17Jul 4, 2022Updated 4 years ago
- PDF2Pod converts PDF documents into short, multi-speaker podcasts—with up to five voices—using OpenAI's GPT-4 API and Eleven Labs TTS.☆20Oct 13, 2024Updated last year
- DSH(DeepSeek Harness)插件:把 ego-lite 浏览器(给 AI Agent 用的 Chromium)接入 HARNESS——13 个结构化 ego_* 工具(文本语义快照、语义定位点击、表单填充、截图、CDP 控制、任务空间隔离),内置 ego 运行…☆125Updated this week
- Alibaba Cloud DDNS for PHP☆12Apr 2, 2022Updated 4 years ago
- Speech detection using silero vad in Rust☆34Dec 16, 2024Updated last year
- Python tools for text to speech (TTS), speech to text (STT), and speech to speech (STS) powered by MLX☆47May 9, 2026Updated 3 months ago
- [Interspeech 2024] Hold Me Tight: Stable Encoder-Decoder Design for Speech Enhancement☆43Jul 25, 2025Updated last year