OpenAI compatible API servers for the Qwen3 TTS models
☆86May 19, 2026Updated 2 months ago
Alternatives and similar repositories for qwen3_audio_api
Users that are interested in qwen3_audio_api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Rust implementation of Qwen3-ASR automatic speech recognition☆250Mar 28, 2026Updated 4 months ago
- Android sound localization and classification app.☆15Jul 4, 2025Updated last year
- Native Mac OS GUI for Using mlx-lm-lora.☆65Dec 19, 2025Updated 7 months ago
- A pipecat bot demo implementation of a Spotify assistant for creating playlists☆20Oct 14, 2025Updated 10 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Onset-and-Offset-Aware Sound Event Detection☆21Feb 10, 2025Updated last year
- vibevoice real time 0.5B swift port☆31Dec 12, 2025Updated 8 months ago
- Multi-talker ASR based on DiCoW with Serialized Output Training☆21Sep 18, 2025Updated 11 months ago
- Code for our tutorial on Discrete Variational Autoencoders☆33May 19, 2025Updated last year
- Supercharge your Gaianet node by generating a vector knowledge base from any API. Demo slides: https://hackmd.io/@santteegt/ByoykY4nC#/ L…☆11Nov 29, 2024Updated last year
- language for describing user interfaces in 3D space using pmndrs/uikit☆20May 28, 2026Updated 2 months ago
- A Discord ChatGPT Bot☆10Aug 29, 2023Updated 2 years ago
- Bleeding edge vLLM Docker image for the NVIDIA DGX Spark (GB10 / sm_121a).☆47Updated this week
- The original BabyAGI, updated with LiteLLM and no vector database reliance (csv instead)☆22Oct 2, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SSVM Node.js Addon☆13Aug 27, 2021Updated 4 years ago
- Get aid from local LLMs right in your PowerShell☆17May 2, 2025Updated last year
- wasm-interface-types supplement & compiler of wasmedge☆18Aug 25, 2023Updated 2 years ago
- a naive 3d human pose editor GUI.☆21Jul 12, 2023Updated 3 years ago
- Stacks -- an infinite imaginative canvas☆22Jun 11, 2026Updated 2 months ago
- Make the bot work so you don't have to learn a language.☆10Jun 23, 2021Updated 5 years ago
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆593Feb 17, 2026Updated 6 months ago
- ☆12Apr 23, 2025Updated last year
- Prediction of sound event bounding boxes (SEBBs)☆35Aug 2, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Rust library for using stable diffusion functions when the Wasi is being executed on WasmEdge.☆13Oct 31, 2024Updated last year
- Real-time text-to-speech with Qwen3-TTS☆1,312Jul 17, 2026Updated last month
- ☆12Dec 23, 2024Updated last year
- ☆15Oct 27, 2020Updated 5 years ago
- Claw but for every user, and computer.☆224Feb 18, 2026Updated 6 months ago
- Towards a general language-audio model for computational paralinguistic tasks☆31Dec 14, 2024Updated last year
- Websocket hook for React☆19Aug 4, 2026Updated 2 weeks ago
- Official page of "DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis"☆28Apr 15, 2026Updated 4 months ago
- ☆15Aug 29, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A FastAPI wrapper for NVIDIA's new parakeet 0.6b v3 TTS 600m model designed for high-quality multilingual speech recognition, beating Whi…☆212Jul 23, 2026Updated 3 weeks ago
- [NeurIPS 2023] AV-NeRF: Learning Neural Fields for Real-World Audio-Visual Scene Synthesis☆36Feb 15, 2024Updated 2 years ago
- Call phone number for speaking with an AI assistant that lets you book an appointment with your preferred doctor over voice.☆12Aug 31, 2023Updated 2 years ago
- ☆13Apr 22, 2025Updated last year
- A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speec…☆7,744Updated this week
- CLI to capture snapshots, short clips, and run motion detection against RTSP/ONVIF came ras☆107Aug 6, 2026Updated last week
- VAD (Voice Activity Detection) as a service, written in Rust☆17Dec 6, 2025Updated 8 months ago