OpenAI compatible API servers for the Qwen3 TTS models
☆88May 19, 2026Updated 3 months ago
Alternatives and similar repositories for qwen3_audio_api
Users that are interested in qwen3_audio_api are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Rust implementation of Qwen3-ASR automatic speech recognition☆252Mar 28, 2026Updated 5 months ago
- A Rust implementation of the Qwen3 Text-to-Speech (TTS) model inference.☆233May 19, 2026Updated 3 months ago
- A white-box hacking guide: From Google Colab to Qwen3.5☆26Mar 10, 2026Updated 6 months ago
- A pipecat bot demo implementation of a Spotify assistant for creating playlists☆20Oct 14, 2025Updated 10 months ago
- Evals meant to evaluate language models' ability to reason over long contexts.☆10Sep 12, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Onset-and-Offset-Aware Sound Event Detection☆22Feb 10, 2025Updated last year
- vibevoice real time 0.5B swift port☆31Dec 12, 2025Updated 9 months ago
- Multi-talker ASR based on DiCoW with Serialized Output Training☆21Sep 18, 2025Updated 11 months ago
- Supercharge your Gaianet node by generating a vector knowledge base from any API. Demo slides: https://hackmd.io/@santteegt/ByoykY4nC#/ L…☆11Nov 29, 2024Updated last year
- language for describing user interfaces in 3D space using pmndrs/uikit☆20May 28, 2026Updated 3 months ago
- Production ready, bleeding edge vLLM Docker image for the NVIDIA DGX Spark (GB10 / sm_121a).☆48Updated this week
- The original BabyAGI, updated with LiteLLM and no vector database reliance (csv instead)☆22Oct 2, 2024Updated last year
- Get aid from local LLMs right in your PowerShell☆17May 2, 2025Updated last year
- A general ChatGPT bot for your slack channel☆16Jul 7, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- tiny_fnc_engine is a minimal python library that provides a flexible engine for calling functions extracted from a LLM.☆37Aug 13, 2026Updated last month
- C inference for Qwen3-ASR 0.6b and 1.7b transcriptions models☆611Feb 17, 2026Updated 6 months ago
- This repository contains the checkpoints and training code for the few-shot adaptation speech models in the SpidR-Adapt paper.☆24Dec 29, 2025Updated 8 months ago
- chat with a live2d model Immersive☆11Updated this week
- ☆12Apr 23, 2025Updated last year
- Super-Mailer Send Inbox To All☆10Jun 8, 2020Updated 6 years ago
- Prediction of sound event bounding boxes (SEBBs)☆35Aug 2, 2024Updated 2 years ago
- A Rust library for using stable diffusion functions when the Wasi is being executed on WasmEdge.☆13Oct 31, 2024Updated last year
- Real-time text-to-speech with Qwen3-TTS☆1,347Aug 25, 2026Updated 2 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆15Oct 27, 2020Updated 5 years ago
- [ICLR 2025] NeRAF jointly learns acoustic and radiance fields, enabling realistic audio-visual generation.☆37Mar 11, 2026Updated 6 months ago
- Towards a general language-audio model for computational paralinguistic tasks☆31Dec 14, 2024Updated last year
- Official page of "DeepASA: An Object-Oriented Multi-Purpose Network for Auditory Scene Analysis"☆31Apr 15, 2026Updated 4 months ago
- ☆15Aug 29, 2021Updated 5 years ago
- ☆37Updated this week
- ☆13Apr 22, 2025Updated last year
- A text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speec…☆7,876Updated this week
- ☆41May 12, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- CLI to capture snapshots, short clips, and run motion detection against RTSP/ONVIF came ras☆141Updated this week
- datatochart.com☆12Mar 29, 2025Updated last year
- A SOTA Industrial-Grade Voice Activity Detection & Audio Event Detection, supporting 100+ languages, outperforming Silero-VAD, TEN-VAD, F…☆526May 6, 2026Updated 4 months ago
- NVFP4 Gemma-4 26B-A4B MoE for DGX Spark — optimal recipe: DFlash n=10 (flex) on AEON vLLM Ultimate. 144 tok/s single / 1,724 peak (Coding…☆44Jun 28, 2026Updated 2 months ago
- ☆38Oct 10, 2025Updated 11 months ago
- A template project for building high-performance, portable, and safe serverless functions in Netlify.☆18Aug 2, 2021Updated 5 years ago
- ☆12Mar 11, 2025Updated last year