Quantized text-audio foundation model from Boson AI
☆43Aug 13, 2025Updated last year
Alternatives and similar repositories for higgs-audio_quantized
Users that are interested in higgs-audio_quantized are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Getting VibeVoice 7b working with 10 gb of vram.☆15Aug 31, 2025Updated last year
- ☆59Feb 8, 2026Updated 7 months ago
- A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!☆120Nov 24, 2025Updated 10 months ago
- Keyword spaCy is a spaCy pipeline component for extracting keywords from text using cosine similarity.☆14Dec 7, 2023Updated 2 years ago
- SoTA open-source TTS for Audiobook and Podcast Generation☆207Jun 19, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The first full-screen and multi-monitor solution for video and image batch previewing in ComfyUI☆19Aug 27, 2026Updated last month
- High fidelity neural audio codec for TTS models☆37Dec 22, 2025Updated 9 months ago
- A highly optimized engine for maya-1 tts model to generate minutes of audio in seconds.☆65Nov 17, 2025Updated 10 months ago
- ☆46Oct 28, 2025Updated 10 months ago
- Sesame Converse - Real Time Conversations - Powered by Gemma 3☆64Mar 19, 2025Updated last year
- Proxy for OpenAI☆16Sep 2, 2025Updated last year
- Super simple python connectors for llama.cpp, including vision models (Gemma 3, Qwen2-VL). Compile llama.cpp and run!☆30Dec 11, 2025Updated 9 months ago
- VLLM Port of the Chatterbox TTS model☆385Oct 18, 2025Updated 11 months ago
- Dia-JAX: A JAX port of Dia, the text-to-speech model for generating realistic dialogue from text with emotion and tone control.☆30May 7, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Modelo de Inteligencia Artificial utilizando Computer Vision para la detección y segmentacion de plantas medicinales en la ciudad de Sucr…☆11Apr 10, 2024Updated 2 years ago
- TensorRT implementation of the waifu2x super-resolution model for faster image and video upscaling.☆18Nov 24, 2024Updated last year
- Distributed AI Orchestrator☆15Apr 25, 2026Updated 5 months ago
- VibeVoice: Expressive, longform conversational speech synthesis. (Community fork)☆1,593Aug 29, 2026Updated 3 weeks ago
- Streaming and Fine-tuning for Chatterbox TTS☆294Jun 15, 2025Updated last year
- Welcome to Kokoro, a high-quality text-to-speech synthesis program powered by deep learning. This tool converts any text into high-fideli…☆15Jan 5, 2026Updated 8 months ago
- Loop your image from output to input in your ComfyUI workflow☆14Jan 16, 2026Updated 8 months ago
- ☆18Jul 27, 2025Updated last year
- ☆83Feb 28, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A box containing all necessary components to play with ultrasound.☆10Nov 8, 2020Updated 5 years ago
- ☆66Jun 11, 2026Updated 3 months ago
- Realtime demo, Streaming and Finetuning code for CSM☆459Sep 17, 2025Updated last year
- A Golang Text to Speech System☆17Feb 16, 2026Updated 7 months ago
- ☆19Dec 17, 2024Updated last year
- Shared configuration files☆13Aug 26, 2026Updated last month
- An implementation of the ECS paradigm in Go.☆13Sep 9, 2022Updated 4 years ago
- Sesame CSM 1B Voice Cloning☆339Mar 15, 2025Updated last year
- ☆10Jun 26, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The aim of the project was to convert an image to speech. An image is processed and segmented to identify the text in the image. Then the…☆13Sep 12, 2018Updated 8 years ago
- SoTA open-source TTS☆171Dec 16, 2025Updated 9 months ago
- Realtime Scraper for Twitter based on the Twitter API☆16Oct 31, 2014Updated 11 years ago
- ☆16Jun 22, 2026Updated 3 months ago
- Getting of model metadata, images/videos and keywords from civitai and visualizing them should be easy.☆19Aug 20, 2026Updated last month
- Ren'py GUI☆10Mar 9, 2023Updated 3 years ago
- ComfyUI nodes useful to Tweak Z-Image base models by adjusting specific 'slices'. These nodes break the model and the TextEncoder into ar…☆15Jan 28, 2026Updated 7 months ago