Official PyTorch inference for Breeze TTS 2
☆498Sep 9, 2026Updated 2 weeks ago
Alternatives and similar repositories for breeze-tts
Users that are interested in breeze-tts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- InteractAvatar is a novel dual-stream DiT framework that enables talking avatars to perform Grounded Human-Object Interaction (GHOI)☆24Jun 3, 2026Updated 3 months ago
- Brand new TTS solution☆11Dec 7, 2024Updated last year
- VITS2 using Phoneme-Level Japanese BERT☆14Dec 17, 2023Updated 2 years ago
- 大量の音声データから笑い声部分を集めるやつ☆14May 23, 2024Updated 2 years ago
- Official PyTorch implementation of "EdVAE: Mitigating Codebook Collapse with Evidential Discrete Variational Autoencoders"☆14Sep 20, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- AI agent running in local environment.☆13Apr 24, 2024Updated 2 years ago
- Comfy EverAnimate WIP☆17Jun 5, 2026Updated 3 months ago
- Deep Latent Gamma Model / Gamma VAE☆14May 9, 2018Updated 8 years ago
- Template for creating Ultravox demo that gets deployed to Vercel.☆20Mar 6, 2025Updated last year
- A fork of Rope with webcam support☆13Mar 13, 2024Updated 2 years ago
- Full code for all experiments in the paper "Denoising Hamiltonian Network for Physical Reasoning"☆15Mar 21, 2025Updated last year
- 44100Hz日本語HuBERTに対応した QuickVC: Any-to-many Voice Conversion Using Inverse Short-time Fourier Transform for Faster Conversion です。☆16May 21, 2023Updated 3 years ago
- ComfyUI-AniSora is now available in ComfyUI, Index-AniSora is the most powerful open-source animated video generation model. It enables o…☆57May 27, 2025Updated last year
- ☆12Feb 19, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- GUI-focused roop☆15Mar 6, 2025Updated last year
- This repo contains script to download MUSIC dataset from youtube☆13Jan 19, 2024Updated 2 years ago
- Sound field reconstruction using neural processes with dynamic kernels☆16Mar 25, 2025Updated last year
- singing voice conversion without f0☆22May 10, 2023Updated 3 years ago
- Real-time end-to-end singing voice convertion☆25Nov 3, 2024Updated last year
- ☆15Jul 18, 2026Updated 2 months ago
- RVCのWebUIを補助するTampermonkeyスクリプト☆18May 23, 2023Updated 3 years ago
- Real-time face swap for PC streaming or video calls☆18Jan 8, 2025Updated last year
- Official Codebase of "A Unified Audio-Visual Learning Framework for Localization, Separation, and Recognition" (ICML 2023)☆12Jun 1, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆78Mar 18, 2024Updated 2 years ago
- This is the official implementation of RL-Chord (TNNLS).☆13Jan 2, 2024Updated 2 years ago
- [ICLR 2026] AudioMCQ: A 571k audio multiple-choice question dataset for post-training Large Audio Language Models with dual CoT annotatio…☆52Apr 21, 2026Updated 5 months ago
- Open TTS implementation for ViiTorVoice-NAR☆21Jul 2, 2026Updated 2 months ago
- ☆13Jan 31, 2021Updated 5 years ago
- Comfy UI node to prompt build for https://huggingface.co/cagliostrolab/animagine-xl-4.0 model☆16Jul 20, 2025Updated last year
- A custom ComfyUI node for MiniCPM vision-language models, supporting v4, v4.5, and v4 GGUF formats, enabling high-quality image captionin…☆156Aug 28, 2025Updated last year
- ComfyUI Nodes that integrate GLSL shader support.☆20Aug 25, 2025Updated last year
- ☆14Jan 26, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Python scripts to create noisy and reverberant 2-speaker mixture audio with Libri-Light and WHAM☆17Nov 7, 2024Updated last year
- A ComfyUI custom node enabling **Flash Attention 1** on legacy NVIDIA GPUs (Tesla V100, T4) that lack Compute Capability 8.0+ required by…☆28Feb 9, 2026Updated 7 months ago
- ☆27Mar 20, 2024Updated 2 years ago
- You can use SHMT method to apply makeup to the characters when use ComfyUI☆29Jan 9, 2025Updated last year
- ComfyUI Node wrapper of SunoAI API☆22Dec 17, 2024Updated last year
- ☆21Apr 24, 2025Updated last year
- TensorFlow implementation of "Finite Scalar Quantization: VQ-VAE Made Simple" (ICLR 2024)☆21Dec 3, 2023Updated 2 years ago