☆35Jan 25, 2026Updated 7 months ago
Alternatives and similar repositories for dataset-maker
Users that are interested in dataset-maker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆146Nov 15, 2025Updated 9 months ago
- High-performance ASR tool using Faster Whisper, supporting custom models, multi-language transcription, and real-time processing feedback…☆10Sep 17, 2025Updated 11 months ago
- ☆16Aug 24, 2025Updated last year
- A simple module for making a request to the tortoise gradio page.☆15Jun 10, 2024Updated 2 years ago
- IndexTTS Fine-tuning notebooks☆139Jun 17, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of the paper PitchFlower: A flow-based neural audio codec with pitch controllability☆37Nov 3, 2025Updated 10 months ago
- A standalone trainer for AceStep 1.5☆22Feb 28, 2026Updated 6 months ago
- 自分用のカスタムノード☆15Jun 6, 2026Updated 3 months ago
- Simple image and video captioning app with a Gradio UI, powered by Qwen2.5/3 VL Instruct.☆28Apr 1, 2026Updated 5 months ago
- ☆11Mar 11, 2025Updated last year
- PuLID face identity preservation for FLUX and Chroma models in ComfyUI☆24Jun 14, 2025Updated last year
- The implementation of the paper *SinGS: Animatable Single-Image Human Gaussian Splats with Kinematic Priors* [CVPR 2025]☆22Nov 11, 2025Updated 9 months ago
- ☆17Sep 4, 2025Updated last year
- TTS pipeline that uses RVC to enhance audio quality and cloning☆150Jan 25, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆12Jul 30, 2026Updated last month
- Finetune Sesame's CSM 1B model, for fun and profit☆17Mar 24, 2025Updated last year
- Watch and hear endless conversations between two ollamas, hence the Two-Way Conversation Engine (TWICE)☆25Oct 13, 2023Updated 2 years ago
- Personal GPEN scripts within the GPEN-Windows stand-alone package.☆20Jun 5, 2022Updated 4 years ago
- Unofficial implementation of AnyText for ComfyUI(EXP)☆57May 22, 2024Updated 2 years ago
- MLT (Kdenlive, others) to FCP (Final Cut Pro, Davinci Resolve, others) video project converter☆11Apr 30, 2019Updated 7 years ago
- PyTorch implementation of Miipher-2 [2025] which is a speech restoration model by Google DeepMind☆70Sep 22, 2025Updated 11 months ago
- Converts Krita .kra files into common used files☆11Mar 15, 2019Updated 7 years ago
- ☆101Aug 14, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Ollama Modelfiles - Discover more at OllamaHub☆22Dec 2, 2023Updated 2 years ago
- ☆44Nov 19, 2025Updated 9 months ago
- Taiwanese Translation with BERT based model and RNN. Collection of Taiwanese text corpus☆13Oct 15, 2022Updated 3 years ago
- Bot de Telegram que facilita pagamentos via Mercado Pago e envia produtos digitais após confirmação. Gerencia produtos, notifica pagament…☆13Jul 6, 2024Updated 2 years ago
- ☆29May 13, 2026Updated 3 months ago
- Dual implementation of reference-based video colorization: ColorMNet (2024) + Deep Exemplar (2019) for ComfyUI☆26Nov 21, 2025Updated 9 months ago
- Landing Page for Divide and Remaster v3☆27Jul 29, 2025Updated last year
- ComfyUI custom nodes for LTXV2, FLUX, other models.☆15Jul 5, 2026Updated 2 months ago
- A guide to help newcomers to the Piper TTS system create voices for NVDA and other screen readers down the line.☆28Dec 5, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Turn dials. Summon bangers! A feature-rich GUI for Ace-Step v1.5☆43Apr 28, 2026Updated 4 months ago
- Incremental Disentanglement for Environment-Aware Zero-Shot Text-to-Speech Synthesis☆27Mar 21, 2025Updated last year
- Python app created with the purpose of speeding up and greatly facilitating the task of cleaning and adjusting Booru-style tags, aimed at…☆12Dec 2, 2023Updated 2 years ago
- Create storybooks using CrewAI, Groq, and Ollama☆18Mar 17, 2024Updated 2 years ago
- Gradio UI for training video models using finetrainers☆35Apr 18, 2025Updated last year
- High-quality speech synthesis with LoRA fine-tuning on index-tts, enhancing prosody and naturalness for single and multi-speaker voices.