☆23Aug 1, 2026Updated 3 weeks ago
Alternatives and similar repositories for gepard-train
Users that are interested in gepard-train are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆26Nov 3, 2025Updated 9 months ago
- ☆13Aug 24, 2023Updated 3 years ago
- ☆53Feb 19, 2026Updated 6 months ago
- A high-performance batch audio transcription tool using nvidia/parakeet-tdt-0.6b-v2 to generate accurate, well-segmented SRT subtitles, w…☆18Dec 9, 2025Updated 8 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆254Jan 13, 2026Updated 7 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆22Aug 21, 2025Updated last year
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 2 months ago
- Kubernetes Pod File System Explorer☆12Feb 12, 2024Updated 2 years ago
- Native End-to-End Full-Duplex Spoken Language Model☆124Updated this week
- LLM-based ASR recipe with Zipformer encoder and Qwen LLM☆35Sep 25, 2025Updated 10 months ago
- From scratch Python implementation of a genetic algorithm that recreates a target image☆11Jan 30, 2023Updated 3 years ago
- PyPi package for KaniTTS-2 model☆62Jun 24, 2026Updated 2 months ago
- ☆16Jul 23, 2026Updated last month
- ☆25Oct 22, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated 11 months ago
- ☆60Dec 17, 2025Updated 8 months ago
- BackstageCon 2023☆10Jul 14, 2025Updated last year
- ☆58Feb 8, 2026Updated 6 months ago
- A beginner-friendly inference to finetune & run inference on open TTS models 🗣️☆30Feb 4, 2026Updated 6 months ago
- Open-source speech AI models from KRAFTON, including Raon-Speech and Raon-SpeechChat for speech understanding, generation, and real-time …☆115Apr 7, 2026Updated 4 months ago
- Self-hosted voice for coding agents. Talk from any browser or a Telegram call, interrupt mid-sentence, clone any voice, and hand real wor…☆36Aug 17, 2026Updated last week
- ☆75Jul 29, 2026Updated 3 weeks ago
- Fast audio super resolution from 16khz to 48khz.☆217Jan 3, 2026Updated 7 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆46Oct 30, 2025Updated 9 months ago
- SLT 2024 Challenge: Post-ASR-Speaker-Tagging☆16Jun 16, 2024Updated 2 years ago
- ☆61Apr 1, 2026Updated 4 months ago
- Official repository for Mamba-based Segmentation Model for Speaker Diarization☆47May 13, 2025Updated last year
- This repository facilitates the creation of Python wheel files (.whl) from the tiny-cuda-nn project to streamline the installation proces…☆12Jul 2, 2025Updated last year
- Variations of L1 SNR Loss function for training audio source separation machine learning models☆45Aug 11, 2026Updated last week
- This repository contains a series of works on diffusion-based speech tokenizers, including the official implementation of the paper: "TaD…☆197Jan 25, 2026Updated 6 months ago
- [ICLR 2026] Official code for BézierFlow: Learning Bézier Stochastic Interpolant Schedulers for Few-Step Generation☆23Apr 13, 2026Updated 4 months ago
- ☆459Nov 2, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆52Jul 5, 2026Updated last month
- VyvoTTS: LLM-Based Text-to-Speech Training Framework☆261Aug 9, 2026Updated 2 weeks ago
- Train your own speech AI model from scratch☆152May 23, 2026Updated 3 months ago
- Gemma-based Multilingual Machine Translation Models☆89Aug 14, 2026Updated last week
- Update ASR paper everyday☆512May 16, 2026Updated 3 months ago
- Open-source text-to-speech model from KRAFTON trained exclusively on public speech data, with curated datasets and reproducible training …☆92May 21, 2026Updated 3 months ago
- A high quality and fast TTS repository☆521Dec 22, 2025Updated 8 months ago