☆11Dec 24, 2024Updated last year
Alternatives and similar repositories for Language-Group
Users that are interested in Language-Group are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- egrecho project☆11Apr 30, 2026Updated 4 months ago
- LAFMA: A Latent Flow Matching Model for Text-to-Audio Generation (INTERSPEECH 2024)☆44Jun 13, 2024Updated 2 years ago
- Code release for AccDiffusionV2 (TPAMI)☆33Nov 4, 2025Updated 9 months ago
- Official PyTorch implementation of the paper "Robust Training for Speaker Verification against Noisy Labels" in INTERSPEECH 2023.☆12Oct 23, 2023Updated 2 years ago
- ☆45Jan 26, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models☆116Oct 30, 2025Updated 10 months ago
- Official Repository of IJCAI 2024 Paper: "BATON: Aligning Text-to-Audio Model with Human Preference Feedback"☆32Mar 4, 2025Updated last year
- Auto-KWS 2021 Challenge 1st place solution.☆11Jul 20, 2021Updated 5 years ago
- Research code for the paper "Training speaker recognition systems with limited data" at https://arxiv.org/abs/2203.14688☆13Dec 2, 2024Updated last year
- Data manipulation and transformation for audio signal processing, powered by PyTorch☆10Sep 30, 2024Updated last year
- Code release for VTW (AAAI 2025 Oral)☆66Nov 4, 2025Updated 9 months ago
- 儿童故事常识推理与寓意理解评测(Commonsense Reasoning and Moral Understanding Evaluation in Children's Stories,CRMU)☆18Oct 22, 2024Updated last year
- ☆34Apr 5, 2026Updated 4 months ago
- One command to start a streaming ASR server.☆12Oct 2, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Jun 4, 2026Updated 2 months ago
- Pytorch implementation of our paper accepted by NeurIPS 2022 -- Learning Best Combination for Efficient N:M Sparsity☆22Jan 13, 2023Updated 3 years ago
- sherpa with mlx☆15Aug 2, 2025Updated last year
- ☆12Jul 11, 2024Updated 2 years ago
- wenet_LLM_from_ASLP☆15Nov 26, 2024Updated last year
- A repo containing download guidance and corresponding scripts of the VoxBlink dataset.☆30Apr 16, 2024Updated 2 years ago
- Code associated with the paper: CTC-DRO: Robust Optimization for Reducing Language Disparities in Speech Recognition.☆17May 16, 2025Updated last year
- ICASSP 2022: 'Self-supervised Speaker Recognition with Loss-gated Learning'☆91May 29, 2023Updated 3 years ago
- 🤖 Node.js auto deploy demo based on GitHub Webhook☆29Oct 23, 2017Updated 8 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- funasr语音转文字的简单api版本,funasr+fastapi,方便部署在服务器上☆13Aug 10, 2024Updated 2 years ago
- Survey on speech generation work.☆21Nov 26, 2023Updated 2 years ago
- NeurIPS 2020, "A Topological Filter for Learning with Label Noise".☆31Apr 11, 2025Updated last year
- [INTERSPEECH 2026 Oral]Official code for "Semantic-VAE: Semantic-Alignment Latent Representation for Better Speech Synthesis"☆124Jun 21, 2026Updated 2 months ago
- An open source community implementation of the model MELLE from the paper: "Autoregressive Speech Synthesis without Vector Quantization"☆16Updated this week
- Landing Page for Divide and Remaster v3☆27Jul 29, 2025Updated last year
- 基于wenet的短时在线语音识别服务☆11Feb 25, 2023Updated 3 years ago
- ☆14Aug 9, 2021Updated 5 years ago
- <综合> Funasr语音识别,调用Qwen大模型回答,通过GPTSovits输出语音的ai程序,其中调用模型还是在线,后续将添加离线大模型☆13Nov 30, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ASR_LLM_TTS前端项目☆15Dec 3, 2024Updated last year
- What Is a Good Caption? A Comprehensive Visual Caption Benchmark for Evaluating Both Correctness and Thoroughness☆28May 16, 2025Updated last year
- ☆30Jan 7, 2023Updated 3 years ago
- Train no-reference speech quality estimators with multiple datasets via learned, per-dataset alignments.☆19Aug 1, 2025Updated last year
- ☆16Nov 9, 2023Updated 2 years ago
- ☆31Aug 28, 2022Updated 4 years ago
- auto scrawl for arrive data☆16Jan 24, 2022Updated 4 years ago