☆12May 23, 2023Updated 3 years ago
Alternatives and similar repositories for whisper_child_asr
Users that are interested in whisper_child_asr are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- public child-adult speaker diarization/classification model and codes☆19Apr 24, 2025Updated last year
- A list of papers for child ASR☆54Oct 8, 2024Updated last year
- Official implementation of "Unsupervised Pre-training for Data-Efficient Text-to-Speech on Low Resource Languages", ICASSP 2023☆27Apr 27, 2023Updated 3 years ago
- 🐍📦 Ultra-fast Python package for calculating and analyzing the Word Error Rate (WER). Built for the scalable evaluation of speech and t…☆29Mar 30, 2026Updated 3 months ago
- Simplified recipes for preparing commonly used speech datasets, and a PyTorch-compatible Python data loader that can perform standard fea…☆15Jun 25, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Simple LPC vocoder in Python☆13Jan 7, 2022Updated 4 years ago
- Companion repository for the "Muddling through SDL GPU" series of posts at www.jonathanfischer.net☆15Apr 2, 2025Updated last year
- ☆21May 26, 2023Updated 3 years ago
- NTI Buddhist Text Reader with Taisho canon and embedded Chinese-English Buddhist dictionary☆16Updated this week
- ☆10Oct 25, 2019Updated 6 years ago
- Cross-Linguistic Transcription Systems☆17Mar 20, 2026Updated 4 months ago
- Network Analysis Interface for Literature Studies☆23May 12, 2023Updated 3 years ago
- A SQL interface for the CHILDES child language corpora☆14Sep 30, 2022Updated 3 years ago
- Optimizing speaker verification and spoofing countermeasure systems together with REINFORCE☆13Mar 31, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Learnable Gammatone Filterbank (LGTFB) and Equal-loudness Normalization (EN)☆13Apr 24, 2020Updated 6 years ago
- Code for ICLR 2024 Paper: CompA: Addressing the Gap in Compositional Reasoning in Audio-Language Models☆23Jul 10, 2024Updated 2 years ago
- Real-time GPT-4o video/photo/voice chat☆41Jul 12, 2026Updated 2 weeks ago
- APL machine learning library☆18Sep 1, 2025Updated 10 months ago
- A list of books about mathematical subjects, using array languages like APL and J for their presentation.☆18Jan 19, 2026Updated 6 months ago
- Language independent SSL-based Speaker Anonymization system☆20May 28, 2024Updated 2 years ago
- An implement of SPEECHSPLIT☆15Sep 12, 2020Updated 5 years ago
- Toy O☆16Sep 21, 2024Updated last year
- ☆19May 19, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Template that combines PyTorch Lightning and Hydra☆16Aug 15, 2023Updated 2 years ago
- A simple APL derivative, built on fixed-arity functions☆27May 24, 2026Updated 2 months ago
- 修复funasr中seaco-paraformer导出onnx后没有时间戳的bug☆25Sep 12, 2024Updated last year
- Multi-Modal Language Modeling with Image, Audio and Text Integration, included multi-images and multi-audio in a single multiturn.☆18Feb 20, 2024Updated 2 years ago
- 语音识别 语音前端处理 语音合成 语音转换等等语音技术的资料汇总☆23Nov 8, 2019Updated 6 years ago
- Face recognition using Facenet☆18May 17, 2019Updated 7 years ago
- R interface to childes-db☆14Updated this week
- Gender stereotypes are reflected in the distributional structure of 25 languages☆18Feb 6, 2023Updated 3 years ago
- Visual Speech Recongnition☆21Dec 24, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- An example implementation of RLHF (or, more accurately, RLAIF) built on MLX and HuggingFace.☆37Jun 21, 2024Updated 2 years ago
- Paper2Code: Automating Code Generation from Scientific Papers in Machine Learning☆14Apr 25, 2025Updated last year
- Repository contains code to fine-tune WhisperASR model☆23Dec 16, 2022Updated 3 years ago
- ☆26Mar 31, 2026Updated 3 months ago
- A simple tool that generates a bunch of 50 second clips from one video. Easily create youtube shorts and tiktoks from longer videos.☆11Feb 10, 2023Updated 3 years ago
- Local only speech to text cli with diarization☆27Sep 14, 2025Updated 10 months ago
- ModernBERT model optimized for Apple Neural Engine.☆38Jan 10, 2025Updated last year