Extensions to YAML syntax for better python interaction
☆80Jan 1, 2026Updated 8 months ago
Alternatives and similar repositories for HyperPyYAML
Users that are interested in HyperPyYAML are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Speech enhancement using mimic loss☆16Oct 25, 2019Updated 6 years ago
- ☆16Mar 12, 2024Updated 2 years ago
- This repository contains the baseline system for CHiME-8 MMCSG challenge focusing on transcribing both sides of a conversation where one …☆41Mar 13, 2024Updated 2 years ago
- SMIT: A Simple Modality Integration Tool☆16Mar 31, 2024Updated 2 years ago
- kaldi based x-vector trained on Cn-Celeb☆13Sep 22, 2020Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Streaming Audio Models Examples in JS☆21Mar 29, 2024Updated 2 years ago
- NVV-SuperBench: Beyond Words, Beyond Quality—Benchmarking Nonverbal Vocalizations in Speech Generation (Interspeech 2026 long paper)☆18Jun 21, 2026Updated 2 months ago
- Segment an audio file and obtain utterance alignments. (Python package)☆348May 15, 2024Updated 2 years ago
- Official Implementation of TSELM: Target speaker extraction using discrete tokens and language models☆63Apr 14, 2025Updated last year
- A recipe for disfluency detection on the LibriStutter dataset using SpeechBrain☆11Mar 13, 2021Updated 5 years ago
- iSTFTNet : Fast and Lightweight Mel-spectrogram Vocoder Incorporating Inverse Short-time Fourier Transform☆14Aug 25, 2023Updated 3 years ago
- Sequence algorithms for use in Flashlight.☆14Jan 12, 2026Updated 7 months ago
- ☆18Mar 13, 2024Updated 2 years ago
- Project page for paper Self-supervised Representation Learning with Relative Predictive Coding☆19Jul 8, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Collection of scripts from mHuBERT-147.☆35Nov 19, 2024Updated last year
- Dan's repository of OpenFst (manually created by downloading certain versions of OpenFst), created to track certain patches.☆13Mar 8, 2016Updated 10 years ago
- simple version of our torch kaldi toolkit, developed at the LIA by 2 apprentices. (@Chaanks & @vbrignatz)☆10Oct 10, 2021Updated 4 years ago
- This repository contains the SpeechBrain Benchmarks☆141Feb 3, 2026Updated 7 months ago
- A simple package for Guided source separation (GSS)☆133May 20, 2024Updated 2 years ago
- End-to-End Korean Automatic Speech Recognition leveraging PyTorch and Hydra.☆10Jan 21, 2022Updated 4 years ago
- ☆14Aug 9, 2018Updated 8 years ago
- Scripts for data generation, scoring and data manifest preparation for CHiME-8 DASR task.☆27Aug 10, 2026Updated 3 weeks ago
- ☆12Jan 16, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- SLMTokBench for paper "SpeechTokenizer: Unified Speech Tokenizer for Speech Large Language Models"☆37Aug 29, 2023Updated 3 years ago
- ☆12Mar 24, 2024Updated 2 years ago
- NOTSOFAR-1 Challenge: Distant Diarization and ASR☆66Feb 12, 2025Updated last year
- A spoken version of the textual story cloze benchmark☆22Aug 6, 2023Updated 3 years ago
- Metadata and versioning details for the Common Voice dataset☆175Jun 16, 2026Updated 2 months ago
- ☆17Updated this week
- ⚡ Blazing fast audio augmentation in Python, powered by GPU for high-efficiency processing in machine learning and audio analysis tasks.☆38May 8, 2026Updated 3 months ago
- Julia package for Hidden Markov Model☆33Sep 11, 2023Updated 2 years ago
- Code for DeCoAR (ICASSP 2020) and BERTphone (Odyssey 2020)☆104Nov 26, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A fast and lightweight python-based CTC beam search decoder for speech recognition.☆468Jul 13, 2023Updated 3 years ago
- Command line utility for forced alignment using Kaldi☆1,879Aug 20, 2026Updated 2 weeks ago
- PESQ (Perceptual Evaluation of Speech Quality) Wrapper for Python Users (narrow band and wide band)☆630Mar 18, 2026Updated 5 months ago
- Download AudioSet for Vision-Audio-Text Pre-training☆13May 16, 2022Updated 4 years ago
- This is unofficial repository for Towards Efficient and Scalable Sharpness-Aware Minimization.☆37Apr 15, 2024Updated 2 years ago
- [NAACL 2025] WaveFM: A High-Fidelity and Efficient Vocoder Based on Flow Matching☆133Apr 8, 2026Updated 4 months ago
- Image and video processing toolbox☆10Jun 12, 2020Updated 6 years ago