A light-weight Python library for computing Kaldi-style acoustic features based on NumPy
☆14Aug 17, 2020Updated 6 years ago
Alternatives and similar repositories for kaldifeat
Users that are interested in kaldifeat are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert kaldi feature extraction and nnet3 models into Tensorflow Lite models. Currently aimed at converting kaldi's x-vector models and …☆20Oct 6, 2022Updated 4 years ago
- Stellenbosch University ZeroSpeech 2019 System☆10Apr 4, 2019Updated 7 years ago
- Audio-only Emotion Detection using Federated Learning☆10Dec 8, 2022Updated 3 years ago
- PyTorch implementation of Densely Connected Time Delay Neural Network☆91May 4, 2023Updated 3 years ago
- Privacy-preserving Voice Analysis via Disentangled Representations☆13Aug 30, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆22Sep 24, 2018Updated 8 years ago
- ☆37May 8, 2021Updated 5 years ago
- Tacotron 2 - PyTorch implementation with faster-than-realtime inference☆30May 28, 2020Updated 6 years ago
- ☆11May 4, 2020Updated 6 years ago
- ☆15Sep 13, 2022Updated 4 years ago
- Entropy Minimization vs. Diversity Maximization for Domain Adaptation☆15Feb 9, 2020Updated 6 years ago
- Voice Framework☆19Jan 21, 2026Updated 8 months ago
- Source code for ICLR 2019 paper☆24Oct 6, 2020Updated 6 years ago
- A set of github actions to build WebRTC as a single static library.☆20Jun 20, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Audio Entailment: Deductive Reasoning for Audio Understanding☆17Dec 10, 2024Updated last year
- This is the official implementation of ClusTR: Clustering Training for Robustness paper.☆21Oct 20, 2021Updated 4 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- MTGAN: Speaker Verification through Multitasking Triplet Generative Adversarial Networks☆19Feb 29, 2020Updated 6 years ago
- Beat and downbeat tracking on symbolic music data☆35May 31, 2022Updated 4 years ago
- Baselines and Classifiers for speaker anti-spoofing detection☆18Jul 25, 2024Updated 2 years ago
- WebRTC Research☆14Jul 15, 2020Updated 6 years ago
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆17May 8, 2025Updated last year
- Implementation of "Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition"☆18Aug 26, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Document Scanner with OCR iOS app written in Swift☆11Nov 22, 2021Updated 4 years ago
- Predicting Biomedical Interactions with Higher-Order Graph Convolutional Networks☆16Nov 9, 2021Updated 4 years ago
- SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY☆16Nov 28, 2024Updated last year
- This is application for dysarthria to improve their pronunciation by using deep learning☆10Dec 29, 2020Updated 5 years ago
- Sina News Crawler and Word Segmentation☆13Dec 20, 2017Updated 8 years ago
- Lightweight multiple sound source localization, based on a triangular microphone array.☆17Dec 5, 2023Updated 2 years ago
- Subband Adaptive System with Crossterms for aliasing reduction☆18Jul 31, 2022Updated 4 years ago
- speex aec kalman filter☆17Mar 17, 2024Updated 2 years ago
- Tencent Cloud Websocket Serverless Component☆12Oct 22, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Webapp for creating interactive pronunciation guides for any English word.☆10Feb 26, 2025Updated last year
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆18Feb 5, 2026Updated 8 months ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement☆16Jul 11, 2025Updated last year
- Computation Graph framework implemented using only NumPy☆11Mar 31, 2024Updated 2 years ago
- An online NLP tool/application which will correct grammar mistakes (like Grammarly) and will also rewrite the sentences in a different fo…☆10Nov 25, 2019Updated 6 years ago
- FFT-based windowed spectrum analyzer☆13Mar 10, 2017Updated 9 years ago