A light-weight Python library for computing Kaldi-style acoustic features based on NumPy
☆14Aug 17, 2020Updated 6 years ago
Alternatives and similar repositories for kaldifeat
Users that are interested in kaldifeat are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert kaldi feature extraction and nnet3 models into Tensorflow Lite models. Currently aimed at converting kaldi's x-vector models and …☆20Oct 6, 2022Updated 3 years ago
- Example python scripts to evaluate various ASR methods☆11Dec 22, 2021Updated 4 years ago
- code for paper "learning to fool the speaker recognition"☆10Jun 12, 2020Updated 6 years ago
- ☆12Jun 17, 2017Updated 9 years ago
- Audio-only Emotion Detection using Federated Learning☆10Dec 8, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch implementation of Densely Connected Time Delay Neural Network☆91May 4, 2023Updated 3 years ago
- Privacy-preserving Voice Analysis via Disentangled Representations☆13Aug 30, 2021Updated 5 years ago
- A recursive forced aligner built on Gentle.☆16Mar 20, 2019Updated 7 years ago
- ☆22Sep 24, 2018Updated 7 years ago
- ☆37May 8, 2021Updated 5 years ago
- Tacotron 2 - PyTorch implementation with faster-than-realtime inference☆30May 28, 2020Updated 6 years ago
- ☆11May 4, 2020Updated 6 years ago
- Belief Revision based Caption Re-ranker with Visual Semantic Information. COLING 2022☆11Apr 13, 2025Updated last year
- ☆15Sep 13, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Source code for ICLR 2019 paper☆24Oct 6, 2020Updated 5 years ago
- Audio Keyword Search☆12May 5, 2019Updated 7 years ago
- This is the official implementation of ClusTR: Clustering Training for Robustness paper.☆21Oct 20, 2021Updated 4 years ago
- PyTorch speech2text inference script for the NVidia openseq2seq wav2letter model variant☆10Aug 12, 2019Updated 7 years ago
- MTGAN: Speaker Verification through Multitasking Triplet Generative Adversarial Networks☆19Feb 29, 2020Updated 6 years ago
- Beat and downbeat tracking on symbolic music data☆35May 31, 2022Updated 4 years ago
- Baselines and Classifiers for speaker anti-spoofing detection☆18Jul 25, 2024Updated 2 years ago
- WebRTC Research☆14Jul 15, 2020Updated 6 years ago
- Subband system identification using generalized Weighted Overlap-Add (WOLA) filter bank for improved acoustic echo cancellation.☆17May 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of "Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition"☆18Aug 26, 2026Updated 3 weeks ago
- Document Scanner with OCR iOS app written in Swift☆11Nov 22, 2021Updated 4 years ago
- Adversarial Auto-encoders for Speech Based Emotion Recogntion☆15Sep 22, 2018Updated 7 years ago
- Predicting Biomedical Interactions with Higher-Order Graph Convolutional Networks☆16Nov 9, 2021Updated 4 years ago
- SALT: STANDARDIZED AUDIO EVENT LABEL TAXONOMY☆16Nov 28, 2024Updated last year
- Sina News Crawler and Word Segmentation☆13Dec 20, 2017Updated 8 years ago
- Lightweight multiple sound source localization, based on a triangular microphone array.☆17Dec 5, 2023Updated 2 years ago
- Subband Adaptive System with Crossterms for aliasing reduction☆18Jul 31, 2022Updated 4 years ago
- speex aec kalman filter☆17Mar 17, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- Webapp for creating interactive pronunciation guides for any English word.☆10Feb 26, 2025Updated last year
- SynTTS-Commands is a large-scale, multilingual (English & Chinese) synthetic speech command dataset designed for low-power Keyword Spotti…☆18Feb 5, 2026Updated 7 months ago
- LLaSE: Maximizing Acoustic Preservation for LLaMA based Speech Enhancement☆16Jul 11, 2025Updated last year
- Computation Graph framework implemented using only NumPy☆11Mar 31, 2024Updated 2 years ago
- An online NLP tool/application which will correct grammar mistakes (like Grammarly) and will also rewrite the sentences in a different fo…☆10Nov 25, 2019Updated 6 years ago
- FFT-based windowed spectrum analyzer☆13Mar 10, 2017Updated 9 years ago