Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡
☆11Jan 23, 2025Updated last year
Alternatives and similar repositories for Automatic-Speech-Recognition-with-PyTorch
Users that are interested in Automatic-Speech-Recognition-with-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- Official implementation of TISDiSS, a scalable framework for discriminative source separation.☆16Jul 31, 2026Updated last month
- SepPrune: Structured Pruning for Efficient Deep Speech Separation-AAAI'26☆15May 31, 2025Updated last year
- ☆18Nov 27, 2024Updated last year
- ☆20Jun 16, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [ICON 2020] TensorFlow Code for "End-to-End Automatic Speech Recognition System for Gujarati"☆13Jul 26, 2021Updated 5 years ago
- ☆24Jul 16, 2025Updated last year
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- ☆17Mar 27, 2026Updated 5 months ago
- Blood Cell Detection and Classification using CNN and LSTMs☆11Jan 12, 2021Updated 5 years ago
- CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning [Official PyTorch implementation]☆29Jun 12, 2025Updated last year
- It is a specialization course of Python in coursera hosted by **University of Michigan**. This repository contains the solutions of the …☆20Jun 9, 2020Updated 6 years ago
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)☆61Nov 1, 2024Updated last year
- Audio Preprocessing and finetuning of wav2vec2-large-xlsr model on AI4D Baamtu Datamation - Automatic Speech Recognition in WOLOF Data.☆18Nov 13, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆27Mar 31, 2026Updated 5 months ago
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆31Jul 20, 2026Updated last month
- From Python basics to Machine Learning and PyTorch Deep Learning - one day at a time, explore it all☆10May 25, 2025Updated last year
- Detecting Broken Glass Insulators for Automated UAV Power Line Inspection Based on an Improved YOLOv8 Model☆22Sep 22, 2024Updated last year
- Sound Separation, Omni modal☆30Sep 15, 2025Updated 11 months ago
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 9 months ago
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago
- wsj0-{2, 3, 4, 5} mix generation scripts, in Python.☆79Mar 17, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- kNN-SVC: Robust Zero-Shot Singing Voice Conversion with Additive Synthesis and Concatenation Smoothness Optimization☆16Nov 7, 2025Updated 10 months ago
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆42Aug 7, 2024Updated 2 years ago
- Keypoint Detection Using Detectron 2 on Custom Dataset☆24Sep 10, 2021Updated 4 years ago
- Efficient voice activity detection algorithm using long-term spectral flatness measurement☆15Feb 21, 2017Updated 9 years ago
- Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, int…☆36Updated this week
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 3 years ago
- ☆36Feb 19, 2025Updated last year
- Official Repository of Smule Renaissance, Smule's Vocal Restoration Models☆43Oct 27, 2025Updated 10 months ago
- ☆16Jul 14, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code repository for paper "DAS-N2N: Machine learning Distributed Acoustic Sensing (DAS) signal denoising without clean data" (https://arx…☆43Jan 23, 2025Updated last year
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆46Oct 30, 2025Updated 10 months ago
- Both audio-only and audio-visual speaker diarization datasets are listed here.☆16Feb 22, 2023Updated 3 years ago
- # TurboQuant v3 (INT4 + AWQ + Protected Channels + Low-Rank) This notebook demonstrates a **TurboQuant-like** quantization algorithm: - G…☆17Mar 29, 2026Updated 5 months ago
- The power-law compressed phase-aware asymmetric (PLCPA-ASYM) loss☆15Sep 4, 2023Updated 3 years ago
- Apply Score diffusion to improve speech signals recorded under various adverse conditions and distortions, including noise, reverberation…☆83Jul 29, 2024Updated 2 years ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆16Jun 27, 2026Updated 2 months ago