Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡
☆11Jan 23, 2025Updated last year
Alternatives and similar repositories for Automatic-Speech-Recognition-with-PyTorch
Users that are interested in Automatic-Speech-Recognition-with-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- Official implementation of TISDiSS, a scalable framework for discriminative source separation.☆16Jul 31, 2026Updated 2 weeks ago
- SepPrune: Structured Pruning for Efficient Deep Speech Separation-AAAI'26☆15May 31, 2025Updated last year
- ☆18Nov 27, 2024Updated last year
- ☆20Jun 16, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICON 2020] TensorFlow Code for "End-to-End Automatic Speech Recognition System for Gujarati"☆13Jul 26, 2021Updated 5 years ago
- ☆24Jul 16, 2025Updated last year
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- ☆17Mar 27, 2026Updated 4 months ago
- Blood Cell Detection and Classification using CNN and LSTMs☆11Jan 12, 2021Updated 5 years ago
- CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning [Official PyTorch implementation]☆29Jun 12, 2025Updated last year
- It is a specialization course of Python in coursera hosted by **University of Michigan**. This repository contains the solutions of the …☆20Jun 9, 2020Updated 6 years ago
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)☆61Nov 1, 2024Updated last year
- Audio Preprocessing and finetuning of wav2vec2-large-xlsr model on AI4D Baamtu Datamation - Automatic Speech Recognition in WOLOF Data.☆18Nov 13, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆27Mar 31, 2026Updated 4 months ago
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆30Jul 20, 2026Updated 3 weeks ago
- From Python basics to Machine Learning and PyTorch Deep Learning - one day at a time, explore it all☆10May 25, 2025Updated last year
- Detecting Broken Glass Insulators for Automated UAV Power Line Inspection Based on an Improved YOLOv8 Model☆22Sep 22, 2024Updated last year
- Sound Separation, Omni modal☆29Sep 15, 2025Updated 11 months ago
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 9 months ago
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated 2 years ago
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago
- wsj0-{2, 3, 4, 5} mix generation scripts, in Python.☆79Mar 17, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- kNN-SVC: Robust Zero-Shot Singing Voice Conversion with Additive Synthesis and Concatenation Smoothness Optimization☆16Nov 7, 2025Updated 9 months ago
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆42Aug 7, 2024Updated 2 years ago
- Keypoint Detection Using Detectron 2 on Custom Dataset☆24Sep 10, 2021Updated 4 years ago
- Efficient voice activity detection algorithm using long-term spectral flatness measurement☆15Feb 21, 2017Updated 9 years ago
- Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, int…☆34Updated this week
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 2 years ago
- ☆36Feb 19, 2025Updated last year
- Official Repository of Smule Renaissance, Smule's Vocal Restoration Models☆43Oct 27, 2025Updated 9 months ago
- ☆16Jul 14, 2020Updated 6 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code repository for paper "DAS-N2N: Machine learning Distributed Acoustic Sensing (DAS) signal denoising without clean data" (https://arx…☆43Jan 23, 2025Updated last year
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆46Oct 30, 2025Updated 9 months ago
- Both audio-only and audio-visual speaker diarization datasets are listed here.☆16Feb 22, 2023Updated 3 years ago
- # TurboQuant v3 (INT4 + AWQ + Protected Channels + Low-Rank) This notebook demonstrates a **TurboQuant-like** quantization algorithm: - G…☆17Mar 29, 2026Updated 4 months ago
- The power-law compressed phase-aware asymmetric (PLCPA-ASYM) loss☆15Sep 4, 2023Updated 2 years ago
- Apply Score diffusion to improve speech signals recorded under various adverse conditions and distortions, including noise, reverberation…☆83Jul 29, 2024Updated 2 years ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆15Jun 27, 2026Updated last month