Real-Time ASR with CNN-BiLSTM: End-to-End Live Streaming Using PyTorch Lightning⚡
☆11Jan 23, 2025Updated last year
Alternatives and similar repositories for Automatic-Speech-Recognition-with-PyTorch
Users that are interested in Automatic-Speech-Recognition-with-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toolkit for researchers in the multimodal sound separation.☆16Oct 20, 2023Updated 2 years ago
- Official implementation of TISDiSS, a scalable framework for discriminative source separation.☆16Oct 19, 2025Updated 9 months ago
- SepPrune: Structured Pruning for Efficient Deep Speech Separation-AAAI'26☆15May 31, 2025Updated last year
- ☆18Nov 27, 2024Updated last year
- ☆18Jun 16, 2026Updated last month
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICON 2020] TensorFlow Code for "End-to-End Automatic Speech Recognition System for Gujarati"☆13Jul 26, 2021Updated 5 years ago
- ☆24Jul 16, 2025Updated last year
- Official implementation of the paper "Distilling a Pretrained Language Model to a Multilingual ASR Model" (Interspeech 2022)☆12Mar 12, 2024Updated 2 years ago
- ☆17Mar 27, 2026Updated 4 months ago
- Blood Cell Detection and Classification using CNN and LSTMs☆11Jan 12, 2021Updated 5 years ago
- CleanUMamba: A Compact Mamba Network for Speech Denoising using Channel Pruning [Official PyTorch implementation]☆28Jun 12, 2025Updated last year
- It is a specialization course of Python in coursera hosted by **University of Michigan**. This repository contains the solutions of the …☆20Jun 9, 2020Updated 6 years ago
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)☆61Nov 1, 2024Updated last year
- Audio Preprocessing and finetuning of wav2vec2-large-xlsr model on AI4D Baamtu Datamation - Automatic Speech Recognition in WOLOF Data.☆18Nov 13, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆26Mar 31, 2026Updated 3 months ago
- Power-Guided Grouped SRU for Real-Time Causal Audio-Visual Speech Separation☆26Jul 20, 2026Updated last week
- From Python basics to Machine Learning and PyTorch Deep Learning - one day at a time, explore it all☆10May 25, 2025Updated last year
- Detecting Broken Glass Insulators for Automated UAV Power Line Inspection Based on an Improved YOLOv8 Model☆22Sep 22, 2024Updated last year
- Sound Separation, Omni modal☆29Sep 15, 2025Updated 10 months ago
- 🌼 Daisy-TTS: Simulating Wider Spectrum of Emotions via Prosody Embedding Decomposition☆14Nov 15, 2025Updated 8 months ago
- ASLP Summer Inter@NPU☆13Jul 30, 2024Updated last year
- AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. I…☆11Nov 21, 2023Updated 2 years ago
- wsj0-{2, 3, 4, 5} mix generation scripts, in Python.☆79Mar 17, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- kNN-SVC: Robust Zero-Shot Singing Voice Conversion with Additive Synthesis and Concatenation Smoothness Optimization☆16Nov 7, 2025Updated 8 months ago
- Model configurations for scaling SE models in the paper "Beyond Performance Plateaus: A Comprehensive Study on Scalability in Speech Enha…☆41Aug 7, 2024Updated last year
- Keypoint Detection Using Detectron 2 on Custom Dataset☆24Sep 10, 2021Updated 4 years ago
- Efficient voice activity detection algorithm using long-term spectral flatness measurement☆15Feb 21, 2017Updated 9 years ago
- Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, int…☆32Updated this week
- The implementation of "End-to-End Neural Speaker Diarization with an Iterative Adaptive Attractor Estimation", which is accepted by Neura…☆11Aug 27, 2023Updated 2 years ago
- ☆35Feb 19, 2025Updated last year
- Official Repository of Smule Renaissance, Smule's Vocal Restoration Models☆43Oct 27, 2025Updated 9 months ago
- ☆16Jul 14, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code repository for paper "DAS-N2N: Machine learning Distributed Acoustic Sensing (DAS) signal denoising without clean data" (https://arx…☆43Jan 23, 2025Updated last year
- Official Repository for "Efficient Vocal Source Separation Through Windowed RoFormer"☆45Oct 30, 2025Updated 8 months ago
- Both audio-only and audio-visual speaker diarization datasets are listed here.☆16Feb 22, 2023Updated 3 years ago
- # TurboQuant v3 (INT4 + AWQ + Protected Channels + Low-Rank) This notebook demonstrates a **TurboQuant-like** quantization algorithm: - G…☆18Mar 29, 2026Updated 4 months ago
- The power-law compressed phase-aware asymmetric (PLCPA-ASYM) loss☆15Sep 4, 2023Updated 2 years ago
- Apply Score diffusion to improve speech signals recorded under various adverse conditions and distortions, including noise, reverberation…☆83Jul 29, 2024Updated 2 years ago
- A high-performance C/C++ inference server for Qwen3-ASR , optimized for CPU/GPU real-time streaming speech recognition.☆15Jun 27, 2026Updated last month