PyTorch implementation of Sequence Transduction with Recurrent Neural Networks (RNN-T) speech recognition paper
☆16Mar 4, 2022Updated 4 years ago
Alternatives and similar repositories for RNN-Transducer
Users that are interested in RNN-Transducer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ASR, End-to-End, end2end, Speech Recognition, 端到端语音识别☆12Oct 25, 2020Updated 5 years ago
- Example code for a neural transducer model.☆68Feb 10, 2024Updated 2 years ago
- Speech command recognition with capsule network & various NNs / KWS on Google Speech Command Dataset.☆25Jan 28, 2019Updated 7 years ago
- Context-Sensitive Neural Spelling Checker☆20Sep 25, 2024Updated last year
- The first Dialectal Arabic Code Switching - DACS corpus from broadcast speech. Annotated at the token-level, considering both the linguis…☆15Apr 3, 2022Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Source code of paper "Adapting pretrained speech model for Mandarin lyrics transcription and alignment"☆19Dec 14, 2023Updated 2 years ago
- An implementation for "Conformer: Convolution-augmented Transformer for Speech Recognition" Paper☆20Aug 16, 2022Updated 4 years ago
- This repository contains a demonstrative implementation for pooling-based models, e.g., DeepPyramidion complementing our paper "Sparsifyi…☆14May 15, 2022Updated 4 years ago
- PyTorch implementation of Listen, Attend and Spell (LAS) speech recognition paper☆12Mar 4, 2022Updated 4 years ago
- Attention-based model for keywords spotting☆19Aug 9, 2021Updated 5 years ago
- Mining effective negative training samples for keyword spotting (PyTorch)☆67May 23, 2020Updated 6 years ago
- ☆16Jul 4, 2024Updated 2 years ago
- Fine-tuning Wav2Vec2.0 on Common Voice(zh-HK)☆16May 8, 2022Updated 4 years ago
- ☆12Aug 9, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch Implementations for End-to-End Automatic Speech Recognition☆127Jun 10, 2019Updated 7 years ago
- ☆12Feb 9, 2021Updated 5 years ago
- A Pytorch Implementation of Transducer Model for End-to-End Speech Recognition☆238May 12, 2020Updated 6 years ago
- ☆16Aug 1, 2025Updated last year
- Implementation of the paper "Confidence estimation for attention based sequence to sequence models for speech recognition"☆16May 9, 2021Updated 5 years ago
- LSTM CTC End2End Speech Recognition.☆38Apr 2, 2019Updated 7 years ago
- The project is related to the development of labs for the ITMO Speaker Recognition Course.☆16Jul 3, 2026Updated 2 months ago
- 🎯 Speech Recognition Challenge by Speech Lab - IIT Madras☆10Nov 5, 2020Updated 5 years ago
- ☆10Apr 2, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ASR project with pytorch-lightning☆20Mar 21, 2025Updated last year
- Working online speech recognition based on RNN Transducer. ( Trained model release available in release )☆292Aug 5, 2021Updated 5 years ago
- Electrophysiology practicals for undergraduate students☆13Mar 8, 2021Updated 5 years ago
- Deep Neural Networks for audio classification☆10Apr 11, 2024Updated 2 years ago
- Conformer RNN-Transducer☆14May 25, 2022Updated 4 years ago
- ☆17Nov 30, 2021Updated 4 years ago
- [arXiv 2024] PyTorch implementation of RRD: https://arxiv.org/abs/2407.12073☆15Dec 2, 2025Updated 9 months ago
- EC499: Major Project☆11Jun 25, 2023Updated 3 years ago
- This is my CS 763 Computer Vision Course Project , Here we try to label Amazon Satelite Images. Here we try to implement the Show and Tel…☆12May 10, 2018Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Simple python script for downloading AVSpeech Dataset☆47Mar 16, 2024Updated 2 years ago
- Unofficial PyTorch implementation of "Keyword Transformer: A Self-Attention Model for Keyword Spotting", Berg et al. 2021.☆41Oct 11, 2022Updated 3 years ago
- The implementation for "Large Language Model Can Transcribe Speech in Multi-Talker Scenarios with Versatile Instructions"☆51Apr 7, 2025Updated last year
- Code for paper "Dual-Path Style Learning for End-to-End Noise-Robust Speech Recognition"☆44May 23, 2023Updated 3 years ago
- PyTorch implementation of "Transformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss" (ICASS…☆114Feb 27, 2022Updated 4 years ago
- Enable RNNLM lattice rescoring with Pytorch [kaldi]☆12Jun 5, 2020Updated 6 years ago
- What part of a song is better at determining it's music genre - the music (audio features) or the lyrics (NLP) ?☆14Jan 2, 2023Updated 3 years ago