End-to-End Speech Recognition using Neural Networks.
☆35Aug 30, 2024Updated 2 years ago
Alternatives and similar repositories for Speech-Recognition
Users that are interested in Speech-Recognition are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Apr 18, 2019Updated 7 years ago
- Predicting Political Instability and Social Conflicts Using Multimodal Data☆10Jun 6, 2016Updated 10 years ago
- acnn for text-independent speaker recognition☆10Feb 8, 2022Updated 4 years ago
- Template and steps to build your personal blog using Jekyll and Minimal Mistake☆10Feb 24, 2020Updated 6 years ago
- ☆10Dec 28, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- an app that makes your personalized newsletter based on your bookmarks☆10Oct 15, 2017Updated 8 years ago
- Speech signal processing, supra-segmental phonology, prosody☆10Jul 24, 2018Updated 8 years ago
- This repo is for a fun side project to build a personalized restaurant recommender using LightFM☆12Mar 31, 2019Updated 7 years ago
- ☆13Nov 15, 2024Updated last year
- Speaker Diarization using GRU in PyTorch☆11Aug 29, 2020Updated 6 years ago
- https://www.kaggle.com/c/tensorflow-speech-recognition-challenge/☆21Mar 1, 2018Updated 8 years ago
- A mini, simple, and fast end-to-end automatic speech recognition toolkit.☆53Dec 6, 2022Updated 3 years ago
- [AAAI2021] A repository of Contrastive Adversarial Learning for Person-independent FER☆16Jan 4, 2022Updated 4 years ago
- GBDT结合LR的二分类模型,封装成了一个类。scikit-learn风格,可以fit和predict。有run_demo☆11Sep 5, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository contains code for a tutorial on end to end automatic speech recognition.☆18Sep 10, 2019Updated 7 years ago
- The project tries to solve a speaker diarization problem using audio features, face recognition and video feature extraction from face im…☆16Feb 10, 2019Updated 7 years ago
- Efficient Speech Processing Tookit for Automatic Speaker Recognition☆18Feb 8, 2023Updated 3 years ago
- Pytorch implementation of DeepLab V3+☆13Apr 13, 2019Updated 7 years ago
- Automatic Speech Recognition using Tensorflow☆46Aug 9, 2017Updated 9 years ago
- Official implementation of Character Region Awareness for Text Detection (CRAFT)☆15Jan 14, 2025Updated last year
- reproduction of the CVPR'21 paper Distilling Knowledge via Knowledge Review for the ML Reproducibility Challenge 2021☆11Apr 16, 2022Updated 4 years ago
- [ICCV'21] The Right to Talk: An Audio-Visual Transformer Approach☆20Aug 2, 2021Updated 5 years ago
- A PyTorch implementation of Speech Transformer with multi-GPUs, an End-to-End ASR with Transformer network on Mandarin Chinese. This code…☆10Dec 25, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official PyTorch code for "Vector Quantization Prompting for Continual Learning (NeurIPS2024)".☆11Oct 16, 2024Updated last year
- ☆13Apr 17, 2018Updated 8 years ago
- ☆12Aug 5, 2022Updated 4 years ago
- Efficient Pose Machine for Multi-Person Pose Estimation☆15Dec 20, 2019Updated 6 years ago
- YOLOv3 in pytorch, trained on Pascal VOC 2007 dataset☆10May 16, 2020Updated 6 years ago
- A High-Quality and Large-Scale Dataset for English-Vietnamese Speech Translation (INTERSPEECH 2022)☆26Jun 5, 2025Updated last year
- On-the-fly Definition Augmentation of LLMs for Biomedical NER☆14Apr 14, 2025Updated last year
- ☆14Sep 7, 2022Updated 4 years ago
- ConvNCF (Convolutional Neural Collaborative Filtering) model implement in pytorch☆13Jul 10, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository follows papers and reports on discrete speech representation learning and speech tokenization methods for speech language…☆15Dec 1, 2023Updated 2 years ago
- ☆12Apr 18, 2021Updated 5 years ago
- DenseCL + regionCL-D☆15Dec 27, 2022Updated 3 years ago
- ☆12May 22, 2023Updated 3 years ago
- ☆16Dec 14, 2022Updated 3 years ago
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- TDY-CNN for text-independent speaker verification☆19Nov 7, 2022Updated 3 years ago