Automated Lip Reading using Deep Reinforcement Learning
☆34Jun 24, 2018Updated 8 years ago
Alternatives and similar repositories for lips-reading
Users that are interested in lips-reading are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- My experiments in lip reading using deep learning with the LRW dataset☆54Mar 14, 2021Updated 5 years ago
- End-to-end pipeline for lip reading at the word level using a tensorflow CNN implementation.☆35Feb 15, 2020Updated 6 years ago
- Using an LSTM and 4d convolutional network for lip reading☆12May 11, 2018Updated 8 years ago
- ☆64Oct 8, 2018Updated 7 years ago
- Automated Lip reading from real-time videos in tensorflow in python☆162Mar 20, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Audio-Visual Speech Recognition using Deep Learning☆61Nov 14, 2018Updated 7 years ago
- Speech Recognition without audio input☆144May 5, 2026Updated 4 months ago
- Lip Reading in the Wild using ResNet and LSTMs in PyTorch☆57Apr 23, 2018Updated 8 years ago
- demo code for lip reading☆21Dec 9, 2016Updated 9 years ago
- Code and models for evaluating a state-of-the-art lip reading network☆196Mar 24, 2023Updated 3 years ago
- "LipNet: End-to-End Sentence-level Lipreading" in PyTorch☆70Sep 9, 2019Updated 7 years ago
- Pytorch code for End-to-End Audiovisual Speech Recognition☆182Nov 18, 2022Updated 3 years ago
- CNN for visual speech recognition☆23Dec 5, 2016Updated 9 years ago
- ☆14Feb 5, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Deep Visual Speech Recognition in arabic words☆17Oct 18, 2023Updated 2 years ago
- Matlab offscreen rendering toolbox.☆16Feb 23, 2015Updated 11 years ago
- The implementation of 'Watch, Listen, Attend and Spell’ (WLAS) network that learns to transcribe videos of mouth motion to character on p…☆11Mar 23, 2018Updated 8 years ago
- SQL Tutorials using Jupyter Notebook☆17Apr 9, 2023Updated 3 years ago
- Code for Fast Bird Part Localization (FGVC 2015)☆21Apr 19, 2016Updated 10 years ago
- A Question Generation Application leveraging RAG and Weaviate vector store to be able to retrieve relative contexts and generate a more u…☆17Feb 3, 2025Updated last year
- Facial-Expression Recognition with Deep Neural Networks☆10Mar 6, 2016Updated 10 years ago
- A tool to collect/validate audio recordings from workers on Amazon Mechanical Turk. Written in Python/Flask. (originally hosted on github…☆15Dec 19, 2022Updated 3 years ago
- Octave port of the Fast Image Source Model by Eric A. Lehmann. Used for room acoustic modeling and impulse response simulation.☆12Aug 2, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- GPUI with background web sockets☆17Feb 25, 2024Updated 2 years ago
- ICASSP'22 Training Strategies for Improved Lip-Reading; ICASSP'21 Towards Practical Lipreading with Distilled and Efficient Models; ICASS…☆438May 18, 2023Updated 3 years ago
- #DNN #CNN #LSTM #Classification #Sequential_data #Lip_reading☆28Jun 3, 2018Updated 8 years ago
- ☆10Sep 19, 2022Updated 4 years ago
- Chinese words classification using lipnet with pytorch☆40Nov 18, 2019Updated 6 years ago
- A simple streaming framework for Python☆15Feb 27, 2018Updated 8 years ago
- Lip Reading - Cross Audio-Visual Recognition using 3D Architectures☆1,902Nov 7, 2022Updated 3 years ago
- Building rust and protobufs with bazel.☆18Jul 16, 2023Updated 3 years ago
- A PyTorch implementation of the Deep Audio-Visual Speech Recognition paper.☆244Feb 15, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Aranizer: A Custom Tokenizer based on SentencePiece and BPE tailored for Arabic Language Modeling☆22Aug 4, 2024Updated 2 years ago
- Voice Activity Detection: In this first assignment, we will create a dataset that simulates speech in every-day scenarios. We train a cla…☆18May 3, 2015Updated 11 years ago
- Visual Speech Recognition for Multiple Languages☆483Aug 17, 2023Updated 3 years ago
- The code of '3D-Aware Semantic-Guided Generative Model for Human Synthesis' (ECCV 2022)☆36Jul 18, 2022Updated 4 years ago
- create timer videos at any speed.☆15Sep 25, 2023Updated 3 years ago
- Gait recognition system based on YOLOv8☆16Jan 26, 2024Updated 2 years ago
- ODAS: Open embeddeD Audition System☆11Mar 20, 2021Updated 5 years ago