The proposed method in LRW-1000: A Naturally-Distributed Large-Scale Benchmark for Lip Reading in the Wild
☆26Nov 23, 2018Updated 7 years ago
Alternatives and similar repositories for D3D
Users that are interested in D3D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code and model for paper <Mutual Information Maximization for Effective Lip Reading>☆19Sep 4, 2020Updated 5 years ago
- Pytorch code for End-to-End Audiovisual Speech Recognition☆182Nov 18, 2022Updated 3 years ago
- "LipNet: End-to-End Sentence-level Lipreading" in PyTorch☆70Sep 9, 2019Updated 6 years ago
- The PyTorch Code and Model In "Learn an Effective Lip Reading Model without Pains", (https://arxiv.org/abs/2011.07557), which reaches the…☆169Sep 12, 2025Updated 11 months ago
- ☆15Dec 11, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Finalist entry for the M2CAI Workflow Challenge 2016☆10Nov 25, 2016Updated 9 years ago
- Audio-Visual Speech Recognition using Deep Learning☆61Nov 14, 2018Updated 7 years ago
- Torch code for using Residual Networks with LSTMs for Lipreading☆99Oct 8, 2018Updated 7 years ago
- This dataset is presented in the paper Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel's Weekly Video…☆12Sep 21, 2022Updated 3 years ago
- Visual speech recognition with face inputs: code and models for F&G 2020 paper "Can We Read Speech Beyond the Lips? Rethinking RoI Select…☆19Apr 12, 2021Updated 5 years ago
- processing and extracting of face and mouth image files out of the TCDTIMIT database☆47Sep 22, 2020Updated 5 years ago
- 2019年“创青春·交子杯”新网银行高校金融科技挑战赛初赛、决赛思路代码分享☆28Dec 11, 2019Updated 6 years ago
- Active appearance model toolbox☆14Nov 2, 2015Updated 10 years ago
- Automated Lip Reading using Deep Reinforcement Learning☆34Jun 24, 2018Updated 8 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Lip Reading in the Wild using ResNet and LSTMs in PyTorch☆57Apr 23, 2018Updated 8 years ago
- ☆11Sep 16, 2014Updated 11 years ago
- [IROS 2024 Oral Pitch] PyTorch Implementation of "Dual-Branch Graph Transformer Network for 3D Human Mesh Reconstruction from Video"☆15Jul 19, 2024Updated 2 years ago
- An ugly tool for labeling segmentations given images and the corresponding superpixels.☆14May 18, 2016Updated 10 years ago
- Using an LSTM and 4d convolutional network for lip reading☆12May 11, 2018Updated 8 years ago
- ☆19Jul 14, 2019Updated 7 years ago
- ☆16Apr 20, 2020Updated 6 years ago
- sk-cnn is proposed in Skeleton based action recognition with convolutional neural network(PR 2016). Here implemented in Keras☆19Apr 10, 2018Updated 8 years ago
- pytorch implementation of SOSELETO☆15Sep 5, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACM MM 2024] PyTorch Implementation of "ARTS: Semi-Analytical Regressor using Disentangled Skeletal Representations for Human Mesh Recov…☆16Feb 27, 2025Updated last year
- The code used to create the ARCA23K and ARCA23K-FSD datasets☆16Nov 9, 2021Updated 4 years ago
- Portal of Johannes and Felix's RNN implementation and further modifications for ASR☆21Nov 27, 2014Updated 11 years ago
- The 1st place solution for AutoSpeech 2019.☆17Jun 9, 2020Updated 6 years ago
- PyTorch implementation of Human Action Recognition Based on Spatial-Temporal Attention at ICLR 2019☆14Dec 12, 2018Updated 7 years ago
- Structured Receptive Fields in Convolutional Neural Networks☆48Feb 20, 2018Updated 8 years ago
- ☆12Oct 5, 2022Updated 3 years ago
- ☆13May 10, 2022Updated 4 years ago
- Translating Torch model to other framework such as Caffe, MxNet ...☆22Dec 16, 2016Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Detects lip movement and check if a person is speaking☆19May 4, 2018Updated 8 years ago
- Spatially and Temporally Efficient Non-local Attention Network for Video-based Person Re-Identification (BMVC 2019)☆144Jun 13, 2021Updated 5 years ago
- Keras implementation of 'LipNet: End-to-End Sentence-level Lipreading'☆691Nov 22, 2022Updated 3 years ago
- modified version of src☆17Jan 13, 2018Updated 8 years ago
- Code for our submision on ICCV2017. A fork from https://github.com/rbgirshick/py-faster-rcnn☆21Sep 18, 2017Updated 8 years ago
- Code and dataset release for "PACS: A Dataset for Physical Audiovisual CommonSense Reasoning" (ECCV 2022)☆18Dec 20, 2022Updated 3 years ago
- Use human pose information to help action recognition, explored with attention-pooling method, C3D method and two-stream architecture, im…☆18Jun 7, 2018Updated 8 years ago