☆15Dec 11, 2021Updated 4 years ago
Alternatives and similar repositories for Deformation-Flow-Based-Two-stream-Network-for-Lip-Reading
Users that are interested in Deformation-Flow-Based-Two-stream-Network-for-Lip-Reading are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code and model for paper <Mutual Information Maximization for Effective Lip Reading>☆19Sep 4, 2020Updated 5 years ago
- PyTorch implementation of "Multi-modality Associative Bridging through Memory: Speech Sound Recollected from Face Video" (ICCV2021)☆22Apr 11, 2022Updated 4 years ago
- The PyTorch Code and Model In "Learn an Effective Lip Reading Model without Pains", (https://arxiv.org/abs/2011.07557), which reaches the…☆169Sep 12, 2025Updated 11 months ago
- PyTorch implementation of "Distinguishing Homophenes using Multi-Head Visual-Audio Memory" (AAAI2022)☆27Mar 9, 2024Updated 2 years ago
- Pytorch code for End-to-End Audiovisual Speech Recognition☆182Nov 18, 2022Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Visual speech recognition with face inputs: code and models for F&G 2020 paper "Can We Read Speech Beyond the Lips? Rethinking RoI Select…☆19Apr 12, 2021Updated 5 years ago
- This dataset is presented in the paper Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel's Weekly Video…☆12Sep 21, 2022Updated 3 years ago
- TF code for our CVPR2020 paper "Discriminative Multi-modality Speech Recognition"☆26Apr 27, 2022Updated 4 years ago
- Official Implementation of Visual Transformer Pooling for Lip reading☆42Aug 8, 2022Updated 4 years ago
- Collection of works from VIPL-AVSU☆50Jul 21, 2026Updated 3 weeks ago
- The official implementation of OpenSR (ACL2023 Oral)☆17Nov 29, 2023Updated 2 years ago
- Official project of DiverseSampling (ACMMM2022 Paper)☆16Feb 25, 2023Updated 3 years ago
- ☆24Mar 30, 2024Updated 2 years ago
- [ICCV2023] Spatio-temporal Prompting Network for Robust Video Feature Extraction☆11Aug 17, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DenseNet3D Model In "LRW-1000: A Naturally-Distributed Large-Scale Benchmark for Lip Reading in the Wild", https://arxiv.org/abs/1810.069…☆123Mar 13, 2026Updated 5 months ago
- 2019年“创青春·交子杯”新网银行高校金融科技挑战赛初赛、决赛思路代码分享☆28Dec 11, 2019Updated 6 years ago
- End to End Multiview Lip Reading☆10Jan 26, 2018Updated 8 years ago
- Official PyTorch implementation of paper Leveraging Unimodal Self Supervised Learning for Multimodal Audio-Visual Speech Recognition (ACL…☆67Jul 13, 2022Updated 4 years ago
- ICASSP'22 Training Strategies for Improved Lip-Reading; ICASSP'21 Towards Practical Lipreading with Distilled and Efficient Models; ICASS…☆438May 18, 2023Updated 3 years ago
- Official PyTorch Implementation of paper EAN: Event Adaptive Network for Efficient Action Recognition https://arxiv.org/abs/2107.10771☆33Oct 24, 2023Updated 2 years ago
- Code for our CICAI 2022 paper "3D Face Cartoonizer: Generating Personalized 3D Cartoon Faces from 2D Real Photos with a Hybrid Dataset".☆10Aug 9, 2022Updated 4 years ago
- Panoramic audiovisual salient object segmentation☆30Jul 9, 2023Updated 3 years ago
- Official implementation of RAVEn (ICLR 2023) and BRAVEn (ICASSP 2024)☆82Feb 27, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- In this notebook, I am updating NLP notebooks, and projects☆10Jun 29, 2023Updated 3 years ago
- [ACM MM 2023] KeyPosS: Plug-and-Play Facial Landmark Detection through GPS-Inspired True-Range Multilateration☆12Nov 21, 2023Updated 2 years ago
- Transfer learning approach to pronunciation scoring☆12Jan 17, 2024Updated 2 years ago
- ☆16Sep 20, 2022Updated 3 years ago
- Pytorch implementation of deep fill v2 (original by Jiayu et al.)☆10Jun 26, 2019Updated 7 years ago
- ☆11Nov 9, 2023Updated 2 years ago
- Self-supervised Siamese network (SSiam), FG 2019☆27Apr 21, 2023Updated 3 years ago
- Semi-Supervised Contrastive Learning for music classification - towards HIL-representation learning.☆17Jul 24, 2024Updated 2 years ago
- ☆15Aug 21, 2020Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Face de-occlusion using 3D morphable model and generative adversarial network☆34Oct 22, 2021Updated 4 years ago
- ☆19Jun 22, 2026Updated last month
- [AAAI 2022] DCAN: Improving Temporal Action Detection via Dual Context Aggregation☆17Nov 13, 2022Updated 3 years ago
- Pytorch implementation and comparison of Fourier Feature Networks and Sinusoidal Representation Networks☆13Jun 27, 2020Updated 6 years ago
- Visual Speech Recognition for Multiple Languages☆480Aug 17, 2023Updated 3 years ago
- This is an OCR program designed for travel document. It can now support 23 types of documents with pre-defined template. You can add what…☆10Nov 22, 2022Updated 3 years ago
- High-resolution facial landmark detection in artworks☆23Dec 17, 2023Updated 2 years ago