Chinese words classification using lipnet with pytorch
☆40Nov 18, 2019Updated 6 years ago
Alternatives and similar repositories for LipNet_ChineseWordsClassification
Users that are interested in LipNet_ChineseWordsClassification are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 2019年“创青春·交子杯”新网银行高校金融科技挑战赛初赛、决赛思路代码分享☆28Dec 11, 2019Updated 6 years ago
- 2019年“创青春.交子杯”新网银行高校金融科技挑战赛-AI算法赛道比赛_代码分享☆89Jul 15, 2020Updated 6 years ago
- The state-of-art PyTorch implementation of the method described in the paper "LipNet: End-to-End Sentence-level Lipreading" (https://arxi…☆238Sep 21, 2022Updated 3 years ago
- "LipNet: End-to-End Sentence-level Lipreading" in PyTorch☆70Sep 9, 2019Updated 6 years ago
- ☆11May 31, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Huggingface Implementation of AV-HuBERT on the MuAViC Dataset☆19Mar 6, 2025Updated last year
- Official code for the paper "Scaling Multilingual Visual Speech Recognition"☆20Aug 15, 2025Updated 11 months ago
- An OpenCV demo on detecting whether a person is speaking or not.☆23Mar 21, 2012Updated 14 years ago
- lip_reading_demo_net☆32Oct 22, 2019Updated 6 years ago
- ☆64Oct 8, 2018Updated 7 years ago
- Keras implementation of 'LipNet: End-to-End Sentence-level Lipreading'☆691Nov 22, 2022Updated 3 years ago
- Visual Speech Recognition For Low-Resource Languages with Automatic Labels (ICASSP 2024)☆17Mar 17, 2025Updated last year
- Automated Lip Reading using Deep Reinforcement Learning☆33Jun 24, 2018Updated 8 years ago
- Audio-Visual Speech Recognition using Deep Learning☆61Nov 14, 2018Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official Pytorch implementation of "Large Language Models are Strong Audio-Visual Speech Recognition Learners" [ICASSP 2025] and "Mitigat…☆64Jan 18, 2026Updated 6 months ago
- PyTorch implementation of Human Action Recognition Based on Spatial-Temporal Attention at ICLR 2019☆14Dec 12, 2018Updated 7 years ago
- Code and models for evaluating a state-of-the-art lip reading network☆197Mar 24, 2023Updated 3 years ago
- Automated Lip reading from real-time videos in tensorflow in python☆162Mar 20, 2018Updated 8 years ago
- Use human pose information to help action recognition, explored with attention-pooling method, C3D method and two-stream architecture, im…☆18Jun 7, 2018Updated 8 years ago
- A version of Obamanet that you won't go insane setting up.☆17Nov 21, 2022Updated 3 years ago
- The code for AAAI 2025 “Large Language Models Are Read/Write Policy-Makers for Simultaneous Generation”☆15Jan 3, 2025Updated last year
- ☆12Sep 19, 2021Updated 4 years ago
- ☆12Sep 1, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Building Pytorch Server with Flask☆31Mar 12, 2018Updated 8 years ago
- Rewrite the cmakefile to install and run it on Ubuntu☆15Sep 11, 2024Updated last year
- [CVPR'26] MetroGS: Efficient and Stable Reconstruction of Geometrically Accurate High-Fidelity Large-Scale Scenes☆20Apr 15, 2026Updated 3 months ago
- The speaker-labeled information of LRW dataset, which is the outcome of the paper "Speaker-adaptive Lip Reading with User-dependent Paddi…☆10Oct 12, 2023Updated 2 years ago
- Temporal Denoising Mask Synthesis Network for Learning Blind Video Temporal Consistency☆32Jun 6, 2021Updated 5 years ago
- Code for Self-and-Collaborative Attention Network from "SCAN: Self-and-Collaborative Attention Network for Video Person Re-identification…☆26Jun 1, 2019Updated 7 years ago
- ☆18May 6, 2019Updated 7 years ago
- A Pytorch (support batch and channel) implementation of GoogleBrain's SpecAugment: A Simple Data Augmentation Method for Automatic Speech…☆11Jul 24, 2024Updated 2 years ago
- ☆12Sep 14, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Deep Variational Information Bottleneck (DVIB) in PyTorch.☆10Apr 25, 2020Updated 6 years ago
- Companion code for Awe the Audience: How the Narrative Trajectories Affect Audience Perception in Public Speaking☆14Jan 6, 2018Updated 8 years ago
- ☆14Jul 27, 2022Updated 4 years ago
- Training code for the ACAM action detection model.☆29Feb 2, 2023Updated 3 years ago
- A structured parsing technique for NER☆15May 26, 2023Updated 3 years ago
- SpringBoot学习系列☆24Aug 26, 2024Updated last year
- QuickSplat: Fast 3D Surface Reconstruction via Learned Gaussian Initialization☆25Nov 11, 2025Updated 9 months ago