sectum1919/cncvs_data_collector

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/sectum1919/cncvs_data_collector)

sectum1919 / cncvs_data_collector

☆27

Alternatives and similar repositories for cncvs_data_collector

Users that are interested in cncvs_data_collector are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

SpringHuo / MAVD
View on GitHub
The MAVD represents Mandarin Audio-Visual dataset with Depth information. MAVD has a rich variety of modal data, including audio, RGB ima…
☆20Apr 22, 2024Updated 2 years ago
DataoceanAI / CNVSRC2023Baseline
View on GitHub
Baseline system for CNVSRC2023 (Chinese Continuous Visual Speech Recognition Challenge 2023)
☆23Apr 27, 2024Updated 2 years ago
liu12366262626 / CNVSRC2025
View on GitHub
Official CNVSRC2025 Competition Baseline
☆17Jun 27, 2025Updated last year
web3aivc / wav2lip_vq
View on GitHub
wav2lip in a Vector Quantized (VQ) space
☆27Jun 20, 2023Updated 3 years ago
leohku / faceformer-emo
View on GitHub
FaceFormer Emo: Speech-Driven 3D Facial Animation with Emotion Embedding
☆27Jul 15, 2023Updated 3 years ago
Managed hosting for WordPress and PHP on Cloudways • Ad
Managed hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
arxrean / LipRead-seq2seq
View on GitHub
An unofficial (PyTorch) implementation for the paper Deep Lip Reading: A comparison of models and an online application.
☆10May 13, 2020Updated 6 years ago
KVDmitrieva / source_sep_hifi
View on GitHub
☆20Jun 29, 2025Updated last year
ex3ndr / supervoice-hybrid
View on GitHub
My hybrid TTS network that combines, VALL-E, VoiceBox, SpeechFlow, Seamless and TortoiseTTS into one
☆26Aug 5, 2024Updated last year
emanuelalvaradog / wav2lip-api
View on GitHub
wav2lip-api
☆11Mar 16, 2023Updated 3 years ago
liu12366262626 / AlignVSR
View on GitHub
Visual Speech Recongnition
☆21Dec 24, 2024Updated last year
YoungSeng / Speech-driven-expressions
View on GitHub
Speech-Driven Expression Blendshape Based on Single-Layer Self-attention Network (AIWIN 2022)
☆77Oct 21, 2022Updated 3 years ago
rogerle / wav2lipup
View on GitHub
optimized wav2lip
☆18Jan 6, 2024Updated 2 years ago
amazon-science / iwslt-autodub-task
View on GitHub
☆21Mar 4, 2024Updated 2 years ago
fclearner / Personal-vad-2.0
View on GitHub
Implementation of "Personal VAD 2.0: Optimizing Personal Voice Activity Detection for On-Device Speech Recognition"
☆16Jun 9, 2026Updated last month
1-Click AI Models by DigitalOcean Gradient • Ad
Deploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
zf223669 / DiffmotionGG-beta
View on GitHub
☆21Dec 9, 2023Updated 2 years ago
JeongHun0716 / Personalized-Lip-Reading
View on GitHub
Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language (AAAI 2025)
☆24Jun 29, 2026Updated 3 weeks ago
aizhiqi-work / OpenKWS
View on GitHub
开源自定义唤醒词
☆17Dec 24, 2025Updated 7 months ago
wuhaozhe / 4d_reconstruction
View on GitHub
☆56Dec 20, 2023Updated 2 years ago
uuembodiedsocialai / ProbTalk3D
View on GitHub
☆104Nov 26, 2025Updated 7 months ago
monk-after-90s / wav2lip_288x288
View on GitHub
☆10Nov 19, 2023Updated 2 years ago
Cocoxili / CMPC
View on GitHub
[IJCAI2022] Unsupervised Voice-Face Representation Learning by Cross-Modal Prototype Contrast
☆21Oct 25, 2023Updated 2 years ago
wujinzhong / Wav2Lip_TensorRT
View on GitHub
☆29Oct 1, 2023Updated 2 years ago
hyx16 / SPMIArray
View on GitHub
Tsinghua University SPMI Lab array processing toolkit
☆18Nov 23, 2016Updated 9 years ago
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
dfki-av / G3FA
View on GitHub
[BMVC'24] G3FA: Geometry-guided GAN for Face Animation
☆20Mar 14, 2025Updated last year
umbertocappellazzo / Llama-AVSR
View on GitHub
Official Pytorch implementation of "Large Language Models are Strong Audio-Visual Speech Recognition Learners" [ICASSP 2025] and "Mitigat…
☆64Jan 18, 2026Updated 6 months ago
FreedomIntelligence / TalkVid
View on GitHub
TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis [CVPR 2026 Findings]
☆193Jun 8, 2026Updated last month
Ego4DSounds / Ego4DSounds
View on GitHub
Ego4DSounds: A diverse egocentric dataset with high action-audio correspondence
☆21Jun 14, 2024Updated 2 years ago
cilkim1 / speech_ani_gan
View on GitHub
An implementation of http://openaccess.thecvf.com/content_CVPRW_2019/papers/Sight%20and%20Sound/Konstantinos_Vougioukas_End-to-End_Speech…
☆18Mar 19, 2020Updated 6 years ago
wonjune-kang / llm-speech-summarization
View on GitHub
Prompting Large Language Models with Audio for General-Purpose Speech Summarization
☆20May 14, 2025Updated last year
madhavlab / 2022_syncnet
View on GitHub
SyncNet for Time Synchronization
☆30Mar 13, 2023Updated 3 years ago
julianyulu / Wav2LipHD
View on GitHub
☆24Oct 8, 2021Updated 4 years ago
my-yy / sl_icmr2022
View on GitHub
Code for "Self-Lifting: A Novel Framework For Unsupervised Voice-Face Association Learning,ICMR,2022"
☆15Oct 25, 2024Updated last year
GPUs on demand by Runpod - Special Offer Available • Ad
Run AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
Mombin / speech2vid
View on GitHub
☆15Oct 28, 2019Updated 6 years ago
choyingw / Cross-Modal-Perceptionist
View on GitHub
CVPR 2022: Cross-Modal Perceptionist: Can Face Geometry be Gleaned from Voices?
☆130Dec 11, 2024Updated last year
wanganran / MilliSonic
View on GitHub
Pushing the limits of acoustic motion tracking
☆14Jul 31, 2020Updated 5 years ago
uniBruce / Mead
View on GitHub
MEAD: A Large-scale Audio-visual Dataset for Emotional Talking-face Generation [ECCV2020]
☆306Jul 7, 2024Updated 2 years ago
oneCodeSuperman / wav2lip_hq_trt
View on GitHub
这是一个在wav2lip，使用wav2lip、gfpgan、yolov5等模型用RT加速的超快推理！经测试在2070显卡上可达到0.03秒每帧实现实时推理。
☆31Sep 23, 2025Updated 10 months ago
BAI-Yeqi / SF2F_PyTorch
View on GitHub
☆16Apr 27, 2025Updated last year
yliess86 / PaintsTorch
View on GitHub
PaintsTorch: Automatic Lineart Colorization
☆10Jun 21, 2019Updated 7 years ago