[CVPR'25] AVF-MAE++ : Scaling Affective Video Facial Masked Autoencoders via Efficient Audio-Visual Self-Supervised Learning
☆22Jun 11, 2026Updated last month
Alternatives and similar repositories for AVF-MAE
Users that are interested in AVF-MAE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2026] Facial-R1: Aligning Reasoning and Recognition for Facial Emotion Analysis☆19May 29, 2026Updated last month
- [CVPR 2024] EmoVIT: Revolutionizing Emotion Insights with Visual Instruction Tuning☆40Apr 20, 2025Updated last year
- ☆22Jan 17, 2025Updated last year
- ☆48Jun 27, 2022Updated 4 years ago
- ☆17Jul 16, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Independent Vector Analysis (IVA-G and IVA-L-SOS) implemented in Python☆21Nov 24, 2025Updated 7 months ago
- This repository provides the codes for MMA-DFER: multimodal (audiovisual) emotion recognition method. This is an official implementation …☆57Sep 16, 2024Updated last year
- ☆10Dec 3, 2024Updated last year
- Repo for "Uncertain Multimodal Intention and Emotion Understanding in the Wild"☆19Oct 20, 2025Updated 9 months ago
- [NeurIPS 2025] This is the official repository for "RAD: Towards Trustworthy Retrieval-Augmented Multi-modal Clinical Diagnosis"☆27Nov 21, 2025Updated 8 months ago
- EmoCapCLIP: Learning Transferable Facial Emotion Representations from Large-Scale Semantically Rich Captions☆22Jul 29, 2025Updated 11 months ago
- ☆29Aug 2, 2023Updated 2 years ago
- ☆20Mar 5, 2021Updated 5 years ago
- ☆24Mar 25, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- This is the code for EEGDnet: Fusing non-local and local self-similarity for EEG signal denoising with transformer☆17Jun 3, 2024Updated 2 years ago
- ☆23Aug 11, 2020Updated 5 years ago
- GUI for plotting/storing data from ProtoCentral breakouts☆24Jul 14, 2026Updated last week
- ☆19Oct 29, 2024Updated last year
- MoIIE: Mixture of Intra- and Inter-Modality Experts for Large Vision Language Models☆19Aug 14, 2025Updated 11 months ago
- ☆11Aug 20, 2025Updated 11 months ago
- awesome video-based self-supervised learning methods in recently years☆10Nov 26, 2020Updated 5 years ago
- MAE-DFER: Efficient Masked Autoencoder for Self-supervised Dynamic Facial Expression Recognition (ACM MM 2023)☆149Nov 16, 2025Updated 8 months ago
- Semantic-decoupled Spatial Partition Guided Point-supervised Oriented Object Detection☆13Jul 7, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆16Jul 8, 2026Updated last week
- Official repo and evaluation implementation of KnowRecall and VisRecall☆10May 22, 2025Updated last year
- MMFformer: Multimodal Fusion Transformer Network for Depression Detection☆20Sep 26, 2025Updated 9 months ago
- ☆13Dec 2, 2024Updated last year
- Code for ICASSP 2024 Paper: RECAP: Retrieval-Augmented Audio Captioning☆16Jun 23, 2024Updated 2 years ago
- [ACL 2024] A Multimodal, Multigenre, and Multipurpose Audio-Visual Academic Lecture Dataset☆24May 29, 2025Updated last year
- ☆19Dec 3, 2025Updated 7 months ago
- ABAW3 (CVPRW): A Joint Cross-Attention Model for Audio-Visual Fusion in Dimensional Emotion Recognition☆50Jan 15, 2024Updated 2 years ago
- This repository is related to 'Intriguing Properties of Hyperbolic Embeddings in Vision-Language Models', published at TMLR (2024), https…☆21Jul 5, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- GPT-4V with Emotion☆97Dec 8, 2023Updated 2 years ago
- Multimodal Genuine Emotion and Expression Detection database☆13Jul 15, 2024Updated 2 years ago
- ☆13Sep 3, 2024Updated last year
- Pipeline for the preprocessing of Micro-Expressions.☆12Apr 27, 2026Updated 2 months ago
- ☆19May 21, 2025Updated last year
- [CVPR 2023] Micron-BERT: BERT-based Facial Micro-Expression Recognition☆186May 12, 2026Updated 2 months ago
- [MAC 2024] The baseline code for MAC 2024.☆12Jun 3, 2025Updated last year