An effective multimodal representation and fusion method for multimodal intent recognition
☆19Jun 7, 2024Updated 2 years ago
Alternatives and similar repositories for EMRFM
Users that are interested in EMRFM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MIntRec: A New Dataset for Multimodal Intent Recognition (ACM MM 2022)☆138May 2, 2025Updated last year
- An official pytorch implementation for the paper: SDR-GNN: Spectral Domain Reconstruction Graph Neural Network for incomplete multimodal …☆18Dec 30, 2024Updated last year
- [ICME 2023 Oral] Pytorch implementation for Multimodal Sentiment Analysis with Preferential Fusion and Distance-aware Contrastive Learnin…☆24Jan 2, 2024Updated 2 years ago
- The implementation codes of paper: Multimodal Sentiment Analysis with Mutual Information-based Disentangled Representation Learning☆23May 8, 2025Updated last year
- [AAAI 2024] DTF-AT: Decoupled Time-Frequency Audio Transformer for Event Classification☆12Mar 10, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and models for the Action Recognition benchmark of Assembly101☆15Mar 26, 2023Updated 3 years ago
- [ACMMM 2023] BMMAL: Towards Balanced Active Learning for Multimodal Classification☆17Sep 25, 2023Updated 2 years ago
- [IROS 2023] Interactive Spatiotemporal Token Attention Network for Skeleton-based General Interactive Action Recognition☆21Jul 12, 2025Updated last year
- Repo for 'VLM-Auto: VLM-based Autonomous Driving Assistant with Human-like Behavior and Understanding for Complex Road Scenes'☆27Oct 10, 2024Updated last year
- A Spatio-temporal Scenario Benchmark for Multi-modal Large Language Models in Autonomous Driving☆23Nov 27, 2025Updated 8 months ago
- [CVPR25] Official Implementation of CAV-MAE Sync☆32Apr 5, 2026Updated 4 months ago
- STI-Bench : Are MLLMs Ready for Precise Spatial-Temporal World Understanding?☆39Jan 12, 2026Updated 7 months ago
- Improving Mamaba performance on Video Understanding task☆49Dec 30, 2025Updated 7 months ago
- Code for dmrnet☆44Jul 16, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆42Nov 22, 2024Updated last year
- Official codebase for "Unveiling the Power of Audio-Visual Early Fusion Transformers with Dense Interactions through Masked Modeling".☆43Aug 2, 2024Updated 2 years ago
- CoRL2024 | Hint-AD: Holistically Aligned Interpretability for End-to-End Autonomous Driving☆74Oct 30, 2024Updated last year
- The official code for Improving Multimodal Learning via Imbalanced Learning☆41Mar 26, 2026Updated 4 months ago
- ☆66Sep 3, 2024Updated last year
- ☆68Mar 27, 2024Updated 2 years ago
- [FG 2025] official implementation for the paper 'Representation Learning and Identity Adversarial Training for Facial Behavior Understand…☆77Jun 13, 2025Updated last year
- TCL-MAP is a powerful method for multimodal intent recognition (AAAI 2024)☆62Jan 25, 2024Updated 2 years ago
- MISA: Modality-Invariant and -Specific Representations for Multimodal Sentiment Analysis☆294Mar 14, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2025] Multi-modal representation learning of shared, unique and synergistic features between modalities☆71May 6, 2025Updated last year
- An unofficial implementation of TubeViT in "Rethinking Video ViTs: Sparse Video Tubes for Joint Image and Video Learning"☆95Jul 15, 2026Updated last month
- [NeurIPS 2025] SURDS: Benchmarking Spatial Understanding and Reasoning in Driving Scenarios with Vision Language Models☆84Sep 23, 2025Updated 10 months ago
- [AAAI 2020] Official implementation of VAANet for Emotion Recognition☆84Oct 3, 2023Updated 2 years ago
- ☆108Dec 27, 2024Updated last year
- [CVPR 2025] Bridging Past and Future: End-to-End Autonomous Driving with Historical Prediction and Planning☆108Apr 7, 2025Updated last year
- [CVPR 2025] DyCoke: Dynamic Compression of Tokens for Fast Video Large Language Models☆114Nov 22, 2025Updated 8 months ago
- Official Code Release of "FusionAD"☆165Jul 9, 2024Updated 2 years ago
- Our ECCV 2022 paper Human Trajectory Prediction via Neural Social Physics☆151Mar 8, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This repo contains the code for paper "LightEMMA: Lightweight End-to-End Multimodal Model for Autonomous Driving"☆145Nov 19, 2025Updated 8 months ago
- [TAFFC 2024] The official implementation of paper: From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Rec…☆125Aug 9, 2026Updated last week
- [CVPR 2023] Micron-BERT: BERT-based Facial Micro-Expression Recognition☆185May 12, 2026Updated 3 months ago
- NeuroNCAP benchmark for end-to-end autonomous driving☆253Oct 14, 2024Updated last year
- A curated list of balanced multimodal learning methods.☆172Mar 26, 2026Updated 4 months ago
- ☆173Sep 18, 2023Updated 2 years ago
- CVPR 2024 Papers Autonomous Driving☆257Aug 12, 2024Updated 2 years ago