[IJCNN2026] Official code for "Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations"
☆35Apr 7, 2025Updated last year
Alternatives and similar repositories for VLN-MBA-VisualPerturbations
Users that are interested in VLN-MBA-VisualPerturbations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for "Exploring and exploiting model uncertainty for robust visual question answering"☆30Apr 7, 2025Updated last year
- [AAAI 2026] Official code for "Agent Journey Beyond RGB: Unveiling Hybrid Semantic-Spatial Environmental Representations for Vision-and-L…☆43Mar 22, 2026Updated 4 months ago
- [MM 2024] The official implementation code for "DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimatio…☆37Oct 31, 2024Updated last year
- [MM 2025] The official implementation code for "VAEmo: Efficient Representation Learning for Visual-Audio Emotion with Knowledge Injectio…☆38Apr 4, 2026Updated 4 months ago
- [ICMR 2025] The official implementation for the paper titled: "Concept Drift Guided LayerNorm Tuning for Efficient Multimodal Metaphor Id…☆28Apr 27, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SDGTALK: STRUCTURED FACIAL PRIORS AND DUAL-BRANCH MOTION FIELDS FOR GENERALIZABLE GAUSSIAN TALKING HEAD SYNTHESIS☆26May 11, 2026Updated 2 months ago
- The official implementation code for "Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic Mapping acros…☆28Aug 2, 2026Updated last week
- [MM 2025] The official implementation for the paper titled: "Listening to the Unspoken: Exploring '365' Aspects of Multimodal Interview P…☆31Jul 31, 2025Updated last year
- [TCSS 2025] The official implementation code for "PhysioSync: Temporal and Cross-Modal Contrastive Learning Inspired by Physiological Syn…☆46Dec 29, 2025Updated 7 months ago
- [CVPRW 2023]The Winner's Solution of CVPR2023-ABAW5 Emotional Reaction Intensity (ERI) Estimation Challenge☆27Mar 19, 2023Updated 3 years ago
- [TAFFC 2025] The offical implementation of paper: Static for Dynamic: Towards a Deeper Understanding of Dynamic Facial Expressions Using…☆69Jul 22, 2026Updated 2 weeks ago
- [TAFFC 2024] The official implementation of paper: From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Rec…☆126Aug 3, 2026Updated last week
- Code of the paper "Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation"…☆20Nov 11, 2025Updated 8 months ago
- Official implementation of Layout-aware Dreamer for Embodied Referring Expression Grounding [AAAI 23].☆16Apr 13, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for the paper "Spectrum Guided Topology Augmentation for Graph Contrastive Learning"☆11Jul 18, 2023Updated 3 years ago
- Projects Student Engagement Detection System in E-Learning Environment using OpenCV and CNN☆17Dec 6, 2023Updated 2 years ago
- Portal for resources for the Stretch community☆13May 14, 2025Updated last year
- ☆16Oct 13, 2025Updated 9 months ago
- ☆22Jan 17, 2025Updated last year
- ☆10Nov 16, 2023Updated 2 years ago
- ☆22Nov 12, 2025Updated 8 months ago
- Official Implementation of Frequency-enhanced Data Augmentation for Vision-and-Language Navigation (NeurIPS2023)☆14Jan 8, 2024Updated 2 years ago
- ☆20Jun 26, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ACM MM 2022] Target-Driven Structured Transformer Planner for Vision-Language Navigation☆16Nov 1, 2022Updated 3 years ago
- Implemention of "Realtime Multi Person Pose-Estimation" in pytorch with data from AI Challenger☆13Nov 24, 2017Updated 8 years ago
- Code and Data for Paper: PanoGen: Text-Conditioned Panoramic Environment Generation for Vision-and-Language Navigation☆83May 31, 2023Updated 3 years ago
- ☆27Apr 29, 2025Updated last year
- This is our code for EmotiW_2019 Student Engagement Regression Task.☆19Jul 12, 2019Updated 7 years ago
- [AAAI 2023] AVCAffe: A Large Scale Audio-Visual Dataset of Cognitive Load and Affect for Remote Work☆22Dec 7, 2025Updated 8 months ago
- This is the official repository for MAGIC: Meta-Ability Guided Interactive Chain-of-Distillation Learning towards Efficient Vision-and-La…☆17May 17, 2026Updated 2 months ago
- Eye closure detection based on EAR☆14Apr 14, 2020Updated 6 years ago
- OVSegDT, a lightweight transformer policy to solve Open-vocabulary Object Goal Navigation☆17May 25, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19May 31, 2023Updated 3 years ago
- Several approaches for bumping up image resolution with NNs (GANs)☆12Feb 17, 2019Updated 7 years ago
- Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose Estimation☆12Oct 6, 2020Updated 5 years ago
- ☆19Mar 11, 2022Updated 4 years ago
- Repository to train models on AffectNet for 231N☆12Jun 8, 2018Updated 8 years ago
- Dataset and baseline for Scenario Oriented Object Navigation (SOON)☆25Nov 23, 2021Updated 4 years ago
- [ICCV 2023] Official repo of "BEVBert: Multimodal Map Pre-training for Language-guided Navigation"☆259Apr 27, 2026Updated 3 months ago