[IJCNN2026] Official code for "Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations"
☆35Apr 7, 2025Updated last year
Alternatives and similar repositories for VLN-MBA-VisualPerturbations
Users that are interested in VLN-MBA-VisualPerturbations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [MM 2024] The official implementation code for "DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimatio…☆37Oct 31, 2024Updated last year
- [TAFFC2026] The official implementation code for "Bidirectional Learning of Facial Action Units and Expressions via Structured Semantic M…☆28Aug 20, 2026Updated last week
- [MM 2025] The official implementation for the paper titled: "Listening to the Unspoken: Exploring '365' Aspects of Multimodal Interview P…☆31Jul 31, 2025Updated last year
- [SPL2026] CLAIP-Emo: Parameter-Efficient Adaptation of Language-supervised models for In-the-Wild Audiovisual Emotion Recognition☆29Aug 20, 2026Updated last week
- [TCSS 2024] MAE pre-training models (ViT and ConvNeXt) using AffectNet images for static facial expression recognition (SFER).☆43Jun 3, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [TAFFC 2025] The offical implementation of paper: Static for Dynamic: Towards a Deeper Understanding of Dynamic Facial Expressions Using…☆70Aug 20, 2026Updated last week
- [TAFFC 2024] The official implementation of paper: From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Rec…☆125Aug 20, 2026Updated last week
- Official implementation of Layout-aware Dreamer for Embodied Referring Expression Grounding [AAAI 23].☆15Apr 13, 2023Updated 3 years ago
- Portal for resources for the Stretch community☆13May 14, 2025Updated last year
- ☆10Nov 16, 2023Updated 2 years ago
- ☆22Nov 12, 2025Updated 9 months ago
- Official Implementation of Frequency-enhanced Data Augmentation for Vision-and-Language Navigation (NeurIPS2023)☆15Jan 8, 2024Updated 2 years ago
- [ACM MM 2022] Target-Driven Structured Transformer Planner for Vision-Language Navigation☆16Nov 1, 2022Updated 3 years ago
- Code and Data for Paper: PanoGen: Text-Conditioned Panoramic Environment Generation for Vision-and-Language Navigation☆83May 31, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆28Apr 29, 2025Updated last year
- This is our code for EmotiW_2019 Student Engagement Regression Task.☆19Jul 12, 2019Updated 7 years ago
- This is the official repository for MAGIC: Meta-Ability Guided Interactive Chain-of-Distillation Learning towards Efficient Vision-and-La…☆17May 17, 2026Updated 3 months ago
- Eye closure detection based on EAR☆14Apr 14, 2020Updated 6 years ago
- Code for A Dual Semantic-Aware Recurrent Global-Adaptive Network For Vision-and-Language Navigation☆17Apr 25, 2024Updated 2 years ago
- OVSegDT, a lightweight transformer policy to solve Open-vocabulary Object Goal Navigation☆19May 25, 2026Updated 3 months ago
- Several approaches for bumping up image resolution with NNs (GANs)☆12Feb 17, 2019Updated 7 years ago
- Simple Pose: Rethinking and Improving a Bottom-up Approach for Multi-Person Pose Estimation☆12Oct 6, 2020Updated 5 years ago
- ☆19Mar 11, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Repository to train models on AffectNet for 231N☆12Jun 8, 2018Updated 8 years ago
- Dataset and baseline for Scenario Oriented Object Navigation (SOON)☆25Nov 23, 2021Updated 4 years ago
- [ICCV 2023] Official repo of "BEVBert: Multimodal Map Pre-training for Language-guided Navigation"☆260Apr 27, 2026Updated 4 months ago
- Implementation of the CVPR'17 paper, Reliable Crowdsourcing and Deep Locality-Preserving Learning for Expression Recognition in the Wild☆13Sep 28, 2022Updated 3 years ago
- Official implementation of "Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation" (NeurIPS'25 Oral)☆92Dec 22, 2025Updated 8 months ago
- Code for the paper "3D Human Pose Estimation with Siamese Equivariant Embedding"☆19Sep 26, 2018Updated 7 years ago
- Demo code for the paper Structure-Aware and Temporally Coherent 3D Human Pose Estimation☆16Jan 9, 2020Updated 6 years ago
- Spatial-X: Zero-Shot Vision-and-Language Navigation with Spatial Scene Priors☆23Apr 5, 2026Updated 4 months ago
- Pytorch implementation of Realtime_Multi-Person_Pose_Estimation☆19Mar 6, 2018Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- GQA-OOD is a new dataset and benchmark for the evaluation of VQA models in OOD (out of distribution) settings.☆33Mar 1, 2021Updated 5 years ago
- Mechanistic Interpretability toolkit for Vision-Language-Action models☆20Aug 3, 2026Updated 3 weeks ago
- ☆22Jul 22, 2025Updated last year
- Utility files for human pose estimation in python☆23Jan 28, 2020Updated 6 years ago
- The official PyTorch implementation of "Following the Human Thread in Social Navigation", International Conference on Learning Representa…☆22Jun 23, 2025Updated last year
- [CVPR 2026] Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation☆31Jun 11, 2026Updated 2 months ago
- Generative Bias for Robust Visual Question Answering ( CVPR 2023 )☆28Jul 4, 2023Updated 3 years ago