[AAAI 2026] Official code for "Agent Journey Beyond RGB: Unveiling Hybrid Semantic-Spatial Environmental Representations for Vision-and-Language Navigation"
☆43Mar 22, 2026Updated 4 months ago
Alternatives and similar repositories for VLN-SUSA
Users that are interested in VLN-SUSA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [IJCNN2026] Official code for "Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations"☆35Apr 7, 2025Updated last year
- Official code for "Exploring and exploiting model uncertainty for robust visual question answering"☆30Apr 7, 2025Updated last year
- [MM 2024] The official implementation code for "DAT: Dialogue-Aware Transformer with Modality-Group Fusion for Human Engagement Estimatio…☆37Oct 31, 2024Updated last year
- [MM 2025] The official implementation code for "VAEmo: Efficient Representation Learning for Visual-Audio Emotion with Knowledge Injectio…☆38Apr 4, 2026Updated 4 months ago
- [MM 2025] The official implementation for the paper titled: "Traits Run Deep: Enhancing Personality Assessment via Psychology-Guided LLM …☆26Jul 30, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [TCSS 2025] The official implementation code for "PhysioSync: Temporal and Cross-Modal Contrastive Learning Inspired by Physiological Syn…☆46Dec 29, 2025Updated 7 months ago
- [TCSS 2024] MAE pre-training models (ViT and ConvNeXt) using AffectNet images for static facial expression recognition (SFER).☆42Jun 3, 2025Updated last year
- [CVPRW 2023]The Winner's Solution of CVPR2023-ABAW5 Emotional Reaction Intensity (ERI) Estimation Challenge☆27Mar 19, 2023Updated 3 years ago
- [TAFFC 2024] The official implementation of paper: From Static to Dynamic: Adapting Landmark-Aware Image Models for Facial Expression Rec…☆126Aug 3, 2026Updated last week
- This is the source code to paper “DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation”.☆35Aug 13, 2025Updated 11 months ago
- Official implementation of: Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel☆35Jun 10, 2025Updated last year
- ☆14Apr 28, 2025Updated last year
- Code of the paper "Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation"…☆20Nov 11, 2025Updated 9 months ago
- Official implementation of "Dynam3D: Dynamic Layered 3D Tokens Empower VLM for Vision-and-Language Navigation" (NeurIPS'25 Oral)☆92Dec 22, 2025Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official implementation of "g3D-LF: Generalizable 3D-Language Feature Fields for Embodied Tasks" (CVPR'25).☆57Jul 14, 2025Updated last year
- Dataset and baseline for Scenario Oriented Object Navigation (SOON)☆25Nov 23, 2021Updated 4 years ago
- Projects Student Engagement Detection System in E-Learning Environment using OpenCV and CNN☆17Dec 6, 2023Updated 2 years ago
- Non-IID Transfer Learning on Graphs☆13Jul 4, 2023Updated 3 years ago
- A real time system for classrooms for attendance and gathering attention data of each student☆11Apr 16, 2025Updated last year
- YOLO-NAS for ROS 2☆14Jun 5, 2023Updated 3 years ago
- [ISER 2023] The official implementation of Audio Visual Language Maps for Robot Navigation☆69May 11, 2024Updated 2 years ago
- [ICRA 25] InsCMPR: Efficient Cross-Modal Place Recognition via Instance-Aware Hybrid Mamba-Transformer☆26Sep 29, 2025Updated 10 months ago
- Portal for resources for the Stretch community☆13May 14, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆16Oct 13, 2025Updated 9 months ago
- This is the official PyTorch implementation of the CVPR 2023 paper: "GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot A…☆10Mar 17, 2024Updated 2 years ago
- ☆22Jan 17, 2025Updated last year
- ☆22Nov 12, 2025Updated 8 months ago
- Official Implementation of Frequency-enhanced Data Augmentation for Vision-and-Language Navigation (NeurIPS2023)☆14Jan 8, 2024Updated 2 years ago
- ☆24Apr 12, 2025Updated last year
- [TASE 2025] Efficient Alignment of Unconditioned Action Prior for Language-conditioned Pick and Place in Clutter☆36Oct 27, 2025Updated 9 months ago
- ☆20Jun 26, 2025Updated last year
- [ACM MM 2022] Target-Driven Structured Transformer Planner for Vision-Language Navigation☆16Nov 1, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implemention of "Realtime Multi Person Pose-Estimation" in pytorch with data from AI Challenger☆13Nov 24, 2017Updated 8 years ago
- GaRLIO: Gravity enhanced Radar-LiDAR-Inertial Odometry [ICRA 2025]☆104Jan 3, 2026Updated 7 months ago
- Code of the paper "EvolveNav: Empowering LLM-Based Vision-Language Navigation via Self-Improving Embodied Reasoning" (TPAMI 2026)☆37Oct 14, 2025Updated 9 months ago
- Official implementation of "NavRAG: Generating User Demand Instructions for Embodied Navigation through Retrieval-Augmented LLM" (ACL'25 …☆59Mar 6, 2025Updated last year
- [CVPR Workshop 2025 - OpenSun3D] ForesightNav: Learning Scene Imagination for Efficient Exploration☆79Apr 23, 2025Updated last year
- ☆46Sep 30, 2023Updated 2 years ago
- Human pose estimation☆12Mar 20, 2018Updated 8 years ago