Baseline for REVERIE-Challenge using HOP
☆10Jul 4, 2022Updated 4 years ago
Alternatives and similar repositories for HOP-REVERIE-Challenge
Users that are interested in HOP-REVERIE-Challenge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code of the CVPR 2022 paper "HOP: History-and-Order Aware Pre-training for Vision-and-Language Navigation"☆31Aug 21, 2023Updated 2 years ago
- Know What and Know Where: An Object-and-Room Informed Sequential BERT for Indoor Vision-Language Navigation☆16Feb 7, 2022Updated 4 years ago
- Official REVERIE Grounding Model of REVERIE Challenge @ CSIG 2022☆19Oct 17, 2022Updated 3 years ago
- A human-annotated, fine-grained dataset for Vision-and-Language Navigation☆17Jan 20, 2022Updated 4 years ago
- Official implementation of WebVLN: Vision-and-Language Navigation on Websites☆35Jan 2, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code of the NeurIPS 2021 paper: Language and Visual Entity Relationship Graph for Agent Navigation☆47Oct 31, 2021Updated 4 years ago
- REVERIE: Remote Embodied Visual Referring Expression in Real Indoor Environments☆158May 15, 2026Updated 2 months ago
- ☆10Nov 16, 2023Updated 2 years ago
- Code of the ICCV 2023 paper "March in Chat: Interactive Prompting for Remote Embodied Referring Expression"☆26May 22, 2024Updated 2 years ago
- Code of the CVPR 2021 Oral paper: A Recurrent Vision-and-Language BERT for Navigation☆209Aug 13, 2022Updated 3 years ago
- Official implementation of History Aware Multimodal Transformer for Vision-and-Language Navigation (NeurIPS'21).☆147Jun 14, 2023Updated 3 years ago
- ☆35Aug 19, 2023Updated 2 years ago
- Official Implementation of Frequency-enhanced Data Augmentation for Vision-and-Language Navigation (NeurIPS2023)☆14Jan 8, 2024Updated 2 years ago
- Pytorch implementation for “V2C: Visual Voice Cloning”☆35Jan 28, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACM MM 2021 Oral] Official repo of "Neighbor-view Enhanced Model for Vision and Language Navigation"☆78Nov 16, 2022Updated 3 years ago
- ☆20Mar 19, 2026Updated 4 months ago
- Official implementation of Layout-aware Dreamer for Embodied Referring Expression Grounding [AAAI 23].☆16Apr 13, 2023Updated 3 years ago
- ☆23Dec 9, 2021Updated 4 years ago
- Official Implementation for CVPR 2022 paper "Unsupervised Vision-Language Parsing: Seamlessly Bridging Visual Scene Graphs with Language …☆24Oct 19, 2022Updated 3 years ago
- Official implementation of Sim-to-Real Transfer via 3D Feature Fields for Vision-and-Language Navigation (CoRL'24).☆80Dec 26, 2025Updated 7 months ago
- ☆12Sep 25, 2023Updated 2 years ago
- ☆43May 23, 2023Updated 3 years ago
- Code for the paper "Improving Vision-and-Language Navigation with Image-Text Pairs from the Web" (ECCV 2020)☆59Oct 7, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICCV 2023 Oral]: Scaling Data Generation in Vision-and-Language Navigation☆225Jul 2, 2025Updated last year
- ☆13Feb 17, 2025Updated last year
- ☆13May 6, 2025Updated last year
- Official implementation of Think Global, Act Local: Dual-scale GraphTransformer for Vision-and-Language Navigation (CVPR'22 Oral).☆284Jun 27, 2023Updated 3 years ago
- ☆19Mar 11, 2022Updated 4 years ago
- Dataset and baseline for Scenario Oriented Object Navigation (SOON)☆25Nov 23, 2021Updated 4 years ago
- Training code of waypoint predictor in Discrete-to-Continuous VLN.☆32Mar 25, 2024Updated 2 years ago
- [ICCV 2025] Official implementation of SAME: Learning Generic Language-Guided Visual Navigation with State-Adaptive Mixture of Experts☆40Apr 3, 2026Updated 3 months ago
- The repository of ECCV 2020 paper `Active Visual Information Gathering for Vision-Language Navigation`☆44Apr 9, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Feature resources of "Diagnosing the Environment Bias in Vision-and-Language Navigation"☆16May 6, 2020Updated 6 years ago
- Official implementation of the NRNS paper☆37Jun 13, 2022Updated 4 years ago
- CoADNet: Collaborative Aggregation-and-Distribution Networks for Co-Salient Object Detection☆19Jan 8, 2021Updated 5 years ago
- [CVPR2022] Official code for Hierarchical Modular Network for Video Captioning. Our proposed HMN is implemented with PyTorch.☆50Sep 30, 2022Updated 3 years ago
- [ICCV 2023] Official repo of "BEVBert: Multimodal Map Pre-training for Language-guided Navigation"☆260Apr 27, 2026Updated 3 months ago
- Controllable mage captioning model with unsupervised modes☆21Apr 14, 2023Updated 3 years ago
- FELA: Learning Fine-Grained Alignment for Aerial Vision-Dialog Navigation, AAAI 2025.☆38Dec 18, 2024Updated last year