Implementation of "Multimodal Text Style Transfer for Outdoor Vision-and-Language Navigation"
☆27Mar 4, 2021Updated 5 years ago
Alternatives and similar repositories for VLN-Transformer
Users that are interested in VLN-Transformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆10Nov 16, 2023Updated 2 years ago
- Cornell Touchdown natural language navigation and spatial reasoning dataset.☆114Sep 5, 2020Updated 5 years ago
- Code of the CVPR 2022 paper "HOP: History-and-Order Aware Pre-training for Vision-and-Language Navigation"☆31Aug 21, 2023Updated 2 years ago
- Implementation of "Visualize Before You Write: Imagination-Guided Open-Ended Text Generation".☆17Feb 3, 2023Updated 3 years ago
- VELMA agent for VLN in Street View☆31Sep 29, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repo for "Imagination-Augmented Natural Language Understanding", NAACL 2022.☆17Aug 30, 2022Updated 3 years ago
- Vision and Language Agent Navigation☆85Jan 29, 2021Updated 5 years ago
- Training code of waypoint predictor in Discrete-to-Continuous VLN.☆32Mar 25, 2024Updated 2 years ago
- Code and models of MOCA (Modular Object-Centric Approach) proposed in "Factorizing Perception and Policy for Interactive Instruction Foll…☆40Jun 21, 2024Updated 2 years ago
- A visual semantic planner for the ALFRED virtual agent challenge using the GPT-2 language model☆16Oct 1, 2020Updated 5 years ago
- Solving reinforcement learning tasks which require language and vision☆33Apr 4, 2023Updated 3 years ago
- Code for ORAR Agent for Vision and Language Navigation on Touchdown and map2seq☆20Nov 3, 2023Updated 2 years ago
- [ACL2023] Official code repository for VLN-Trans☆14Sep 10, 2023Updated 2 years ago
- Know What and Know Where: An Object-and-Room Informed Sequential BERT for Indoor Vision-Language Navigation☆16Feb 7, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A PyTorch implementation of SSCR☆23Aug 12, 2024Updated 2 years ago
- Official implementation of Layout-aware Dreamer for Embodied Referring Expression Grounding [AAAI 23].☆16Apr 13, 2023Updated 3 years ago
- Code for reproducing the results of NeurIPS 2020 paper "MultiON: Benchmarking Semantic Map Memory using Multi-Object Navigation”☆58Dec 8, 2020Updated 5 years ago
- Repository of our accepted NeurIPS-2022 paper "Towards Versatile Embodied Navigation"☆22Dec 8, 2022Updated 3 years ago
- https://xgxvisnav.github.io/☆22Dec 22, 2023Updated 2 years ago
- A human-annotated, fine-grained dataset for Vision-and-Language Navigation☆17Jan 20, 2022Updated 4 years ago
- REVERIE: Remote Embodied Visual Referring Expression in Real Indoor Environments☆159May 15, 2026Updated 2 months ago
- Official implementation of our EMNLP 2022 paper "CPL: Counterfactual Prompt Learning for Vision and Language Models"☆35Dec 5, 2022Updated 3 years ago
- Code and Data for our CVPR 2021 paper "Structured Scene Memory for Vision-Language Navigation"☆43Jul 31, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Multimodal-Procedural-Planning☆92Jun 1, 2023Updated 3 years ago
- Cooperative Vision-and-Dialog Navigation☆74Nov 22, 2022Updated 3 years ago
- Pytorch Code and Data for EnvEdit: Environment Editing for Vision-and-Language Navigation (CVPR 2022)☆30Aug 2, 2022Updated 4 years ago
- Panoramic Graph Environment Annotation toolkit, for collecting audio and text annotations in panoramic graph environments such as Matterp…☆19Mar 5, 2021Updated 5 years ago
- Official implementation of Learning from Unlabeled 3D Environments for Vision-and-Language Navigation (ECCV'22).☆44Mar 16, 2023Updated 3 years ago
- A multimodal dataset for google map restaurants.☆12Sep 28, 2022Updated 3 years ago
- ☆15Dec 23, 2022Updated 3 years ago
- ☆18Oct 7, 2019Updated 6 years ago
- Source code for paper "Trajectory of Alternating Direction Method of Multipliers and Adaptive Acceleration" of NeurIPS 2019☆10Jan 25, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆19Nov 7, 2020Updated 5 years ago
- Official implementation of the ECCV 2022 Oral paper: Sim-2-Sim Transfer for Vision-and-Language Navigation in Continuous Environments☆35Dec 16, 2023Updated 2 years ago
- Code of the NeurIPS 2021 paper: Language and Visual Entity Relationship Graph for Agent Navigation☆47Oct 31, 2021Updated 4 years ago
- Official implementation of the NRNS paper☆37Jun 13, 2022Updated 4 years ago
- ☆25Mar 9, 2023Updated 3 years ago
- ☆21Mar 19, 2026Updated 4 months ago
- [ECCV 2022] Official pytorch implementation of the paper "FedVLN: Privacy-preserving Federated Vision-and-Language Navigation"☆14Oct 8, 2022Updated 3 years ago