Code for ORAR Agent for Vision and Language Navigation on Touchdown and map2seq
☆20Nov 3, 2023Updated 2 years ago
Alternatives and similar repositories for map2seq_vln
Users that are interested in map2seq_vln are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of "Multimodal Text Style Transfer for Outdoor Vision-and-Language Navigation"☆27Mar 4, 2021Updated 5 years ago
- Official implementation of the ECCV 2022 Oral paper: Sim-2-Sim Transfer for Vision-and-Language Navigation in Continuous Environments☆35Dec 16, 2023Updated 2 years ago
- Vision and Language Agent Navigation☆85Jan 29, 2021Updated 5 years ago
- Nocturnal Visual Place Recognition via Generative and Inherited Knowledge Transfer☆14Dec 3, 2024Updated last year
- [AAAI-25 Oral] Official Implementation of "FLAME: Learning to Navigate with Multimodal LLM in Urban Environments"☆68Nov 2, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- REVERIE: Remote Embodied Visual Referring Expression in Real Indoor Environments☆158May 15, 2026Updated 2 months ago
- ☆15Feb 22, 2023Updated 3 years ago
- Vision-and-Language Navigation in Continuous Environments using Habitat☆835Jan 7, 2025Updated last year
- Official implementation of: Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel☆35Jun 10, 2025Updated last year
- Implementation of Trust Region Policy Optimization and Proximal Policy Optimization algorithms on the objective of Robot Walk.☆12Mar 9, 2021Updated 5 years ago
- Official implementation of Think Global, Act Local: Dual-scale GraphTransformer for Vision-and-Language Navigation (CVPR'22 Oral).☆282Jun 27, 2023Updated 3 years ago
- Dataset for Bilingual VLN☆11Dec 5, 2020Updated 5 years ago
- [ACL2023] Official code repository for VLN-Trans☆14Sep 10, 2023Updated 2 years ago
- ☆10Oct 1, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for the paper "Data Efficient Masked Language Modeling for Vision and Language".☆18Sep 17, 2021Updated 4 years ago
- Code for the ACL 2022 paper "Continual Sequence Generation with Adaptive Compositional Modules"☆39Apr 4, 2022Updated 4 years ago
- Code for the paper "Improving Vision-and-Language Navigation with Image-Text Pairs from the Web" (ECCV 2020)☆59Oct 7, 2022Updated 3 years ago
- Fast-Slow Test-time Adaptation for Online Vision-and-Language Navigation☆35Dec 5, 2025Updated 7 months ago
- ☆11Oct 16, 2023Updated 2 years ago
- ☆24Jul 16, 2024Updated 2 years ago
- [ICCV 2023] Simple Baselines for Interactive Video Retrieval with Questions and Answers☆20Apr 16, 2024Updated 2 years ago
- PyTorch implementation of Vanilla PG, TNPG, TRPO, PPO on Mujoco environment☆12Feb 22, 2019Updated 7 years ago
- Official implementation of Why Only Text: Empowering Vision-and-Language Navigation with Multi-modal Prompts(IJCAI 2024)☆15Oct 16, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Repository for "Who Plays First? Optimizing the Order of Play in Stackelberg Games with Many Robots" - RSS 2024☆18Jun 25, 2024Updated 2 years ago
- ☆17Mar 3, 2025Updated last year
- A fast Ramer-Douglas-Peucker algorithm implementation.☆15Sep 3, 2023Updated 2 years ago
- Enables AI agents to use Google Maps features (geocoding, elevation, search, directions) via the Agent-to-Agent (A2A) protocol.☆17Apr 29, 2025Updated last year
- [TGRS2022] [CorrNet] Lightweight Salient Object Detection in Optical Remote Sensing Images via Feature Correlation☆23Nov 17, 2023Updated 2 years ago
- ☆59Apr 1, 2022Updated 4 years ago
- Code for CVPR22 paper One Step at a Time: Long-Horizon Vision-and-Language Navigation with Milestones☆13Jul 27, 2022Updated 3 years ago
- ☆28Aug 31, 2023Updated 2 years ago
- ☆13Sep 23, 2023Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official implementation of History Aware Multimodal Transformer for Vision-and-Language Navigation (NeurIPS'21).☆146Jun 14, 2023Updated 3 years ago
- ☆16Apr 9, 2021Updated 5 years ago
- [TIP2023] [GeleNet] Salient Object Detection in Optical Remote Sensing Images Driven by Transformer☆36May 12, 2024Updated 2 years ago
- Training code of waypoint predictor in Discrete-to-Continuous VLN.☆32Mar 25, 2024Updated 2 years ago
- ☆27Jun 22, 2024Updated 2 years ago
- Team: bacon-reloaded☆21May 22, 2017Updated 9 years ago
- [ICCV 2025] MMGeo: Multimodal Compositional Geo-Localization for UAVs☆20Oct 20, 2025Updated 9 months ago