[CVPR 2026] Official implementation of FantasyVLN: Unified Multimodal Chain-of-Thought Reasoning for Vision-and-Language Navigation
☆38Feb 23, 2026Updated 5 months ago
Alternatives and similar repositories for fantasy-vln
Users that are interested in fantasy-vln are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 【ICLR 2026】 Official implementation of [OmniNav: A Unified Framework for Prospective Exploration and Visual-Language Navigation]☆210Updated this week
- [ICML 2026 Spotlight] UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models☆29May 1, 2026Updated 3 months ago
- ☆28Dec 19, 2025Updated 7 months ago
- Sparse Video Generation Model for Embodied Navigation conditioned on loose language guidance, 100% real world verification☆115Jul 10, 2026Updated last month
- InternRobotics' open platform for building generalized navigation foundation models.☆1,044Mar 10, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Towards Long-Horizon Vision-Language Navigation: Platform, Benchmark and Method (CVPR-25)☆254Aug 20, 2025Updated 11 months ago
- ☆221Aug 5, 2026Updated last week
- [CoRL 2025] Human-like Navigation in a World Built for Humans☆42Sep 25, 2025Updated 10 months ago
- ☆59Aug 18, 2025Updated last year
- Code of the paper "EvolveNav: Empowering LLM-Based Vision-Language Navigation via Self-Improving Embodied Reasoning" (TPAMI 2026)☆37Oct 14, 2025Updated 10 months ago
- [ICRA 2026] Official implementation of the paper: "StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling"☆578Nov 2, 2025Updated 9 months ago
- [AAAI 2025] Enhancing Multi-Robot Semantic Navigation Through Multimodal Chain-of-Thought Score Collaboration☆33Dec 13, 2024Updated last year
- ☆21Mar 19, 2026Updated 4 months ago
- Nav-R1: Reasoning and Navigation in Embodied Scenes☆131Oct 31, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official repository of General Scene Adaptation for Vision-and-Language Navigation (ICLR'2025)☆73Apr 16, 2025Updated last year
- Official implementation of [AstraNav-World: World Model for Foresight Control and Consistency]☆99Jan 21, 2026Updated 6 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆22Nov 18, 2025Updated 9 months ago
- This is the official repository for VLN-CLASH.☆28Aug 5, 2025Updated last year
- [CVPR 2026] Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation☆28Jun 11, 2026Updated 2 months ago
- Official implementation of the paper: "ActiveVLN: Towards Active Exploration via Multi-Turn RL in Vision-and-Language Navigation"☆75Feb 11, 2026Updated 6 months ago
- [IROS'25 Oral] WMNav: Integrating Vision-Language Models into World Models for Object Goal Navigation☆177Mar 24, 2026Updated 4 months ago
- This is the official repository for MAGIC: Meta-Ability Guided Interactive Chain-of-Distillation Learning towards Efficient Vision-and-La…☆17May 17, 2026Updated 3 months ago
- ☆11Feb 5, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Repository for the "AnywhereVLA: Language-Conditioned Exploration and Mobile Manipulation" paper☆26Oct 25, 2025Updated 9 months ago
- ☆22Mar 7, 2026Updated 5 months ago
- [ICCV 25] Official repository of "Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dial…☆31Apr 1, 2026Updated 4 months ago
- ☆106Jun 12, 2026Updated 2 months ago
- SplatSDF, ICRA 2026☆21Feb 21, 2026Updated 5 months ago
- ☆34May 4, 2026Updated 3 months ago
- [ICRA 2026] Official implementation of the paper: "TagaVLM: Topology-Aware Global Action Reasoning for Vision-Language Navigation"☆20Apr 30, 2026Updated 3 months ago
- Official codebase for the paper "WorldMemArena: Evaluating Multimodal Agent Memory Through Action–World Interaction"☆25May 29, 2026Updated 2 months ago
- Official code of Geometric Autoencoder for Diffusion Models.☆21Mar 12, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official code of "Towards Long-horizon Agentic Multimodal Search"☆29Apr 17, 2026Updated 4 months ago
- [CVPR 2025] UniGoal: Towards Universal Zero-shot Goal-oriented Navigation☆349Sep 16, 2025Updated 11 months ago
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- This is the source code to paper “DAgger Diffusion Navigation: DAgger Boosted Diffusion Policy for Vision-Language Navigation”.☆36Aug 13, 2025Updated last year
- CVPR 2026 - MSGNav: Unleashing the Power of Multi-modal 3D Scene Graph for Zero-Shot Embodied Navigation☆69Mar 23, 2026Updated 4 months ago
- DeSplat: Decomposed Gaussian Splatting for Distractor-Free Rendering☆61Oct 14, 2025Updated 10 months ago
- Code of the paper "NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning" (TPAMI 2025)☆144Jun 4, 2025Updated last year