[CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation
☆85May 24, 2026Updated 3 months ago
Alternatives and similar repositories for AwareVLN
Users that are interested in AwareVLN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 3 months ago
- [ACL 2026 Poster] Code and Benchmark for "Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vi…☆21Jun 3, 2026Updated 3 months ago
- [ICCV 2025] IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation☆68Aug 4, 2025Updated last year
- The code of the paper "Latent Fingerprint Matching via Dense Minutia Descriptor"☆23Aug 27, 2025Updated last year
- F2F-AP: Flow-to-Future Asynchronous Policy for Real-time Dynamic Manipulation☆17Sep 5, 2026Updated 2 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 【ICLR 2026】 Official implementation of [OmniNav: A Unified Framework for Prospective Exploration and Visual-Language Navigation]☆224Sep 11, 2026Updated last week
- ☆21Feb 22, 2024Updated 2 years ago
- [ICRA 2026] Official implementation of the paper: "StreamVLN: Streaming Vision-and-Language Navigation via SlowFast Context Modeling"☆614Nov 2, 2025Updated 10 months ago
- A fingerprint recognition framework featuring fixed-length dense descriptors, robust enhancement, and pose-aware alignment. Code for desc…☆22Apr 24, 2026Updated 4 months ago
- [CoRL 2025] GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation☆83Jun 21, 2026Updated 2 months ago
- End-to-End Visual Language Navigation with Limited Sensing: A Survey☆45Updated this week
- [TPAMI-26] Official Implementation of "Dream to Recall: Imagination-Guided Experience Retrieval for Memory-Persistent Vision-and-Language…☆19Aug 18, 2026Updated last month
- Codes, datasets, and synthetic dataset generator about the paper "LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single…☆17Feb 28, 2026Updated 6 months ago
- [ECCV 2026] Official code repository for : "AgentVLN: Towards Agentic Vision-and-Language Navigation"☆239Aug 31, 2026Updated 2 weeks ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆36Jun 14, 2026Updated 3 months ago
- ☆27Jun 2, 2026Updated 3 months ago
- ☆30Aug 18, 2025Updated last year
- Official implementation of the paper: "ActiveVLN: Towards Active Exploration via Multi-Turn RL in Vision-and-Language Navigation"☆78Feb 11, 2026Updated 7 months ago
- NaVIDA: Vision-Language Navigation with Inverse Dynamics Augmentation☆32Apr 17, 2026Updated 5 months ago
- (CVPR 26) Explore with Long-term Memory: A Benchmark and Multimodal LLM-based Reinforcement Learning Framework for Embodied Exploration☆41Mar 8, 2026Updated 6 months ago
- [IROS 2025] Official implementation of SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigati…☆20Mar 27, 2026Updated 5 months ago
- Breaking Down and Building Up: Mixture of Skill-Based Vision-and-Language Navigation Agents☆34Aug 30, 2026Updated 2 weeks ago
- InternRobotics' open platform for building generalized navigation foundation models.☆1,112Mar 10, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for OctoNav-Bench and OctoNav-R1☆78Apr 29, 2026Updated 4 months ago
- [CVPR 2026] Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment☆28May 11, 2026Updated 4 months ago
- [ECCV 2026] MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots☆117Updated this week
- Official implementation of the paper "ETP-R1: Evolving Topological Planning with Reinforcement Fine-tuning for Vision-Language Navigatio…☆39Dec 25, 2025Updated 8 months ago
- Nav-R1: Reasoning and Navigation in Embodied Scenes☆132Oct 31, 2025Updated 10 months ago
- ☆34May 4, 2026Updated 4 months ago
- ☆34May 13, 2026Updated 4 months ago
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated 2 months ago
- ☆61Aug 18, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICRA 2026] NavDP: Learning Sim-to-Real Navigation Diffusion Policy with Privileged Information Guidance☆792Sep 7, 2026Updated last week
- [ICLR 2026] From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning☆80Updated this week
- [RSS 2026] Official code & data for "OmniNavBench: Beyond Isolation — A Unified Benchmark for General-Purpose Navigation"☆99Aug 9, 2026Updated last month
- Sparse Video Generation Model for Embodied Navigation conditioned on loose language guidance, 100% real world verification☆120Jul 10, 2026Updated 2 months ago
- [ICCV 25] Official repository of "Collaborative Instance Object Navigation: Leveraging Uncertainty-Awareness to Minimize Human-Agent Dial…☆32Apr 1, 2026Updated 5 months ago
- [ICRA 2026] LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision-and-Language Navigation in Continuous Environments☆53Jun 17, 2026Updated 3 months ago
- This repository represents the official implementation of the paper titled "Context-Nav: Context-Driven Exploration and Viewpoint-Aware 3…☆21Jun 23, 2026Updated 2 months ago