☆25Apr 30, 2026Updated 3 months ago
Alternatives and similar repositories for World2VLM
Users that are interested in World2VLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Sep 3, 2025Updated 11 months ago
- S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence☆81Jul 22, 2026Updated last week
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆21Mar 18, 2026Updated 4 months ago
- ☆22Jul 26, 2026Updated last week
- Code for ICCV 2023 work "Generalized Few-Shot Point Cloud Segmentation Via Geometric Words"☆14Sep 26, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆15Apr 3, 2026Updated 4 months ago
- [ECCV 2026🔥] Official code repository for "Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors"☆53Jun 23, 2026Updated last month
- [ICCV 2023] ImGeoNet: Image-induced Geometry-aware Voxel Representation for Multi-view 3D Object Detection☆19Sep 12, 2024Updated last year
- Official PyTorch implementation of CorrespondentDream: Enhancing 3D Fidelity of Text-to-3D using Cross-View Correspondences (CVPR 2024 Po…☆19Apr 29, 2024Updated 2 years ago
- Code for NeurIPS 2024 work "MVSDet: Multi-View Indoor 3D Object Detection via Efficient Plane Sweeps"☆17Dec 11, 2024Updated last year
- ☆27Jun 5, 2025Updated last year
- ☆22Apr 3, 2026Updated 4 months ago
- Geometry Meets Vision: Revisiting Pretrained Semantics in Distilled Fields☆32Oct 3, 2025Updated 10 months ago
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆74Apr 21, 2026Updated 3 months ago
- SPAgent, a foundation agent for understanding, reasoning over, and operating within the physical and spatial world.☆211Jul 25, 2026Updated last week
- A Deep Reinforcement Learning Strategy and Framework for Floating Waste Capture☆13Mar 13, 2025Updated last year
- [Preprint] ViFeEdit: A Video-Free Tuner of Your Video Diffusion Transformer☆67Mar 31, 2026Updated 4 months ago
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆95Jul 13, 2025Updated last year
- repository for "Exploiting Proximity-Aware Tasks for Embodied Social Navigation" paper code☆12Nov 16, 2023Updated 2 years ago
- ☆10Apr 8, 2024Updated 2 years ago
- [ICML 2026] MLLM-4D: Towards Visual-based Spatial-Temporal Intelligence☆38May 1, 2026Updated 3 months ago
- [CVPR 2026 Fingdings] This repo is the official implementation of "Euclid’s Gift: Enhancing Spatial Perception and Reasoning in Vision‑La…☆28Mar 15, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2024] DiffSF: Diffusion Models for Scene Flow Estimation☆30Jan 9, 2025Updated last year
- SimpleVQA: Multimodal Factuality Evaluation for Multimodal Large Language Models☆15Feb 20, 2025Updated last year
- Keras Functional API for multiple inputs and mixed data☆11Feb 18, 2019Updated 7 years ago
- ☆36May 21, 2026Updated 2 months ago
- Search framework for CAD files using the frequency domain☆13May 3, 2024Updated 2 years ago
- Code for ICRA'25 paper: "Differentiable Composite Neural Signed Distance Fields for Robot Navigation in Dynamic Indoor Environments"☆16Apr 2, 2025Updated last year
- ☆17Jun 13, 2025Updated last year
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆19Feb 1, 2026Updated 6 months ago
- Basic codes of ml☆13Dec 2, 2019Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 3 months ago
- Self-supervised representation learning for CAD model☆13May 7, 2025Updated last year
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆16May 21, 2026Updated 2 months ago
- Learning Descriptive Image Captioning via Semipermeable Maximum Likelihood Estimation (NeurIPS 2023)☆23Oct 1, 2023Updated 2 years ago
- VHTest☆16Oct 31, 2024Updated last year
- [CVPR 2025] Official Repository of the paper "On the Consistency of Video Large Language Models in Temporal Comprehension"☆16Oct 13, 2025Updated 9 months ago
- [TMM'25]Geometry-Aware 3D Gaussian Representation for Real-Time Rendering of Large-Scale Scenes☆17Nov 18, 2024Updated last year