[CVPR 2026] This repository is the official implementation of MVGGT: Multimodal Visual Geometry Grounded Transformer for Multiview 3D Referring Expression Segmentation
☆134Mar 24, 2026Updated 5 months ago
Alternatives and similar repositories for mvggt
Users that are interested in mvggt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR'26] IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction☆439Dec 1, 2025Updated 9 months ago
- Official implementation of “4D LangVGGT: 4D Language-Visual Geometry Grounded Transformer”☆92Mar 25, 2026Updated 5 months ago
- [ICML2025 Oral] ReferSplat: Referring Segmentation in 3D Gaussian Splatting☆148May 26, 2026Updated 3 months ago
- [NeurIPS 2025 Spotlight] Official implementation of the SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature Alig…☆164Sep 25, 2025Updated 11 months ago
- [CVPR2026] Generalizing Visual Geometry Priors to Sparse Gaussian Occupancy Prediction☆44Jul 16, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Spa3R: Predictive Spatial Field Modeling for 3D Visual Reasoning☆53Mar 25, 2026Updated 5 months ago
- [AAAI 2026] Official implementation of the paper ”SegDINO3D: 3D Instance Segmentation Empowered by Both Image-Level and Object-Level 2D F…☆72Jul 29, 2026Updated last month
- [ICCV 2025] SAS: Segment Any 3D Scene with Integrated 2D Priors☆40Jun 25, 2025Updated last year
- Supercharge your AI agents by versioning, tracking, and merging overlapping skills.☆41Apr 9, 2026Updated 5 months ago
- M³: Dense Matching Meets Multi-View Foundation Models for Monocular Gaussian Splatting SLAM☆87Mar 18, 2026Updated 5 months ago
- [CVPR'26 Highlight] AMB3R: Accurate Feed-forward Metric-scale 3D Reconstruction with Backend☆489Jun 5, 2026Updated 3 months ago
- PanSt3R: Multi-view Consistent Panoptic Segmentation (official code)☆89Mar 20, 2026Updated 5 months ago
- Code implementation of Pi-Long☆194Apr 16, 2026Updated 4 months ago
- The official implementation of InfiniteVGGT☆387Apr 19, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2026 Hightlight] OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Transformer☆371May 21, 2026Updated 3 months ago
- [CVPR'25] Official repository for "Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration"☆107Jun 10, 2025Updated last year
- [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction☆377Mar 20, 2026Updated 5 months ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆445Jul 15, 2026Updated last month
- [CVPR 2026] Official code of "EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding"☆119Jul 17, 2026Updated last month
- Official code for CVPR 2026 paper: VGGT-Det: Mining VGGT Internal Priors for Sensor-Geometry-Free Multi-View Indoor 3D Object Detection☆160Jul 15, 2026Updated last month
- [CVPR 2025] Insightful Instance Features for 3D Instance Segmentation☆20May 20, 2026Updated 3 months ago
- [SIGGRAPH Asia 2025 (ACM TOG)] AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views☆927Dec 22, 2025Updated 8 months ago
- ☆35Nov 17, 2025Updated 9 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,152Jul 3, 2026Updated 2 months ago
- 启智平台任务管理 CLI:资源查询、任务提交、日志查看和 MCP/agent workflow☆130Updated this week
- [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer☆813Jan 28, 2026Updated 7 months ago
- Geometry-grounded Point Transformer (CVPR 2026)☆150May 6, 2026Updated 4 months ago
- [CVPR 2025] OmniSplat: Taming Feed-Forward 3D Gaussian Splatting for Omnidirectional Images with Editable Capabilities☆40Jun 6, 2025Updated last year
- Official implement of VGGT-Long☆903Mar 20, 2026Updated 5 months ago
- [CVPR25 Highlight] Official implementation of Fun3DU, a method for functional understanding and segmentation in 3D scenes☆53Sep 30, 2025Updated 11 months ago
- [CVPR 2026 (Highlight)] Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction☆532May 11, 2026Updated 4 months ago
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆966Oct 27, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2024] OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding☆220May 27, 2025Updated last year
- The official implementation of the paper “VGGT4D: Mining Motion Cues in Visual Geometry Transformers for 4D Scene Reconstruction.”☆275Dec 2, 2025Updated 9 months ago
- ☆18Dec 25, 2025Updated 8 months ago
- TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction☆355Jun 12, 2026Updated 2 months ago
- [CVPR'2026, Oral] FUSER: Feed-Forward MUltiview 3D Registration Transformer and SE(3)^N Diffusion Refinement☆44Sep 2, 2026Updated last week
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆354Apr 18, 2026Updated 4 months ago
- [CVPR 2026 Highlight] Unified 3D Reconstruction and Semantic Understanding via Generalizable Gaussian Splatting from Unposed Multi-View I…☆207Apr 10, 2026Updated 5 months ago