[CVPR 2026] This repository is the official implementation of MVGGT: Multimodal Visual Geometry Grounded Transformer for Multiview 3D Referring Expression Segmentation
☆128Mar 24, 2026Updated 4 months ago
Alternatives and similar repositories for mvggt
Users that are interested in mvggt are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR'26] IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction☆427Dec 1, 2025Updated 7 months ago
- Official implementation of “4D LangVGGT: 4D Language-Visual Geometry Grounded Transformer”☆91Mar 25, 2026Updated 4 months ago
- [ICML2025 Oral] ReferSplat: Referring Segmentation in 3D Gaussian Splatting☆146May 26, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Official implementation of the SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature Alig…☆164Sep 25, 2025Updated 10 months ago
- [CVPR2026] Generalizing Visual Geometry Priors to Sparse Gaussian Occupancy Prediction☆40Jul 16, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Spa3R: Predictive Spatial Field Modeling for 3D Visual Reasoning☆51Mar 25, 2026Updated 4 months ago
- [AAAI 2026] Official implementation of the paper ”SegDINO3D: 3D Instance Segmentation Empowered by Both Image-Level and Object-Level 2D F…☆67Updated this week
- [ICCV 2025] SAS: Segment Any 3D Scene with Integrated 2D Priors☆38Jun 25, 2025Updated last year
- M³: Dense Matching Meets Multi-View Foundation Models for Monocular Gaussian Splatting SLAM☆83Mar 18, 2026Updated 4 months ago
- [CVPR'26 Highlight] AMB3R: Accurate Feed-forward Metric-scale 3D Reconstruction with Backend☆471Jun 5, 2026Updated last month
- PanSt3R: Multi-view Consistent Panoptic Segmentation (official code)☆80Mar 20, 2026Updated 4 months ago
- Code implementation of Pi-Long☆193Apr 16, 2026Updated 3 months ago
- The official implementation of InfiniteVGGT☆379Apr 19, 2026Updated 3 months ago
- [CVPR 2026 Hightlight] OmniVGGT: Omni-Modality Driven Visual Geometry Grounded Transformer☆353May 21, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR'25] Official repository for "Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration"☆105Jun 10, 2025Updated last year
- [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction☆366Mar 20, 2026Updated 4 months ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆431Jul 15, 2026Updated 2 weeks ago
- [CVPR 2026] Official code of "EmbodiedSplat: Online Feed-Forward Semantic 3DGS for Open-Vocabulary 3D Scene Understanding"☆109Jul 17, 2026Updated last week
- Official code for CVPR 2026 paper: VGGT-Det: Mining VGGT Internal Priors for Sensor-Geometry-Free Multi-View Indoor 3D Object Detection☆145Jul 15, 2026Updated 2 weeks ago
- [CVPR 2025] Insightful Instance Features for 3D Instance Segmentation☆18May 20, 2026Updated 2 months ago
- [SIGGRAPH Asia 2025 (ACM TOG)] AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views☆899Dec 22, 2025Updated 7 months ago
- ☆35Nov 17, 2025Updated 8 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,094Jul 3, 2026Updated 3 weeks ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer☆806Jan 28, 2026Updated 6 months ago
- Geometry-grounded Point Transformer (CVPR 2026)☆145May 6, 2026Updated 2 months ago
- [CVPR 2025] OmniSplat: Taming Feed-Forward 3D Gaussian Splatting for Omnidirectional Images with Editable Capabilities☆39Jun 6, 2025Updated last year
- Official implement of VGGT-Long☆886Mar 20, 2026Updated 4 months ago
- [CVPR25 Highlight] Official implementation of Fun3DU, a method for functional understanding and segmentation in 3D scenes☆51Sep 30, 2025Updated 10 months ago
- [CVPR 2026 (Highlight)] Scal3R: Scalable Test-Time Training for Large-Scale 3D Reconstruction☆515May 11, 2026Updated 2 months ago
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆946Oct 27, 2025Updated 9 months ago
- [NeurIPS 2024] OpenGaussian: Towards Point-Level 3D Gaussian-based Open Vocabulary Understanding☆217May 27, 2025Updated last year
- The official implementation of the paper “VGGT4D: Mining Motion Cues in Visual Geometry Transformers for 4D Scene Reconstruction.”☆268Dec 2, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR'2026, Oral] FUSER: Feed-Forward MUltiview 3D Registration Transformer and SE(3)^N Diffusion Refinement☆38Jun 4, 2026Updated last month
- ☆17Dec 25, 2025Updated 7 months ago
- TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction☆347Jun 12, 2026Updated last month
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆347Apr 18, 2026Updated 3 months ago
- [CVPR 2026 Highlight] Unified 3D Reconstruction and Semantic Understanding via Generalizable Gaussian Splatting from Unposed Multi-View I…☆200Apr 10, 2026Updated 3 months ago
- [NeurIPS 2025] AutoSeg3D, online real-time 3D segmentation as instance tracking with long-short term query memory for embodied perception☆56Dec 18, 2025Updated 7 months ago
- ☆15Nov 25, 2025Updated 8 months ago