[ICCV 2025] LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion
☆302Jul 15, 2025Updated last year
Alternatives and similar repositories for LangScene-X
Users that are interested in LangScene-X are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [CVPR 2025 Highlight] VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step☆353Jul 4, 2025Updated last year
- [CVPR 2026] Official implementation of "MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second".☆462Mar 19, 2026Updated 4 months ago
- The implementation of Extreme Viewpoint 4D Video Generation☆264Sep 6, 2025Updated 10 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆192Mar 10, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- DeepVerse: 4D Autoregressive Video Generation as a World Model☆230Aug 11, 2025Updated 11 months ago
- [ICLR'26] IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction☆427Dec 1, 2025Updated 7 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,095Jul 3, 2026Updated 3 weeks ago
- Cameras as Relative Positional Encoding☆742Dec 18, 2025Updated 7 months ago
- [ICCV 2025] SpatialTrackerV2: 3D Point Tracking Made Easy☆985Feb 27, 2026Updated 5 months ago
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆945Oct 27, 2025Updated 9 months ago
- ☆721May 1, 2025Updated last year
- [ECCV 2026] Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training☆243Jun 19, 2026Updated last month
- [NeurIPS'24] Large Spatial Model: End-to-end Unposed Images to Semantic 3D☆236Feb 11, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR'26] PE3R: Perception-Efficient 3D Reconstruction. Take 2 - 3 photos with your phone, upload them, wait a few minutes, and then star…☆415Feb 28, 2026Updated 5 months ago
- [ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields☆543Oct 31, 2025Updated 8 months ago
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆480Feb 5, 2026Updated 5 months ago
- [SIGGRAPH Asia 2025] WorldExplorer: Towards Generating Fully Navigable 3D Scenes☆190Mar 30, 2026Updated 3 months ago
- 3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding☆414Jul 20, 2026Updated last week
- [ICCV 2025] Official implementation of the paper "DreamCube: 3D Panorama Generation via Multi-plane Synchronization".☆181Feb 4, 2026Updated 5 months ago
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆485Apr 16, 2026Updated 3 months ago
- [ICCV 2025] InstaScene: Towards Complete 3D Instance Decomposition and Reconstruction from Cluttered Scenes☆134Apr 5, 2026Updated 3 months ago
- [ICCV 2025 ⭐highlight⭐] Implementation of VMem: Consistent Interactive Video Scene Generation with Surfel-Indexed View Memory☆442Jul 25, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR2026] Official Implementation of "Dens3R: A Foundation Model for 3D Geometry Prediction"☆395May 14, 2026Updated 2 months ago
- Code of WinT3R: Window-Based Streaming Rrconstruction With Camera Token Pool☆229Mar 4, 2026Updated 4 months ago
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆437Jun 6, 2025Updated last year
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆431Jul 15, 2026Updated 2 weeks ago
- [ECCV 2024] Make-Your-3D: Fast and Consistent Subject-Driven 3D Content Generation☆128Mar 15, 2024Updated 2 years ago
- [ICCV 2025 Oral] SceneSplat - Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining☆355May 25, 2026Updated 2 months ago
- [ICLR 2026] A simple state update rule to enhance length generalization for CUT3R☆711May 11, 2026Updated 2 months ago
- [NeurIPS 2025 (Spotlight)] The implementation for the paper "4DGT Learning a 4D Gaussian Transformer Using Real-World Monocular Videos"☆469Sep 19, 2025Updated 10 months ago
- [CVPR 2026 Highlight & Best Paper of VideoWorldModel Workshop] NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos☆643May 12, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025 Highlight] GEN3C: 3D-Informed World-Consistent Video Generation with Precise Camera Control☆1,389Jun 15, 2026Updated last month
- Official implementation of Continuous 3D Perception Model with Persistent State☆1,470Aug 27, 2025Updated 11 months ago
- [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction☆366Mar 20, 2026Updated 4 months ago
- [ICCV 2025] This is the official implementation of POMATO: Marrying Pointmap Matching with Temporal Motions for Dynamic 3D Reconstruction☆124Aug 9, 2025Updated 11 months ago
- [CVPR2025] Prometheus: 3D-Aware Latent Diffusion Models for Feed-Forward Text-to-3D Scene Generation☆148Jul 5, 2025Updated last year
- [NeurIPS 2025 Spotlight] Official implementation of the SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature Alig…☆164Sep 25, 2025Updated 10 months ago
- [CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations☆589Apr 22, 2026Updated 3 months ago