ICCV 2025 | TesserAct: Learning 4D Embodied World Models
☆405Aug 4, 2025Updated last year
Alternatives and similar repositories for TesserAct
Users that are interested in TesserAct are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆608Oct 26, 2025Updated 9 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 7 months ago
- [ICLR’26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control☆108Feb 8, 2026Updated 6 months ago
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆405Jul 23, 2025Updated last year
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆438Jun 6, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆490Apr 16, 2026Updated 4 months ago
- [ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields☆563Oct 31, 2025Updated 9 months ago
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆957Oct 27, 2025Updated 9 months ago
- ☆257Jan 31, 2026Updated 6 months ago
- Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals☆2,572Apr 19, 2026Updated 4 months ago
- Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io☆407May 17, 2025Updated last year
- [ICCV 2025] SpatialTrackerV2: 3D Point Tracking Made Easy☆994Feb 27, 2026Updated 5 months ago
- DeepVerse: 4D Autoregressive Video Generation as a World Model☆230Aug 11, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official implementation of Continuous 3D Perception Model with Persistent State☆1,484Aug 27, 2025Updated 11 months ago
- ICLR 2026 Paper: Ctrl-World☆553Apr 8, 2026Updated 4 months ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,784Jul 9, 2026Updated last month
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,115Jul 3, 2026Updated last month
- [ICCV 2025] LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion☆302Jul 15, 2025Updated last year
- ☆575Jun 24, 2026Updated last month
- [ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation☆179Jun 19, 2025Updated last year
- [ICML 2024] 3D-VLA: A 3D Vision-Language-Action Generative World Model☆630Oct 29, 2024Updated last year
- [CVPR 2025] Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video☆227May 25, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆71Jun 9, 2026Updated 2 months ago
- Cameras as Relative Positional Encoding☆746Dec 18, 2025Updated 8 months ago
- RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning☆1,808Updated this week
- [NeurIPS 2025] WorldMem: Long-term Consistent World Simulation with Memory☆385Feb 21, 2026Updated 6 months ago
- ☆723May 1, 2025Updated last year
- Stereo4D dataset and processing code☆312Nov 4, 2025Updated 9 months ago
- [ICLR 2025] LAPA: Latent Action Pretraining from Videos☆566Jan 22, 2025Updated last year
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆151Nov 4, 2025Updated 9 months ago
- Official Implementation of paper "MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion"☆1,383Jun 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆484Feb 5, 2026Updated 6 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆194Mar 10, 2026Updated 5 months ago
- Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the …☆1,351Jun 8, 2026Updated 2 months ago
- A curated list of awesome papers for reconstructing 4D spatial intelligence from video. (arXiv 2507.21045)☆516Jun 5, 2026Updated 2 months ago
- [CVPR 2026] Official implementation of "MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second".☆469Mar 19, 2026Updated 5 months ago
- ☆139May 13, 2025Updated last year
- [CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations☆596Apr 22, 2026Updated 3 months ago