ICCV 2025 | TesserAct: Learning 4D Embodied World Models
☆409Aug 4, 2025Updated last year
Alternatives and similar repositories for TesserAct
Users that are interested in TesserAct are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆612Oct 26, 2025Updated 11 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆125Jan 10, 2026Updated 8 months ago
- [ICLR’26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control☆107Feb 8, 2026Updated 7 months ago
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆414Aug 21, 2026Updated last month
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆437Jun 6, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆498Apr 16, 2026Updated 5 months ago
- [ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields☆579Oct 31, 2025Updated 11 months ago
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆975Oct 27, 2025Updated 11 months ago
- ☆261Jan 31, 2026Updated 8 months ago
- Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals☆2,675Apr 19, 2026Updated 5 months ago
- Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io☆413May 17, 2025Updated last year
- [ICCV 2025] SpatialTrackerV2: 3D Point Tracking Made Easy☆1,002Feb 27, 2026Updated 7 months ago
- Official implementation of Continuous 3D Perception Model with Persistent State☆1,497Aug 27, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- DeepVerse: 4D Autoregressive Video Generation as a World Model☆233Aug 11, 2025Updated last year
- ICLR 2026 Paper: Ctrl-World☆568Apr 8, 2026Updated 5 months ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,918Jul 9, 2026Updated 2 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,182Jul 3, 2026Updated 2 months ago
- [ICCV 2025] LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion☆303Jul 15, 2025Updated last year
- ☆586Sep 10, 2026Updated 2 weeks ago
- [ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation☆179Jun 19, 2025Updated last year
- [ICML 2024] 3D-VLA: A 3D Vision-Language-Action Generative World Model☆638Oct 29, 2024Updated last year
- [CVPR 2025] Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video☆228May 25, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆73Jun 9, 2026Updated 3 months ago
- Cameras as Relative Positional Encoding☆758Dec 18, 2025Updated 9 months ago
- RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning☆1,866Updated this week
- [NeurIPS 2025] WorldMem: Long-term Consistent World Simulation with Memory☆397Feb 21, 2026Updated 7 months ago
- [ICLR 2025] LAPA: Latent Action Pretraining from Videos☆577Jan 22, 2025Updated last year
- Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views☆727May 1, 2025Updated last year
- Stereo4D dataset and processing code☆312Nov 4, 2025Updated 10 months ago
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆153Nov 4, 2025Updated 10 months ago
- Official Implementation of paper "MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion"☆1,388Jun 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆491Feb 5, 2026Updated 7 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆199Mar 10, 2026Updated 6 months ago
- Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the …☆1,373Jun 8, 2026Updated 3 months ago
- [CVPR 2026] Official implementation of "MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second".☆482Mar 19, 2026Updated 6 months ago
- ☆140May 13, 2025Updated last year
- A curated list of awesome papers for reconstructing 4D spatial intelligence from video. (arXiv 2507.21045)☆528Sep 9, 2026Updated 3 weeks ago
- [CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations☆601Apr 22, 2026Updated 5 months ago