ICCV 2025 | TesserAct: Learning 4D Embodied World Models
☆408Aug 4, 2025Updated last year
Alternatives and similar repositories for TesserAct
Users that are interested in TesserAct are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆609Oct 26, 2025Updated 10 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 8 months ago
- [ICLR’26] Learning Video Generation for Robotic Manipulation with Collaborative Trajectory Control☆107Feb 8, 2026Updated 7 months ago
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆409Aug 21, 2026Updated 2 weeks ago
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆437Jun 6, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆496Apr 16, 2026Updated 4 months ago
- [ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields☆572Oct 31, 2025Updated 10 months ago
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆967Oct 27, 2025Updated 10 months ago
- ☆259Jan 31, 2026Updated 7 months ago
- Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals☆2,638Apr 19, 2026Updated 4 months ago
- Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io☆410May 17, 2025Updated last year
- [ICCV 2025] SpatialTrackerV2: 3D Point Tracking Made Easy☆1,001Feb 27, 2026Updated 6 months ago
- DeepVerse: 4D Autoregressive Video Generation as a World Model☆230Aug 11, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official implementation of Continuous 3D Perception Model with Persistent State☆1,490Aug 27, 2025Updated last year
- ICLR 2026 Paper: Ctrl-World☆562Apr 8, 2026Updated 5 months ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,870Jul 9, 2026Updated 2 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,151Jul 3, 2026Updated 2 months ago
- [ICCV 2025] LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion☆302Jul 15, 2025Updated last year
- ☆581Jun 24, 2026Updated 2 months ago
- [ICLR 2025] SPA: 3D Spatial-Awareness Enables Effective Embodied Representation☆179Jun 19, 2025Updated last year
- [ICML 2024] 3D-VLA: A 3D Vision-Language-Action Generative World Model☆635Oct 29, 2024Updated last year
- [CVPR 2025] Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video☆228May 25, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆74Jun 9, 2026Updated 3 months ago
- Cameras as Relative Positional Encoding☆752Dec 18, 2025Updated 8 months ago
- RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning☆1,859Updated this week
- [NeurIPS 2025] WorldMem: Long-term Consistent World Simulation with Memory☆392Feb 21, 2026Updated 6 months ago
- [ICLR 2025] LAPA: Latent Action Pretraining from Videos☆568Jan 22, 2025Updated last year
- Feed-forward Geometry, Appearance and Camera Estimation from Uncalibrated Sparse Views☆725May 1, 2025Updated last year
- Stereo4D dataset and processing code☆312Nov 4, 2025Updated 10 months ago
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆152Nov 4, 2025Updated 10 months ago
- Official Implementation of paper "MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion"☆1,387Jun 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆486Feb 5, 2026Updated 7 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆196Mar 10, 2026Updated 6 months ago
- Cosmos-Predict2.5, the latest version of the Cosmos World Foundation Models (WFMs) family, specialized for simulating and predicting the …☆1,365Jun 8, 2026Updated 3 months ago
- A curated list of awesome papers for reconstructing 4D spatial intelligence from video. (arXiv 2507.21045)☆523Updated this week
- [CVPR 2026] Official implementation of "MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second".☆471Mar 19, 2026Updated 5 months ago
- ☆139May 13, 2025Updated last year
- [CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations☆600Apr 22, 2026Updated 4 months ago