FlowWM stochastic world modeling via flow matching in DINOv3 feature space, with the FuturePerception (Waymo) benchmark.
☆81Sep 9, 2026Updated this week
Alternatives and similar repositories for Flow-World-Models
Users that are interested in Flow-World-Models are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Apr 17, 2026Updated 4 months ago
- ☆13Nov 1, 2023Updated 2 years ago
- Official implementation paper OmniX: Any-view and Any-time 4D Reconstruction via Feed-forward Trajectory Fields☆127Jul 12, 2026Updated 2 months ago
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Sep 5, 2026Updated last week
- A set of tools and examples for converting and utilizing powerful vision models, DINOv3 and EdgeTAM (SAM2), within the ONNX ecosystem.☆15Nov 5, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2026 Oral] Learning to Drive via Real-World Simulation at Scale☆323Updated this week
- ☆16Sep 30, 2025Updated 11 months ago
- Implementation of Bridging Scene Generation and Planning: Driving with World Model via Unifying Vision and Motion Representation☆78Aug 3, 2026Updated last month
- Video Depth Propagation [3DV 2026]☆39Jan 23, 2026Updated 7 months ago
- Temporal Self-imitation Learning☆17Jul 3, 2026Updated 2 months ago
- Code for "Steerable Scene Generation with Post Training and Inference-Time Search", CoRL 2025☆88Oct 11, 2025Updated 11 months ago
- ☆24May 27, 2026Updated 3 months ago
- [CVPR 2026 Highlight] A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens☆258Jul 17, 2026Updated last month
- [ICRA 2026] YOPO: A Minimalist’s Detection Transformer for Monocular RGB Category‑level 9D Multi‑Object Pose Estimation☆25Mar 12, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12May 7, 2019Updated 7 years ago
- [CVPR 2026] DrivePI: Spatial-aware 4D MLLM for Unified Autonomous Driving Understanding, Perception, Prediction and Planning☆133Mar 21, 2026Updated 5 months ago
- OSCAR — public inference release☆75Jun 16, 2026Updated 2 months ago
- Apple's Cut Cross Entropy☆36Jan 19, 2025Updated last year
- Exploration into some new research surrounding value networks☆16Feb 10, 2026Updated 7 months ago
- AMoE: Agglomerative Mixture-of-Experts Vision Foundation Models☆59Jun 11, 2026Updated 3 months ago
- Code for Learning Barrier Certificates: Towards Safe Reinforcement Learning with Zero Training-time Violations☆18Mar 4, 2022Updated 4 years ago
- From RGB Generation to Dense Field Readout: Pixel-Space Dense Prediction with Text-to-Image Models (ReChannel)☆40Aug 17, 2026Updated 3 weeks ago
- Code for RepWAM: World Action Modeling with Representation Visual-Action Tokenizers☆66Aug 4, 2026Updated last month
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- UniDriveVLA: Unifying Understanding, Perception, and Action Planning for Autonomous Driving☆241Updated this week
- [ECCV 2024]FipTR: A Simple yet Effective Transformer Framework for Future Instance Prediction in Autonomous Driving☆26Aug 3, 2024Updated 2 years ago
- A unified framework for feed-forward neural networks☆15Nov 28, 2025Updated 9 months ago
- This is the official repo for Do LLM Modules Generalize? A Study on Motion Generation for Autonomous Driving. CoRL 2025☆21Oct 20, 2025Updated 10 months ago
- Zero-Shot Multi-Object Shape Completion (ECCV 2024)☆32Apr 1, 2025Updated last year
- Our inference and training framework to run on the Cosmos Models☆525Updated this week
- Code for paper Feasible Actor-Critic: Constrained Reinforcement Learning for Ensuring Statewise Safety.☆20May 22, 2022Updated 4 years ago
- LSRM is a SOTA, feed-forward 3D reconstruction model that generates high-fidelity, relightable 3D digital twins from sparse 2D views.☆72Jun 5, 2026Updated 3 months ago
- ☆140Oct 21, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026] Official repository of "StableMTL: Repurposing Latent Diffusion Models for Multi-Task Learning from Partially Annotated Synth…☆18Feb 21, 2026Updated 6 months ago
- Skill-based Teleoperation☆51Dec 4, 2025Updated 9 months ago
- [ECCV 26 Spotlight] Official code for Zero-Shot Depth from Defocus (https://arxiv.org/abs/2603.26658)☆57Updated this week
- [CoRL 2024 Oral] FREA: Feasibility-Guided Generation of Safety-Critical Scenarios with Reasonable Adversariality☆66May 5, 2026Updated 4 months ago
- ☆23Sep 1, 2022Updated 4 years ago
- Embedding and readout for simple multi-categorical and gaussian continuous☆20Jul 5, 2026Updated 2 months ago
- Offical implementation of CVPR 2026 paper SpaceDrive: Infusing Spatial Awareness into VLM-based Autonomous Driving.☆92May 19, 2026Updated 3 months ago