Official repo of "FlowWAM: Optical Flow as a Unified Action Representation for World Action Models"
☆28Jul 15, 2026Updated last week
Alternatives and similar repositories for FlowWAM
Users that are interested in FlowWAM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Jul 15, 2026Updated last week
- ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?☆133Jul 9, 2026Updated last week
- [ICCV 2025] Official repo of "EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow"☆27Oct 16, 2025Updated 9 months ago
- Official Repository of "Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads"☆17Oct 6, 2025Updated 9 months ago
- [ICML 2026] OcclusionFormer: Arranging Z-Order for Layout-Grounded Image Generation☆22May 21, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 3 months ago
- OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation☆58Mar 25, 2026Updated 3 months ago
- ☆108Updated this week
- RoboDojo Official Repo☆275Updated this week
- Unofficial Implementation of Training-free Diffusion Model Adaptation for Variable-Sized Text-to-Image Synthesis☆16Sep 27, 2023Updated 2 years ago
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,196Apr 3, 2026Updated 3 months ago
- ☆40Jun 30, 2026Updated 3 weeks ago
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator☆33Apr 15, 2026Updated 3 months ago
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆20Jun 27, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implemetation of the paper "Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising".☆43Jun 10, 2026Updated last month
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆30Apr 26, 2026Updated 2 months ago
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆63May 14, 2026Updated 2 months ago
- [CVPR'25 Highlight] A VQA benchmark for 6D spatial reasoning.☆20Apr 29, 2026Updated 2 months ago
- ☆22Mar 31, 2026Updated 3 months ago
- the official repository of 《ECT: Fine-grained Edge Detection with Learned Cause Tokens》☆16Feb 15, 2024Updated 2 years ago
- [ECCV 2024] QueryCDR: Query-based Controllable Distortion Rectification Network for Fisheye Images☆11Feb 14, 2025Updated last year
- [ICCV 2025 Highlight] PriOr-Flow: Enhancing Primitive Panoramic Optical Flow with Orthogonal View☆19Jul 24, 2025Updated 11 months ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,645Jul 9, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of "Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator"☆16Updated this week
- Learning Visual Feature-Based World Models via Residual Latent Action☆42May 11, 2026Updated 2 months ago
- R3-Avatar: Record and Retrieve Temporal Codebook for Reconstructing Photorealistic Human Avatars☆23Nov 23, 2025Updated 7 months ago
- [RA-L 2026] Official Code of ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning☆45Feb 23, 2026Updated 4 months ago
- [arxiv] BadWAM: When World-Action Models Dream Right but Act Wrong☆43Updated this week
- repository for training action-conditioned latent diffusion world models for robot video generation☆73May 29, 2026Updated last month
- [3DV 2026] Code for KaoLRM: Repurposing Pre-trained Large Reconstruction Models for Parametric 3D Face Reconstruction☆32Mar 17, 2026Updated 4 months ago
- HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation☆51Jun 16, 2026Updated last month
- [CoRL 2025 Oral] ClutterDexGrasp: A Sim-to-Real System for General Dexterous Grasping in Cluttered Scenes☆30Aug 12, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆12Jan 16, 2025Updated last year
- [CVPR 2026] IOMM: Fast Pre-training of Unified Multimodal Models without Text-Image Pairs☆26Apr 11, 2026Updated 3 months ago
- [CVPR 2024] Code and models for pi-ViT, a video transformer for understanding activities of daily living☆31Nov 12, 2025Updated 8 months ago
- [CVPR 2026] Official PyTorch Implementation for "Motion-Aware Animatable Gaussian Avatars Deblurring".☆25May 6, 2026Updated 2 months ago
- Official PyTorch Implementation of Paper "Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures"☆23Jun 30, 2026Updated 3 weeks ago
- Official implementation of MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models☆37Jun 12, 2026Updated last month
- H-OmniStereo: Zero-Shot Omnidirectional Stereo Matching with Heading-Aligned Normal Priors☆50May 15, 2026Updated 2 months ago