Official repo of "FlowWAM: Optical Flow as a Unified Action Representation for World Action Models"
☆46Jul 15, 2026Updated 3 weeks ago
Alternatives and similar repositories for FlowWAM
Users that are interested in FlowWAM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆25Jul 15, 2026Updated 3 weeks ago
- ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?☆153Jul 30, 2026Updated 2 weeks ago
- [ICCV 2025] Official repo of "EC-Flow: Enabling Versatile Robotic Manipulation from Action-Unlabeled Videos via Embodiment-Centric Flow"☆27Oct 16, 2025Updated 9 months ago
- Official Repository of "Fibottention: Inceptive Visual Representation Learning with Diverse Attention Across Heads"☆17Oct 6, 2025Updated 10 months ago
- [ICML 2026] OcclusionFormer: Arranging Z-Order for Layout-Grounded Image Generation☆22May 21, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 4 months ago
- OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation☆59Updated this week
- (ECCV2024) KFD-NeRF: Rethinking Dynamic NeRF with Kalman Filter☆19Jun 27, 2025Updated last year
- Involving over 40 Advanced Manipulation Policies☆151Updated this week
- RoboDojo Official Repo☆379Updated this week
- Unofficial Implementation of Training-free Diffusion Model Adaptation for Variable-Sized Text-to-Image Synthesis☆16Sep 27, 2023Updated 2 years ago
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,288Apr 3, 2026Updated 4 months ago
- ☆43Jun 30, 2026Updated last month
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator☆34Apr 15, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Claude Code skill: read & write Overleaf projects via the git bridge. Works on Mac/Linux/WSL.☆24May 7, 2026Updated 3 months ago
- Official implemetation of the paper "Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising".☆52Jun 10, 2026Updated 2 months ago
- [ICRA 2026 🥳] GarmentPile++: Affordance-Driven Cluttered Garments Retrieval with Vision-Language Reasoning☆18Mar 6, 2026Updated 5 months ago
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆31Apr 26, 2026Updated 3 months ago
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆64May 14, 2026Updated 2 months ago
- [CVPR'25 Highlight] A VQA benchmark for 6D spatial reasoning.☆20Apr 29, 2026Updated 3 months ago
- Official implementation of paper "GAPrompt: Geometry-Aware Point Cloud Prompt for 3D Vision Model", ICML 2025☆17Dec 25, 2025Updated 7 months ago
- ☆23Mar 31, 2026Updated 4 months ago
- the official repository of 《ECT: Fine-grained Edge Detection with Learned Cause Tokens》☆16Feb 15, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ECCV 2024] QueryCDR: Query-based Controllable Distortion Rectification Network for Fisheye Images☆11Feb 14, 2025Updated last year
- [ICCV 2025 Highlight] PriOr-Flow: Enhancing Primitive Panoramic Optical Flow with Orthogonal View☆19Jul 24, 2025Updated last year
- Research sources on graph-based anomaly detection☆13Nov 29, 2022Updated 3 years ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,749Jul 9, 2026Updated last month
- Official implementation of "Image2Sim: Scaling Embodied Navigation via Generative Neural Simulator"☆27Jul 17, 2026Updated 3 weeks ago
- Learning Visual Feature-Based World Models via Residual Latent Action☆44May 11, 2026Updated 3 months ago
- R3-Avatar: Record and Retrieve Temporal Codebook for Reconstructing Photorealistic Human Avatars☆23Nov 23, 2025Updated 8 months ago
- [RA-L 2026] Official Code of ImplicitRDP: An End-to-End Visual-Force Diffusion Policy with Structural Slow-Fast Learning☆50Aug 2, 2026Updated last week
- repository for training action-conditioned latent diffusion world models for robot video generation☆76Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- HorizonDrive: Self-Corrective Autoregressive World Model for Long-horizon Driving Simulation☆53Jun 16, 2026Updated last month
- [CoRL 2025 Oral] ClutterDexGrasp: A Sim-to-Real System for General Dexterous Grasping in Cluttered Scenes☆30Aug 12, 2025Updated last year
- ☆12Jan 16, 2025Updated last year
- [CVPR 2026] Official PyTorch Implementation for "Motion-Aware Animatable Gaussian Avatars Deblurring".☆27May 6, 2026Updated 3 months ago
- Official PyTorch Implementation of Paper "Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures"☆27Jun 30, 2026Updated last month
- Official implementation of MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models☆40Jun 12, 2026Updated 2 months ago
- Repository for WACV23 paper "Automatically Annotating Indoor Images with CAD Models via RGB-D Scans"☆18Aug 5, 2025Updated last year