☆43Jun 30, 2026Updated last month
Alternatives and similar repositories for Flash-WAM
Users that are interested in Flash-WAM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Vision-Language-Action Optimization with Trajectory Ensemble Voting (ICANN2026)☆26Feb 18, 2026Updated 5 months ago
- A comprehensive collection of Vision-Language-Action (VLA) models, benchmarks, and datasets for robotic manipulation and embodied AI res…☆18Nov 5, 2025Updated 8 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 9 months ago
- repository for training action-conditioned latent diffusion world models for robot video generation☆74May 29, 2026Updated 2 months ago
- [ICLR 2025] "Understanding Constraint Inference in Safety-Critical Inverse Reinforcement Learning"☆16Nov 30, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?☆146Updated this week
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,237Apr 3, 2026Updated 4 months ago
- The official implementation of World-Language-Action Model for Unified World Modeling, Language Reasoning, and Action Synthesis☆126Jun 18, 2026Updated last month
- [RSS 2026] Causal video-action world model for generalist robot control☆1,718Jul 9, 2026Updated 3 weeks ago
- [IEEE TPAMI] Code for the paper "Aligning Few-Step Diffusion Models with Dense Reward Difference Learning"☆22Feb 25, 2026Updated 5 months ago
- Official repo of "FlowWAM: Optical Flow as a Unified Action Representation for World Action Models"☆42Jul 15, 2026Updated 2 weeks ago
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆63May 14, 2026Updated 2 months ago
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 2 months ago
- Official repo for "StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation"☆30Jun 29, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆32May 8, 2026Updated 2 months ago
- ☆99Jul 15, 2026Updated 2 weeks ago
- [ICML 2026] 🏂 World Guidance: World Modeling in Condition Space for Action Generation☆162Apr 28, 2026Updated 3 months ago
- A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models☆756Jun 15, 2026Updated last month
- A curated list of academic papers and resources on Vision-Language-Action (VLA) and World Action Models (WAM)☆29Updated this week
- Flow Map OPD for AnyStep Video Diffusion☆404May 23, 2026Updated 2 months ago
- ☆100Jun 2, 2026Updated 2 months ago
- A curated, continuously updated reading list, paper blogs, and resources for World Action Models (WAMs) in embodied AI.☆1,213Updated this week
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆31Apr 26, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICRA 2026] SOE: Sample-Efficient Robot Policy Self-Improvement via On-Manifold Exploration☆18Mar 2, 2026Updated 5 months ago
- Official code of Motus: A Unified Latent Action World Model☆1,220Jan 5, 2026Updated 6 months ago
- OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation☆59Mar 25, 2026Updated 4 months ago
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆28Jun 1, 2026Updated 2 months ago
- CGGS: Consistency-Augmented Geometric Gaussian Splatting for Ego-Centric 3D Scene Generation (TIP 2026)☆24Jul 23, 2026Updated last week
- ☆15Nov 21, 2025Updated 8 months ago
- ICLR 2026☆44May 29, 2026Updated 2 months ago
- [ECCV 2026] Demystifying Video Reasoning☆47Jul 14, 2026Updated 3 weeks ago
- PhysisForcing: Physics Reinforced World Simulator for Robotic Manipulation☆114Jul 11, 2026Updated 3 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of Forge4D: Feed-Forward 4D Human Reconstruction and Interpolation from Uncalibrated Sparse Videos☆53May 2, 2026Updated 3 months ago
- ☆11Nov 27, 2025Updated 8 months ago
- A comprehensive study of related works and research around Google's Genie 3 model - a new frontier for world models☆15Aug 18, 2025Updated 11 months ago
- ☆139Mar 31, 2026Updated 4 months ago
- Unofficial OpenPI extension experiment to build a more complete OpenPI-style VLA engineering stack: pi0.5 semantics, RTC, pi0.6 RECAP/MEM…☆17Jul 8, 2026Updated 3 weeks ago
- TVRBench: Target Viewpoint Reproduction Benchmark for Active Spatial Intelligence☆26Jun 2, 2026Updated 2 months ago
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 6 months ago