ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple and have a scalable dataset.
☆53Nov 27, 2025Updated 9 months ago
Alternatives and similar repositories for ENACT
Users that are interested in ENACT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- THEORY OF SPACE: a benchmark for evaluating whether foundation models can actively explore under partial observability efficiently to bui…☆87Feb 27, 2026Updated 6 months ago
- Benchmark and training code for MindCube: spatial mental modeling in vision-language models from limited views.☆172Aug 23, 2026Updated 3 weeks ago
- [ICLR 2026] OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models☆94Jan 21, 2026Updated 7 months ago
- Cortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon Manipulation☆72Jul 16, 2026Updated 2 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 10 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 3 months ago
- ☆25Jul 30, 2026Updated last month
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 7 months ago
- ☆34May 13, 2026Updated 4 months ago
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 10 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆299Mar 6, 2025Updated last year
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆12May 5, 2025Updated last year
- ☆43May 29, 2025Updated last year
- "World Models in a Closed-Loop World" (ICLR'26 Oral)☆198Apr 3, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆156Aug 27, 2025Updated last year
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- EO: Open-source Unified Embodied Foundation Model Series☆63Jan 15, 2026Updated 8 months ago
- ☆28Aug 25, 2026Updated 3 weeks ago
- 🔥 open-ss2: a third-party open-source implementation of Figure AI's Helix "System 1, System 2" VLA model for high-rate, dexterous humano…☆11Mar 18, 2025Updated last year
- WEAVER is an efficient, consistent and high-fidelity world model enabling state-of-the-art robot policy evaluation, improvement, and fast…☆44Aug 24, 2026Updated 3 weeks ago
- Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow☆46Mar 28, 2026Updated 5 months ago
- ☆46Aug 26, 2024Updated 2 years ago
- A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation☆72Apr 1, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16Jun 19, 2026Updated 3 months ago
- [IEEE RA-L 2026] REALM: A Real-to-Sim Validated Benchmark for Generalization in Robotic Manipulation☆67Sep 6, 2026Updated 2 weeks ago
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆354Apr 18, 2026Updated 5 months ago
- World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).☆503Sep 5, 2026Updated 2 weeks ago
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 6 months ago
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation☆419Mar 11, 2026Updated 6 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 7 months ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- Implementation of paper "Playful Agentic Robot Learning"☆128Jun 20, 2026Updated 3 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code Release for Strap Paper☆26Jan 29, 2026Updated 7 months ago
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL" (EMNLP Findings 2026)☆38Nov 1, 2025Updated 10 months ago
- Implementation of "RoboAgent: Chaining Basic Capabilities for Embodied Task Planning"☆51Apr 12, 2026Updated 5 months ago
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 3 months ago
- [CVPR 2026 Highlight] XL-VLA: Cross-Hand Latent Representation for Vision-Language-Action Models☆126Jul 3, 2026Updated 2 months ago
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation☆532May 17, 2026Updated 4 months ago
- ☆27Jan 16, 2026Updated 8 months ago