ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple and have a scalable dataset.
☆54Nov 27, 2025Updated 10 months ago
Alternatives and similar repositories for ENACT
Users that are interested in ENACT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- THEORY OF SPACE: a benchmark for evaluating whether foundation models can actively explore under partial observability efficiently to bui…☆87Feb 27, 2026Updated 7 months ago
- Benchmark and training code for MindCube: spatial mental modeling in vision-language models from limited views.☆173Aug 23, 2026Updated last month
- [ICLR 2026] OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models☆95Jan 21, 2026Updated 8 months ago
- Cortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon Manipulation☆77Jul 16, 2026Updated 2 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS 2026] When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆20Sep 29, 2026Updated last week
- ☆25Jul 30, 2026Updated 2 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 8 months ago
- ☆35May 13, 2026Updated 4 months ago
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 11 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆298Mar 6, 2025Updated last year
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆12May 5, 2025Updated last year
- ☆44May 29, 2025Updated last year
- "World Models in a Closed-Loop World" (ICLR'26 Oral)☆211Apr 3, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆155Aug 27, 2025Updated last year
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- EO: Open-source Unified Embodied Foundation Model Series☆64Jan 15, 2026Updated 8 months ago
- ☆28Aug 25, 2026Updated last month
- 🔥 open-ss2: a third-party open-source implementation of Figure AI's Helix "System 1, System 2" VLA model for high-rate, dexterous humano…☆11Mar 18, 2025Updated last year
- WEAVER is an efficient, consistent and high-fidelity world model enabling state-of-the-art robot policy evaluation, improvement, and fast…☆47Aug 24, 2026Updated last month
- Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow☆49Mar 28, 2026Updated 6 months ago
- ☆46Aug 26, 2024Updated 2 years ago
- A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation☆76Apr 1, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆18Jun 19, 2026Updated 3 months ago
- [IEEE RA-L 2026] REALM: A Real-to-Sim Validated Benchmark for Generalization in Robotic Manipulation☆68Sep 6, 2026Updated last month
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆357Apr 18, 2026Updated 5 months ago
- World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).☆510Sep 5, 2026Updated last month
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 6 months ago
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation☆420Mar 11, 2026Updated 7 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 8 months ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- Implementation of paper "Playful Agentic Robot Learning"☆133Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code Release for Strap Paper☆27Jan 29, 2026Updated 8 months ago
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL" (EMNLP Findings 2026)☆38Nov 1, 2025Updated 11 months ago
- Implementation of "RoboAgent: Chaining Basic Capabilities for Embodied Task Planning"☆51Apr 12, 2026Updated 5 months ago
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 4 months ago
- [CVPR 2026 Highlight] XL-VLA: Cross-Hand Latent Representation for Vision-Language-Action Models☆128Jul 3, 2026Updated 3 months ago
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation☆540May 17, 2026Updated 4 months ago
- ☆28Jan 16, 2026Updated 8 months ago