ENACT is a benchmark that evaluates embodied cognition through world modeling from egocentric interaction. It is designed to be simple and have a scalable dataset.
☆53Nov 27, 2025Updated 9 months ago
Alternatives and similar repositories for ENACT
Users that are interested in ENACT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- THEORY OF SPACE: a benchmark for evaluating whether foundation models can actively explore under partial observability efficiently to bui…☆86Feb 27, 2026Updated 6 months ago
- Benchmark and training code for MindCube: spatial mental modeling in vision-language models from limited views.☆168Aug 23, 2026Updated last week
- [ICLR 2026] OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models☆94Jan 21, 2026Updated 7 months ago
- Cortex: A Bidirectionally Aligned Embodied Agent Framework for Long-horizon Manipulation☆67Jul 16, 2026Updated last month
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 2 months ago
- ☆22Jul 30, 2026Updated last month
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- ☆34May 13, 2026Updated 3 months ago
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 10 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆298Mar 6, 2025Updated last year
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- ☆40May 29, 2025Updated last year
- "World Models in a Closed-Loop World" (ICLR'26 Oral)☆188Apr 3, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆154Aug 27, 2025Updated last year
- Evaluate Multimodal LLMs as Embodied Agents☆60Feb 14, 2025Updated last year
- EO: Open-source Unified Embodied Foundation Model Series☆62Jan 15, 2026Updated 7 months ago
- ☆26Updated this week
- 🔥 open-ss2: a third-party open-source implementation of Figure AI's Helix "System 1, System 2" VLA model for high-rate, dexterous humano…☆11Mar 18, 2025Updated last year
- WEAVER is an efficient, consistent and high-fidelity world model enabling state-of-the-art robot policy evaluation, improvement, and fast…☆43Updated this week
- In‑Context World‑Action Modeling from Human Videos for Open‑Ended Task Generalization☆41Updated this week
- Dream2Flow: Bridging Video Generation and Open-World Manipulation with 3D Object Flow☆46Mar 28, 2026Updated 5 months ago
- ☆45Aug 26, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Vision-Language-Model for Detecting and Reasoning Over Failures in Robotic Manipulation☆72Apr 1, 2025Updated last year
- ☆16Jun 19, 2026Updated 2 months ago
- [IEEE RA-L 2026] REALM: A Real-to-Sim Validated Benchmark for Generalization in Robotic Manipulation☆67Updated this week
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆352Apr 18, 2026Updated 4 months ago
- World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).☆493Updated this week
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 5 months ago
- PointWorld: Scaling 3D World Models for In-The-Wild Robotic Manipulation☆419Mar 11, 2026Updated 5 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 6 months ago
- [NeurIPS D&B Track 2024] Source code for the paper "Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge…☆25May 2, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code Release for Strap Paper☆25Jan 29, 2026Updated 7 months ago
- Implementation of paper "Playful Agentic Robot Learning"☆124Jun 20, 2026Updated 2 months ago
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL"☆37Nov 1, 2025Updated 9 months ago
- Implementation of "RoboAgent: Chaining Basic Capabilities for Embodied Task Planning"☆48Apr 12, 2026Updated 4 months ago
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 3 months ago
- This is the official evaluation code for Robobench☆24Aug 16, 2026Updated 2 weeks ago
- [CVPR 2026 Highlight] XL-VLA: Cross-Hand Latent Representation for Vision-Language-Action Models☆119Jul 3, 2026Updated last month