Inference, evaluation and analysis code for STEVO-Bench
☆23Jun 21, 2026Updated last month
Alternatives and similar repositories for STEVO-Bench
Users that are interested in STEVO-Bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The first open-domain closed-loop revisited benchmark for evaluating memory consistency and action control in world models.☆76Jul 2, 2026Updated last month
- Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models☆269Jul 23, 2026Updated 2 weeks ago
- ☆21Jul 31, 2026Updated last week
- Code implementation for "Feedforward 3D Editing via Text-Steerable Image-to-3D"☆57Dec 23, 2025Updated 7 months ago
- ☆17Apr 7, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A diagnostic tool and a guideline for advancing next-generation world models capable of robust understanding, forecasting, and purposeful…☆25Apr 12, 2026Updated 3 months ago
- FreeCond: A Free Lunch for Input Conditions in Text-Guided Inpainting. FreeCond introduces a more generalized form💪 of the original inpa…☆15May 22, 2025Updated last year
- [SIGGRAPH 2026] SynthVerse: A Large-Scale Diverse Synthetic Dataset for Point Tracking☆52Jul 17, 2026Updated 3 weeks ago
- [CVPR 2025] Program synthesis for 3D spatial reasoning☆63Jun 16, 2025Updated last year
- Code for "Linear Mechanisms for Spatiotemporal Reasoning in Vision Language Models"☆17Feb 16, 2026Updated 5 months ago
- [SIGGRAPH 2026] OmniRoam: World Wandering via Long-Horizon Panoramic Video Generation☆119Jul 26, 2026Updated 2 weeks ago
- [EMNLP 2025 Findings] 3D-Aware Vision-Language Models Fine-Tuning with Geometric Distillation☆39Jun 12, 2025Updated last year
- [ECCV 2026] Official implementation of paper LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models☆133Updated this week
- [NeurIPS 2025] Video World Models with Long-term Spatial Memory☆73May 11, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆37Jun 13, 2026Updated last month
- A backup of my homework. 统计学习☆11Jan 13, 2022Updated 4 years ago
- [CV4AEC Workshop CVPR 2024] Dataset + Evaluation Metrics☆28Apr 9, 2024Updated 2 years ago
- Repository for WACV23 paper "Automatically Annotating Indoor Images with CAD Models via RGB-D Scans"☆18Aug 5, 2025Updated last year
- ☆183Jun 8, 2026Updated 2 months ago
- VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images☆52Apr 28, 2026Updated 3 months ago
- Repo for visualization of MCSS outputs and its evaluation☆17Apr 26, 2021Updated 5 years ago
- UE5-based Data Engine used in OmniX☆103Jul 12, 2026Updated 3 weeks ago
- [ECCV 2026] Official Implementation of Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction☆18Apr 26, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definitio…☆34Jul 9, 2026Updated last month
- [SIGGRAPH Asia 2024 Conference] PC-Planner: Physics-Constrained Self-Supervised Learning for Robust Neural Motion Planning with Shape-Awa…☆18Mar 19, 2026Updated 4 months ago
- [Preprint] Any 3D Scene is Worth 1K Tokens: 3D-Grounded Representation for Scene Generation at Scale☆58Apr 14, 2026Updated 3 months ago
- Super Mario Bros. (NES) gameplay dataset for machine learning.☆13Jul 22, 2025Updated last year
- [ICLR 2026] Official implementation of "Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillat…☆29May 4, 2026Updated 3 months ago
- [CGF] A curated list of papers on feed-forward 3D reconstruction and novel view synthesis.☆17Mar 14, 2026Updated 4 months ago
- ☆47Feb 25, 2025Updated last year
- A Framework for Open-Vocabulary Object Retrieval and Drawer Manipulation in Point Clouds☆31Jan 19, 2025Updated last year
- ☆20Jul 14, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR'26] This repository is the implementation of "3D Aware Region Prompted Vision Language Model"☆30Feb 19, 2026Updated 5 months ago
- Create Meshes from Depth Maps in Unity☆10Feb 13, 2023Updated 3 years ago
- ☆13Sep 12, 2024Updated last year
- [Official, NeurIPS 2025] TempSamp-R1: Effective Temporal Sampling with Reinforcement Fine-Tuning for Video LLMs.☆19Jun 8, 2026Updated 2 months ago
- (ECCV 2024) MaRINeR: Enhancing Novel Views by Matching Rendered Images with Nearby References☆19Sep 11, 2024Updated last year
- ☆966Jul 24, 2026Updated 2 weeks ago
- This repository contains the code for the paper - "Aligning Text, Images, and 3D Structure Token-by-Token" (CVPR 2026)☆49Jun 11, 2025Updated last year