Official implementation of "Reward Prediction with Factorized World States"
☆20Mar 11, 2026Updated 4 months ago
Alternatives and similar repositories for StateFactory
Users that are interested in StateFactory are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆18Jun 2, 2026Updated 2 months ago
- [ACL 2025 Findings] Text2World: Benchmarking Large Language Models for Symbolic World Model Generation☆29Feb 25, 2025Updated last year
- Official implementation of the paper "TOKENTRIM: INFERENCE-TIME TOKEN PRUNING FOR AUTOREGRESSIVE LONG VIDEO GENERATION"☆15Feb 8, 2026Updated 5 months ago
- [ACM MM 25] Official repo of "UEMM-Air: Enable UAVs to Undertake More Multi-modal Tasks"☆37Aug 20, 2025Updated 11 months ago
- Official repo for "Generative Point Tracking with Flow Matching".☆24Oct 22, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Benchmark for High-level World Modeling and Long-horizon Procedural Planning☆31Dec 17, 2025Updated 7 months ago
- Collections of papers and code for employing MLLM for quality assessment tasks.☆12Apr 18, 2024Updated 2 years ago
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆21Apr 14, 2026Updated 3 months ago
- Can We Predict Before Executing Machine Learning Agents?☆21Jul 7, 2026Updated 3 weeks ago
- This is the project for 'USG'.☆41Jun 21, 2026Updated last month
- Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision☆47Oct 19, 2025Updated 9 months ago
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- Official implementation of the paper: Sub-JEPA: Subspace Gaussian Regularization for Stable End-to-End World Models.☆61Jul 27, 2026Updated last week
- [arXiv 26] RemoteAgent: Bridging Vague Human Intents and Earth Observation with RL-based Agentic MLLMs☆19Jul 5, 2026Updated last month
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Temporally Correlated Episodic Reinforcement Learning, ICLR 24☆12Apr 8, 2024Updated 2 years ago
- [ICML 2026] Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding☆16Mar 13, 2026Updated 4 months ago
- ☆11Dec 11, 2023Updated 2 years ago
- ☆20Jul 22, 2025Updated last year
- A pytorch implementation of "Robust Facial Landmark Detection by Multi-order Multi-constrained Network"☆13Dec 9, 2020Updated 5 years ago
- This repository contains the video files (download links) and corresponding annotations used in the paper "Long-Term Face Tracking for Cr…☆14Dec 18, 2020Updated 5 years ago
- ☆20Apr 2, 2026Updated 4 months ago
- ☆15Feb 25, 2026Updated 5 months ago
- Official repo for "TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders"☆25Apr 9, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICCV 2025] Factorized Learning for Temporally Grounded Video-Language Models☆24Apr 18, 2026Updated 3 months ago
- [EMNLP 2025 Outstanding Paper Award] Official repo for DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph …☆22Nov 16, 2025Updated 8 months ago
- A Benchmark for Efficient and Compositional Visual Reasoning☆25Aug 2, 2023Updated 3 years ago
- Example use cases for the GPT-4 Vision API☆19Nov 26, 2023Updated 2 years ago
- ☆54Feb 2, 2025Updated last year
- Back to Parsimonious Latents: Learning Task-Centric World Models from Visual Foundations☆23Jul 12, 2026Updated 3 weeks ago
- Official implementation of "Goal Force: Teaching Video Models To Accomplish Physics-Conditioned Goals" (CVPR 2026)☆42Feb 25, 2026Updated 5 months ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆68Apr 13, 2026Updated 3 months ago
- [CVPR'2022, TPAMI'2024] LAVT: Language-Aware Vision Transformer for Referring Segmentation☆26Jan 21, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ACM MM 25] Official repo of "RemoteSAM: Towards Segment Anything for Earth Observation"☆246Jan 4, 2026Updated 7 months ago
- [AAAI 26] Official repo of "RemoteReasoner: Towards Unifying Geospatial Reasoning Workflow"☆16Nov 24, 2025Updated 8 months ago
- ☆35Mar 18, 2026Updated 4 months ago
- In-Context Reinforcement Learning for Tool Use in Large Language Models☆48Mar 26, 2026Updated 4 months ago
- ☆11Sep 8, 2016Updated 9 years ago
- Aligning Agentic World Models via Knowledgeable Experience Learning☆38May 15, 2026Updated 2 months ago
- Learn how to use OpenML for reproducible, collaborative machine learning projects☆14Aug 6, 2023Updated 2 years ago