[CVPR 2025] Official Implementation for Optimus-2: Multimodal Minecraft Agent with Goal-Observation-Action Conditioned Policy
☆28Jun 17, 2025Updated last year
Alternatives and similar repositories for CVPR25-Optimus-2
Users that are interested in CVPR25-Optimus-2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Implementation for Optimus-3: Dual-Router Aligned Mixture-of-Experts Agent with Dual-Granularity Reasoning-Aware Policy Optimiza…☆74Apr 14, 2026Updated 5 months ago
- [NeurIPS 2024] Official Implementation for Optimus-1: Hybrid Multimodal Memory Empowered Agents Excel in Long-Horizon Tasks☆104Jun 17, 2025Updated last year
- Paper List of Minecraft Agents☆71May 24, 2026Updated 4 months ago
- Official Implementation of Paper "ROCKET-2: Steering Visuomotor Policy via Cross-View Goal Alignment" (AAAI'26)☆46Jul 2, 2025Updated last year
- Official repository of the "Fine-grained Key-Value Memory Enhanced Predictor for Video Representation Learning" (ACM MM 2023)☆23Jul 11, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆14May 13, 2025Updated last year
- Official implementation of paper "ROCKET-1: Mastering Open-World Interaction with Visual-Temporal Context Prompting" (CVPR'25)☆48Apr 13, 2025Updated last year
- ☆13Apr 28, 2025Updated last year
- Compute surface normal from depth image using d2nt and cross product method☆10Aug 3, 2023Updated 3 years ago
- GROOT: Learning to Follow Instructions by Watching Gameplay Videos (ICLR'24, Spotlight)☆71Dec 18, 2023Updated 2 years ago
- Detection and Reconstruction of Transparent Objects with Infrared Projection-based RGB-D Cameras☆13Jan 17, 2021Updated 5 years ago
- Code of the paper "Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation"…☆19Nov 11, 2025Updated 10 months ago
- Locally run an Instruction-Tuned Chat-Style LLM☆29Apr 3, 2023Updated 3 years ago
- [ICCV 23]This is a Pytorch implementation of our paper "SMMix: Self-Motivated Image Mixing for Vision Transformers"☆16Jul 14, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [IROS'25 Oral & NeurIPSw'24] Official implementation of "MineDreamer: Learning to Follow Instructions via Chain-of-Imagination for Simula…☆104Jun 16, 2025Updated last year
- ☆14Dec 16, 2024Updated last year
- XS-VID: An Extra Small Object Video Detection Dataset☆10Updated this week
- We introduce ADAM, An emboDied causal Agent in Minecraft, that can autonomously navigate the open world, perceive multimodal contexts, le…☆35Apr 7, 2025Updated last year
- ☆10Aug 18, 2026Updated last month
- PyTorch implementation of SegBlocks: Towards Block-Based Adaptive Resolution Networks for Fast Segmentation (ECCV2020 Embedded Vision Wor…☆19Mar 31, 2023Updated 3 years ago
- ☆23Apr 24, 2026Updated 5 months ago
- Contrastive multi-omics association learning☆14Apr 28, 2026Updated 5 months ago
- detecting tennis court keypoints with yolo☆10Apr 19, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This repo is a live list of papers on game playing and large multimodality model - "A Survey on Game Playing Agents and Large Models: Met…☆162Sep 3, 2024Updated 2 years ago
- ☆18Dec 1, 2025Updated 10 months ago
- ☆18Feb 23, 2024Updated 2 years ago
- EMMOE: A Comprehensive Benchmark for Embodied Mobile Manipulation in Open Environments☆28May 15, 2025Updated last year
- Pytorch implementation of Centered Kernel Alignment(CKA) and its minibatch version.☆11May 11, 2022Updated 4 years ago
- HEtero-Assists Distillation for Heterogeneous Object Detectors☆10Jul 3, 2023Updated 3 years ago
- ☆22Mar 7, 2025Updated last year
- MoTIF: Learning Motion Trajectories with Local Implicit Neural Functions for Continuous Space-Time Video Super-Resolution