[CVPR 2025] Official Implementation for Optimus-2: Multimodal Minecraft Agent with Goal-Observation-Action Conditioned Policy
☆27Jun 17, 2025Updated last year
Alternatives and similar repositories for CVPR25-Optimus-2
Users that are interested in CVPR25-Optimus-2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code and benchmark of the paper "MineAnyBuild: Benchmarking Spatial Planning for Open-world AI Agents" (NeurIPS D&B 2025)☆15Oct 13, 2025Updated 9 months ago
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆25Updated this week
- Official repository of the "Fine-grained Key-Value Memory Enhanced Predictor for Video Representation Learning" (ACM MM 2023)☆23Jul 11, 2024Updated 2 years ago
- [CVPR 2022 Oral] Faithful Extreme Rescaling via Generative Prior Reciprocated Invertible Representations☆13Jul 14, 2022Updated 4 years ago
- Mobile network bandwidth traces☆14Oct 13, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code of the paper "Correctable Landmark Discovery via Large Models for Vision-Language Navigation" (TPAMI 2024)☆16Jun 7, 2024Updated 2 years ago
- Detection and Reconstruction of Transparent Objects with Infrared Projection-based RGB-D Cameras☆13Jan 17, 2021Updated 5 years ago
- [CVPR 2025] LION-FS: Fast & Slow Video-Language Thinker as Online Video Assistant☆29Dec 2, 2025Updated 7 months ago
- XS-VID: An Extra Small Object Video Detection Dataset☆10Mar 4, 2025Updated last year
- https://www.datafountain.cn/competitions/518☆13Mar 1, 2023Updated 3 years ago
- Contrastive multi-omics association learning☆13Apr 28, 2026Updated 2 months ago
- Python Implementation of paper "Robust Camera Calibration for Sport Videos using Court Models"☆14Nov 15, 2023Updated 2 years ago
- A customized docker for headless GPU rendering without host-side configuration☆11Aug 22, 2022Updated 3 years ago
- https://www.kaggle.com/competitions/sorghum-id-fgvc-9☆19Mar 1, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Source code for the paper: "Pantheon: Preemptible Multi-DNN Inference on Mobile Edge GPUs"☆16Apr 15, 2024Updated 2 years ago
- ☆12Apr 22, 2025Updated last year
- CVPR18: Learning and Using the Arrow of Time☆40Feb 11, 2022Updated 4 years ago
- HEtero-Assists Distillation for Heterogeneous Object Detectors☆10Jul 3, 2023Updated 3 years ago
- MoTIF: Learning Motion Trajectories with Local Implicit Neural Functions for Continuous Space-Time Video Super-Resolution☆38Sep 30, 2023Updated 2 years ago
- Implemention of "Realtime Multi Person Pose-Estimation" in pytorch with data from AI Challenger☆13Nov 24, 2017Updated 8 years ago
- Code for the C2KD paper (ICASSP 2023)☆20May 15, 2023Updated 3 years ago
- [ICASSP'25] Enhancing Vision-Language Tracking by Effectively Converting Textual Cues into Visual Cues☆19Dec 31, 2024Updated last year
- 这是一个问卷星互填社区刷点数的工具,进而帮您更快采集自己的样本☆11Oct 31, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation of "Open-World Multi-Task Control Through Goal-Aware Representation Learning and Adaptive Horizon Prediction"☆47Aug 15, 2023Updated 2 years ago
- [AAAI 2026 Oral] SemanticVLA: Semantic-Aligned Sparsification and Enhancement for Efficient Robotic Manipulation☆70Apr 5, 2026Updated 3 months ago
- ☆29Jun 30, 2026Updated 2 weeks ago
- A curated list of frameworks, tools, and resources for building and deploying AI agents. From multi-agent systems to autonomous coding as…☆35Jul 13, 2026Updated last week
- 基于ViT模型的医疗图像辅助诊断系统☆11Jan 30, 2024Updated 2 years ago
- ☆22Apr 17, 2026Updated 3 months ago
- This is the official impletations of the EMNLP Findings paper, VideoINSTA: Zero-shot Long-Form Video Understanding via Informative Spatia…☆24Apr 7, 2026Updated 3 months ago
- LagMemo: Language 3D Gaussian Splatting Memory for Multi-modal Open-vocabulary Multi-goal Visual Navigation☆18Jun 17, 2026Updated last month
- Coda and Data for NeurIPS 2025 paper "MuSLR: Multimodal Symbolic Logical Reasoning"☆16Oct 5, 2025Updated 9 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆15Aug 5, 2025Updated 11 months ago
- [IEEE TMM 2025] CRSOT: Cross-Resolution Object Tracking using Unaligned Frame and Event Cameras☆22Jan 18, 2025Updated last year
- TransMDOT☆22Jan 8, 2024Updated 2 years ago
- Source code for book "Image algorithms for low-level vision tasks" (Jia. 2024), including denoising, super-resolution, dehazing, image co…☆20Jul 19, 2025Updated last year
- IVC-Prune: Revealing the Implicit Visual Coordinates in LVLMs for Vision Token Pruning☆16Feb 27, 2026Updated 4 months ago
- Code for Static and Dynamic Concepts for Self-supervised Video Representation Learning.☆11Jul 28, 2022Updated 3 years ago
- Learning Laplacian Representations in Reinforcement Learning☆18Jan 2, 2021Updated 5 years ago