Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation (CVPR-26)
☆25May 19, 2026Updated 2 months ago
Alternatives and similar repositories for TAVP
Users that are interested in TAVP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis☆13Dec 26, 2024Updated last year
- DDP-WM: Disentangled Dynamics Prediction for Efficient World Models (ICML-26)☆18Mar 4, 2026Updated 5 months ago
- The official repository of [CVPR2025] DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering☆28Apr 18, 2025Updated last year
- DART: Differentiable Adaptive Region Tokenizer for Vision Foundation Models☆22Oct 13, 2025Updated 9 months ago
- ☆14Oct 23, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Transferable Feature Representation for Visible-to-Infrared Cross-Dataset Human Action Recognition (Complexity 2018)☆13Dec 14, 2022Updated 3 years ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated 11 months ago
- 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians (ACM MM 25)☆79Jul 21, 2025Updated last year
- [IEEE T-CSVT 2019] Hierarchically Learned View-Invariant Representations for Cross-View Action Recognition☆14Nov 26, 2019Updated 6 years ago
- [IEEE T-PAMI 2023] Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering☆20Jul 6, 2023Updated 3 years ago
- ☆152Aug 27, 2025Updated 11 months ago
- FlowBotHD: History-Aware Diffuser Handling Ambiguities in Articulated Objects Manipulation☆13Dec 13, 2024Updated last year
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆24Dec 19, 2023Updated 2 years ago
- [ICRA 2024] SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs☆21Jun 1, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [CVPR2026] GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation☆27Jul 6, 2026Updated last month
- [ICLR 2026] 🦅 FALCON: an effective vision-language-action model injects rich 3D spatial tokens into the action head, enabling robust spa…☆35May 26, 2026Updated 2 months ago
- Official implementation of Adapt3R: Adaptive 3D Scene Representation for Domain Transfer in Imitation Learning☆52Jul 17, 2026Updated 3 weeks ago
- [IEEE T-IP 2021] Semantics-aware Adaptive Knowledge Distillation for Cross-modal Action Recognition☆29Jan 6, 2025Updated last year
- "CF3: Compact and Fast 3D Feature Fields" (ICCV 2025)☆21Oct 12, 2025Updated 9 months ago
- code for "GeoVideo: Introducing Geometric Regularization into Video Generation Models"☆19Jan 8, 2026Updated 7 months ago
- ☆19Jul 21, 2025Updated last year
- (ICCV 2025) Official repository of paper "Rethinking Bimanual Robotic Manipulation: Learning with Decoupled Interaction Framework☆17Oct 15, 2025Updated 9 months ago
- ☆62Jan 15, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository provides the code for training the position constrained generative grasp sampler from the paper Constrained Generative Sa…☆22Dec 4, 2024Updated last year
- The official implementation of "Cross-modal Causal Relation Alignment for Video Question Grounding. (CVPR 2025 Highlight)"☆52Apr 27, 2025Updated last year
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 7 months ago
- [CVPR-2025] GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding☆47Aug 15, 2025Updated 11 months ago
- ☆86Jun 16, 2026Updated last month
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆38Mar 1, 2026Updated 5 months ago
- TAX-Pose: Methods for Relative Placement☆25Jul 11, 2024Updated 2 years ago
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆51Apr 10, 2026Updated 4 months ago
- Data collection, saving and publishing code for ARIO dataset. Collect multi-sensor rostopic data and store it in a specific structure.☆19Sep 27, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Sparse-Dense Motion Modelling and Tracking for Manipulation without Prior Object Models☆13Jul 28, 2026Updated last week
- The official implementation of “Cross-Modal Causal Representation Learning for Radiology Report Generation” (IEEE T-IP 2025)☆68May 27, 2025Updated last year
- A Pragmatist Robot 🤖☆20Mar 28, 2026Updated 4 months ago
- [CVPR 2026] MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent☆37Apr 30, 2026Updated 3 months ago
- ☆37Apr 10, 2026Updated 4 months ago
- Official Implementation of ARM4R ICML 2025☆54Sep 18, 2025Updated 10 months ago
- ☆16Jun 9, 2024Updated 2 years ago