Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation (CVPR-26)
☆25May 19, 2026Updated 2 months ago
Alternatives and similar repositories for TAVP
Users that are interested in TAVP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis☆13Dec 26, 2024Updated last year
- DDP-WM: Disentangled Dynamics Prediction for Efficient World Models (ICML-26)☆19Mar 4, 2026Updated 4 months ago
- The official repository of [CVPR2025] DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering☆28Apr 18, 2025Updated last year
- DART: Differentiable Adaptive Region Tokenizer for Vision Foundation Models☆21Oct 13, 2025Updated 9 months ago
- ☆14Oct 23, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated 10 months ago
- Transferable Feature Representation for Visible-to-Infrared Cross-Dataset Human Action Recognition (Complexity 2018)☆13Dec 14, 2022Updated 3 years ago
- 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians (ACM MM 25)☆79Jul 21, 2025Updated last year
- [IEEE T-CSVT 2019] Hierarchically Learned View-Invariant Representations for Cross-View Action Recognition☆14Nov 26, 2019Updated 6 years ago
- ☆152Aug 27, 2025Updated 10 months ago
- [IEEE T-PAMI 2023] Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering☆20Jul 6, 2023Updated 3 years ago
- FlowBotHD: History-Aware Diffuser Handling Ambiguities in Articulated Objects Manipulation☆13Dec 13, 2024Updated last year
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆24Dec 19, 2023Updated 2 years ago
- [CVPR2026] GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation☆27Jul 6, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [ICLR 2026] 🦅 FALCON: an effective vision-language-action model injects rich 3D spatial tokens into the action head, enabling robust spa…☆31May 26, 2026Updated last month
- Official implementation of Adapt3R: Adaptive 3D Scene Representation for Domain Transfer in Imitation Learning☆52Updated this week
- [IEEE T-IP 2021] Semantics-aware Adaptive Knowledge Distillation for Cross-modal Action Recognition☆29Jan 6, 2025Updated last year
- "CF3: Compact and Fast 3D Feature Fields" (ICCV 2025)☆20Oct 12, 2025Updated 9 months ago
- code for "GeoVideo: Introducing Geometric Regularization into Video Generation Models"☆18Jan 8, 2026Updated 6 months ago
- ☆18Jul 21, 2025Updated last year
- (ICCV 2025) Official repository of paper "Rethinking Bimanual Robotic Manipulation: Learning with Decoupled Interaction Framework☆15Oct 15, 2025Updated 9 months ago
- Data collection, saving and publishing code for ARIO dataset. Collect multi-sensor rostopic data and store it in a specific structure.☆19Sep 27, 2024Updated last year
- KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation☆22Apr 23, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆62Jan 15, 2024Updated 2 years ago
- This repository provides the code for training the position constrained generative grasp sampler from the paper Constrained Generative Sa…☆22Dec 4, 2024Updated last year
- [ICCV 2025] Official repository of the paper "BlinkTrack: Feature Tracking over 80 FPS via Events and Images". This repository contains t…☆20Aug 27, 2025Updated 10 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 6 months ago
- [CVPR-2025] GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding☆47Aug 15, 2025Updated 11 months ago
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆37Mar 1, 2026Updated 4 months ago
- ☆86Jun 16, 2026Updated last month
- TAX-Pose: Methods for Relative Placement☆25Jul 11, 2024Updated 2 years ago
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆51Apr 10, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Sparse-Dense Motion Modelling and Tracking for Manipulation without Prior Object Models☆13Jun 11, 2026Updated last month
- The official implementation of “Cross-Modal Causal Representation Learning for Radiology Report Generation” (IEEE T-IP 2025)☆68May 27, 2025Updated last year
- [CVPR 2026] MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent☆36Apr 30, 2026Updated 2 months ago
- A Pragmatist Robot 🤖☆19Mar 28, 2026Updated 3 months ago
- ☆37Apr 10, 2026Updated 3 months ago
- ☆16Jun 9, 2024Updated 2 years ago
- Official Implementation of ARM4R ICML 2025☆54Sep 18, 2025Updated 10 months ago