Learning to See and Act: Task-Aware Virtual View Exploration for Robotic Manipulation (CVPR-26)
☆29May 19, 2026Updated 4 months ago
Alternatives and similar repositories for TAVP
Users that are interested in TAVP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DDP-WM: Disentangled Dynamics Prediction for Efficient World Models (ICML-26)☆19Mar 4, 2026Updated 6 months ago
- VisionGRU: A Linear-Complexity RNN Model for Efficient Image Analysis☆14Dec 26, 2024Updated last year
- The official repository of [CVPR2025] DSPNet: Dual-vision Scene Perception for Robust 3D Question Answering☆30Apr 18, 2025Updated last year
- DART: Differentiable Adaptive Region Tokenizer for Vision Foundation Models☆25Oct 13, 2025Updated 11 months ago
- ☆15Oct 23, 2023Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Transferable Feature Representation for Visible-to-Infrared Cross-Dataset Human Action Recognition (Complexity 2018)☆13Dec 14, 2022Updated 3 years ago
- [ICCV 2025] FVP: 4D Visual Pre-training for Robot Learning☆17Sep 5, 2025Updated last year
- 3DAffordSplat: Efficient Affordance Reasoning with 3D Gaussians (ACM MM 25)☆80Jul 21, 2025Updated last year
- [IEEE T-CSVT 2019] Hierarchically Learned View-Invariant Representations for Cross-View Action Recognition☆15Nov 26, 2019Updated 6 years ago
- [IEEE T-PAMI 2023] Cross-Modal Causal Relational Reasoning for Event-Level Visual Question Answering☆21Jul 6, 2023Updated 3 years ago
- ☆156Aug 27, 2025Updated last year
- FlowBotHD: History-Aware Diffuser Handling Ambiguities in Articulated Objects Manipulation☆13Dec 13, 2024Updated last year
- [IEEE T-IP 2022] TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning☆25Dec 19, 2023Updated 2 years ago
- [ICRA 2024] SG-Bot: Object Rearrangement via Coarse-to-Fine Robotic Imagination on Scene Graphs☆21Jun 1, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] 🦅 FALCON: an effective vision-language-action model injects rich 3D spatial tokens into the action head, enabling robust spa…☆37May 26, 2026Updated 3 months ago
- [CVPR2026] GeoPredict: Leveraging Predictive Kinematics and 3D Gaussian Geometry for Precise VLA Manipulation☆30Jul 6, 2026Updated 2 months ago
- Official implementation of Adapt3R: Adaptive 3D Scene Representation for Domain Transfer in Imitation Learning☆52Jul 17, 2026Updated 2 months ago
- [IEEE T-IP 2021] Semantics-aware Adaptive Knowledge Distillation for Cross-modal Action Recognition☆30Jan 6, 2025Updated last year
- "CF3: Compact and Fast 3D Feature Fields" (ICCV 2025)☆21Oct 12, 2025Updated 11 months ago
- ☆89Jun 16, 2026Updated 3 months ago
- code for "GeoVideo: Introducing Geometric Regularization into Video Generation Models"☆21Jan 8, 2026Updated 8 months ago
- A comprehensive list of Implicit Representations and NeRF papers relating to SLAM/Robotics domain, including papers, video, codes, and re…☆13Nov 27, 2023Updated 2 years ago
- ☆19Jul 21, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- (ICCV 2025) Official repository of paper "Rethinking Bimanual Robotic Manipulation: Learning with Decoupled Interaction Framework☆18Oct 15, 2025Updated 11 months ago
- ☆62Jan 15, 2024Updated 2 years ago
- This repository provides the code for training the position constrained generative grasp sampler from the paper Constrained Generative Sa…☆22Dec 4, 2024Updated last year
- The official implementation of "Cross-modal Causal Relation Alignment for Video Question Grounding. (CVPR 2025 Highlight)"☆53Apr 27, 2025Updated last year
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 8 months ago
- [CVPR-2025] GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding☆50Aug 15, 2025Updated last year
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆39Mar 1, 2026Updated 6 months ago
- Embodied Question Answering (EQA) benchmark and method (ICCV 2025)☆61Aug 12, 2025Updated last year
- TAX-Pose: Methods for Relative Placement☆25Jul 11, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- personal practice☆17Apr 2, 2021Updated 5 years ago
- [CVPR 2025] VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation☆53Apr 10, 2026Updated 5 months ago
- Data collection, saving and publishing code for ARIO dataset. Collect multi-sensor rostopic data and store it in a specific structure.☆20Sep 27, 2024Updated last year
- Sparse-Dense Motion Modelling and Tracking for Manipulation without Prior Object Models☆14Jul 28, 2026Updated last month
- The official implementation of “Cross-Modal Causal Representation Learning for Radiology Report Generation” (IEEE T-IP 2025)☆70May 27, 2025Updated last year
- A Pragmatist Robot 🤖☆23Mar 28, 2026Updated 5 months ago
- [CVPR 2026] MergeVLA: Cross-Skill Model Merging Toward a Generalist Vision-Language-Action Agent☆37Aug 19, 2026Updated last month