[CoRL2024] Official repo of `A3VLM: Actionable Articulation-Aware Vision Language Model`
☆122Oct 7, 2024Updated last year
Alternatives and similar repositories for A3VLM
Users that are interested in A3VLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [IROS24 Oral]ManipVQA: Injecting Robotic Affordance and Physically Grounded Information into Multi-Modal Large Language Models☆102Aug 22, 2024Updated last year
- [arXiv 2024] Articulated Object Manipulation using Online Axis Estimation with SAM2-Based Tracking☆18Apr 4, 2025Updated last year
- The official codebase for ManipLLM: Embodied Multimodal Large Language Model for Object-Centric Robotic Manipulation(cvpr 2024)☆150Jul 9, 2024Updated 2 years ago
- [CVPR 2024] Hierarchical Diffusion Policy for Multi-Task Robotic Manipulation☆238Apr 9, 2024Updated 2 years ago
- Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model☆374Jun 23, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for Ditto in the House: Building Articulation Models of Indoor Scenes through Interactive Perception☆17Aug 25, 2023Updated 2 years ago
- [IROS 2023] Open-Vocabulary Affordance Detection in 3d Point Clouds☆89Sep 4, 2024Updated last year
- ☆20Dec 18, 2024Updated last year
- A unified architecture for multimodal multi-task robotic policy learning.☆185Feb 2, 2024Updated 2 years ago
- Code for the paper "3D Diffuser Actor: Policy Diffusion with 3D Scene Representations"☆392Aug 17, 2024Updated last year
- Official Code for RVT-2 and RVT☆409Feb 14, 2025Updated last year
- Official implementation of RAM: Retrieval-Based Affordance Transfer for Generalizable Zero-Shot Robotic Manipulation☆101Dec 30, 2024Updated last year
- [ICML 2024] 3D-VLA: A 3D Vision-Language-Action Generative World Model☆629Oct 29, 2024Updated last year
- [ICLR'25] LLaRA: Supercharging Robot Learning Data for Vision-Language Policy☆229Mar 29, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆64Dec 14, 2024Updated last year
- [CoRL 2024] RoboEXP: Action-Conditioned Scene Graph via Interactive Exploration for Robotic Manipulation☆132Oct 26, 2025Updated 9 months ago
- ☆475Apr 14, 2026Updated 3 months ago
- F3RM: Feature Fields for Robotic Manipulation. Official repo for the paper "Distilled Feature Fields Enable Few-Shot Language-Guided Mani…☆221Apr 26, 2024Updated 2 years ago
- [CoRL 2023] REFLECT: Summarizing Robot Experiences for Failure Explanation and Correction☆107Mar 12, 2024Updated 2 years ago
- [CVPR 2023 Highlight] GAPartNet: Cross-Category Domain-Generalizable Object Perception and Manipulation via Generalizable and Actionable …☆163Oct 29, 2024Updated last year
- [ICML 2024] LEO: An Embodied Generalist Agent in 3D World☆487Apr 20, 2025Updated last year
- Template Code for the Paper: MILES: Making Imitation Learning Easy with Self-Supervision☆19Nov 14, 2024Updated last year
- [RSS 2024] Learning Manipulation by Predicting Interaction☆119Jul 2, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ECCV 2024] 🎉 Official repository of "Robo-ABC: Affordance Generalization Beyond Categories via Semantic Correspondence for Robot Manipu…☆101Nov 26, 2024Updated last year
- VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models☆826Feb 20, 2025Updated last year
- Official implementation of GR-MG☆90Jan 12, 2025Updated last year
- [NeurIPS 2024 D&B] Point Cloud Matters: Rethinking the Impact of Different Observation Spaces on Robot Learning☆92Oct 14, 2024Updated last year
- Voltron: Language-Driven Representation Learning for Robotics☆236Jul 9, 2023Updated 3 years ago
- code implementation of GraspGPT and FoundationGrasp☆151Dec 17, 2025Updated 7 months ago
- ☆47May 13, 2024Updated 2 years ago
- FlowBot3D: Learning 3D Articulation Flow to Manipulate Articulated Objects☆34Sep 18, 2023Updated 2 years ago
- Suite of human-collected datasets and a multi-task continuous control benchmark for open vocabulary visuolinguomotor learning.☆363Jul 2, 2026Updated 3 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Pytorch implementation of the models RT-1-X and RT-2-X from the paper: "Open X-Embodiment: Robotic Learning Datasets and RT-X Models"☆244Updated this week
- Official codebase for "Any-point Trajectory Modeling for Policy Learning"☆278Jun 19, 2025Updated last year
- MOKA: Open-World Robotic Manipulation through Mark-based Visual Prompting (RSS 2024)☆101Jul 16, 2024Updated 2 years ago
- Pre-training Reusable Representations for Robotic Manipulation Using Diverse Human Video Data☆378Mar 21, 2023Updated 3 years ago
- Code for the RSS 2023 paper "Energy-based Models are Zero-Shot Planners for Compositional Scene Rearrangement"☆21Jul 4, 2023Updated 3 years ago
- ☆137Apr 25, 2023Updated 3 years ago
- ☆31Jun 24, 2024Updated 2 years ago