[arXiv 2026] APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
☆38Jun 13, 2026Updated 2 months ago
Alternatives and similar repositories for APT
Users that are interested in APT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [arXiv 2025] Seeing to Act, Prompting to Specify: A Bayesian Factorization of Vision Language Action Policy☆22Mar 16, 2026Updated 5 months ago
- ☆26Oct 18, 2025Updated 10 months ago
- Motion Retargeting from Human Motion to Humanoid Robots Based on Gradient Descent with a User-Friendly Interface☆16Jul 14, 2025Updated last year
- ☆17Aug 3, 2025Updated last year
- Code release of [ICCV2025 Highlight] WonderPlay: Dynamic 3D Scene Generation from a Single Image and Actions☆34Mar 3, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [TASE 2025] Efficient Alignment of Unconditioned Action Prior for Language-conditioned Pick and Place in Clutter☆36Oct 27, 2025Updated 9 months ago
- This Repository will be used by the Concordia Shanghai Swarm Team Club to share & work on different programs for the DJI TT Swarm Kit☆10Dec 19, 2021Updated 4 years ago
- [TRO 2024] Grasp, See and Place: Efficient Unknown Object Rearrangement with Policy Structure Prior☆77Mar 28, 2025Updated last year
- AffordanceVLA: A Vision-Language-Action Model Empowering Action Generation through Affordance-Aware Understanding☆61Jun 5, 2026Updated 2 months ago
- simulations used in "Concept2Robot: Learning Manipulation Concepts from Instructions and Human Demonstrations"☆28Jan 1, 2023Updated 3 years ago
- ManipArena-Sim is the simulation environment for ManipArena, a real-robot benchmark for bimanual manipulation. It provides data collectio…☆21Updated this week
- ☆22Oct 4, 2021Updated 4 years ago
- Author's implementation of DemoDiffusion.☆70Jan 14, 2026Updated 7 months ago
- GUI for intuitively creating 3D reconstructions of the real world☆40Apr 9, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official repository for VCoT-Grasp.☆22Nov 18, 2025Updated 9 months ago
- ☆37May 18, 2026Updated 3 months ago
- [ICCV2023] RICO: Regularizing the Unobservable for Indoor Compositional Reconstruction☆63Sep 17, 2023Updated 2 years ago
- ☆20Mar 2, 2026Updated 5 months ago
- CoRL 2025 TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models☆115Oct 25, 2025Updated 9 months ago
- Official release of the CLVR Jaco Play Dataset, Dass et al. 2023☆17Apr 24, 2023Updated 3 years ago
- Github repository for World-VLA-Loop.☆33Feb 25, 2026Updated 5 months ago
- ActionCodec: What Makes for Good Action Tokenizers☆62Mar 1, 2026Updated 5 months ago
- ☆23Jul 1, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AR actor for specialist policy training☆27May 4, 2026Updated 3 months ago
- RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation☆39Aug 3, 2026Updated 3 weeks ago
- ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation https://procvlm.github.io/☆24May 15, 2026Updated 3 months ago
- [ICRA 2023] A Joint Modeling of Vision-Language-Action for Target-oriented Grasping in Clutter☆163Apr 13, 2025Updated last year
- ☆16Aug 12, 2023Updated 3 years ago
- ☆20Jun 8, 2026Updated 2 months ago
- [ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models☆92May 18, 2026Updated 3 months ago
- ☆33May 16, 2025Updated last year
- LaST-R1☆114May 6, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Official Release of "Mixture of Horizons in Action Chunking", ICML 2026☆62May 4, 2026Updated 3 months ago
- Cross-platform SDK for 3D TOF camera☆19May 8, 2026Updated 3 months ago
- Code for the RSS 2023 paper "Energy-based Models are Zero-Shot Planners for Compositional Scene Rearrangement"☆21Jul 4, 2023Updated 3 years ago
- "Sequential Dexterity: Chaining Dexterous Policies for Long-Horizon Manipulation" code repository☆184Apr 25, 2024Updated 2 years ago
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆77Jul 29, 2026Updated 3 weeks ago
- FieldGen is a semi-automatic data generation framework that enables scalable collection of diverse, high-quality real-world manipulation …☆27Oct 28, 2025Updated 9 months ago
- Official implementation of "AnyPlace: Learning Generalized Object Placement for Robot Manipulation"☆99Mar 25, 2025Updated last year