ActionCodec: What Makes for Good Action Tokenizers
☆70Mar 1, 2026Updated 6 months ago
Alternatives and similar repositories for actioncodec
Users that are interested in actioncodec are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LAP: Language-Action Pre-Training Enables Zero-Shot Cross Embodiment Transfer☆170May 20, 2026Updated 3 months ago
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆56Mar 24, 2026Updated 5 months ago
- [RSS 2026 Finalist] Ordered Action Tokenization☆115Jul 27, 2026Updated last month
- ☆50May 12, 2026Updated 4 months ago
- Pytorch PI-zero and PI-zero-fast. Adapted from LeRobot☆206Sep 2, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆33Jun 7, 2026Updated 3 months ago
- [ACM MM'26 Oral] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆65May 14, 2026Updated 3 months ago
- Official implementation of the paper "Conditioning Matters: Training Diffusion Policies is Faster Than You Think".☆18May 19, 2025Updated last year
- ☆33Mar 10, 2024Updated 2 years ago
- ☆26Feb 19, 2024Updated 2 years ago
- ☆30Apr 21, 2026Updated 4 months ago
- Contact-Anchored Policies: Contact Conditioning Creates Strong Robot Utility Models☆25Apr 2, 2026Updated 5 months ago
- InternVLA-A1: Unifying Understanding, Generation, and Action for Robotic Manipulation☆552Jul 20, 2026Updated last month
- 🔥 The first open-sourced diffusion vision-langauge-action model. [ICLR 2026]☆186Mar 12, 2026Updated 6 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Reshaping Action Error Distributions for Reliable Vision-Language-Action Models☆17Feb 5, 2026Updated 7 months ago
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆133Nov 15, 2025Updated 9 months ago
- [ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models☆95May 18, 2026Updated 3 months ago
- 复旦研究生入学教育测试☆33Aug 28, 2025Updated last year
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆286Jul 7, 2026Updated 2 months ago
- GeRM: A Generalist Robotic Model with Mixture-of-Experts for Quadruped Robot https://songwxuan.github.io/GeRM/☆38Apr 29, 2025Updated last year
- ☆19Jul 7, 2024Updated 2 years ago
- [ICRA'25] Official code repository of "QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning"☆23Jun 25, 2026Updated 2 months ago
- 🦾 A Dual-System VLA with System2 Thinking☆149Aug 21, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- FieldGen is a semi-automatic data generation framework that enables scalable collection of diverse, high-quality real-world manipulation …☆27Oct 28, 2025Updated 10 months ago
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 10 months ago
- ☆18Aug 21, 2024Updated 2 years ago
- A large-scale benchmark for evaluating vision-language-action models, embodied agents, and vision-language models☆470Nov 11, 2025Updated 10 months ago
- [CVPR'2026] "MM-ACT: Learn from Multimodal Parallel Generation to Act"☆119Mar 13, 2026Updated 6 months ago
- [CoRL 2026] APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies☆41Jun 13, 2026Updated 3 months ago
- Official code for "From Seeing to Doing: Bridging Reasoning and Decision for Robotic Manipulation" (ICLR2026)☆39Mar 1, 2026Updated 6 months ago
- Official code of RDT 2☆806Feb 7, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆410Aug 21, 2026Updated 3 weeks ago
- [ICCV2025] Official code repository of "CARP: Visuomotor Policy Learning via Coarse-to-Fine Autoregressive Prediction"☆62Aug 10, 2025Updated last year
- code for CoRL2025 "LaDiWM: A Latent Diffusion-based World Model for Predictive Manipulation"☆60Nov 30, 2025Updated 9 months ago
- AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World | CoRL 2025☆106Mar 26, 2026Updated 5 months ago
- This is the official codebase for paper: Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Acti…☆65Jul 11, 2026Updated 2 months ago
- ViVa: A Video-Generative Value Model for Robot Reinforcement Learning☆94Jun 30, 2026Updated 2 months ago
- LIBERO-PRO is the official repository of the LIBERO-PRO — an evaluation extension of the original LIBERO benchmark☆315Jul 17, 2026Updated last month