[ICLR 2026] Official implemetation of the paper "Policy Contrastive Decoding for Robotic Foundation Models"
☆29Mar 5, 2026Updated 5 months ago
Alternatives and similar repositories for PCD
Users that are interested in PCD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2026] Official implemetation of the paper "InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning"☆51Feb 2, 2026Updated 6 months ago
- 2025 CCF BDCI DeepSearch 赛道 Top 方案☆92Apr 15, 2026Updated 4 months ago
- Recoverable Compression: A Multimodal Vision Token Recovery Mechanism Guided by Text Information☆23Apr 13, 2025Updated last year
- [CVPR 2025] Offical implementation of the paper "Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters The…☆31Mar 12, 2026Updated 5 months ago
- 🔥 open-ss2: a third-party open-source implementation of Figure AI's Helix "System 1, System 2" VLA model for high-rate, dexterous humano…☆11Mar 18, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [WIP] Code for LangToMo☆21Mar 19, 2026Updated 5 months ago
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆34Nov 2, 2025Updated 9 months ago
- ☆25Jun 1, 2021Updated 5 years ago
- MemoryWAM: Efficient World Action Modeling with Persistent Memory☆76Jun 19, 2026Updated 2 months ago
- Official implementation of PriorVLA.☆17May 11, 2026Updated 3 months ago
- ☆44Apr 9, 2026Updated 4 months ago
- Code for the paper Robot Data Curation with Mutual Information Estimators☆42Apr 22, 2025Updated last year
- [CVPR 2026] Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation☆20May 28, 2025Updated last year
- [ICML 2022] Channel Importance Matters in Few-shot Image Classification☆59Apr 19, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [NeurIPS 2025 Spotlight] Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.☆157Mar 31, 2026Updated 4 months ago
- 🦾 A Dual-System VLA with System2 Thinking☆149Aug 21, 2025Updated last year
- OpenVLA: An open-source vision-language-action model for robotic manipulation.☆376Mar 19, 2025Updated last year
- Code for Ditto in the House: Building Articulation Models of Indoor Scenes through Interactive Perception☆16Aug 25, 2023Updated 3 years ago
- Vision-Language-Action Optimization with Trajectory Ensemble Voting (ICANN2026)☆27Feb 18, 2026Updated 6 months ago
- ☆114Mar 23, 2026Updated 5 months ago
- ☆16Jun 11, 2025Updated last year
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 5 months ago
- AFUN: Towards an Affordance Foundation Model for Functionality Understanding☆44Jun 15, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.☆58Jan 10, 2026Updated 7 months ago
- Geometric Problem Solving Integrating FormalGeo Symbolic System and Hypergraph Neural Network.☆16Sep 23, 2025Updated 11 months ago
- [TPAMI 2023] Object Affinity Learning: Towards Annotation-free Instance Segmentation☆14Sep 14, 2023Updated 2 years ago
- Efficiently apply modification functions to RLDS/TFDS datasets.☆44Jun 5, 2024Updated 2 years ago
- [ICCV 2025] MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation☆56Oct 14, 2025Updated 10 months ago
- [RSS 2025] CLIP-RT : Learning Language-Conditioned Robotic Policies from Natural Language Supervision☆37May 13, 2025Updated last year
- VLA-GSE: Boosting Parameter Efficient Finetuning in VLA with Generalized and Specialized Experts☆21Jul 11, 2026Updated last month
- ☆13Jun 20, 2022Updated 4 years ago
- Coarse-to-fine Q-Network☆59Aug 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆31Jun 30, 2026Updated 2 months ago
- Evaluating and reproducing real-world robot manipulation policies (e.g., RT-1, RT-1-X, Octo, and OpenVLA) in simulation under common setu…☆272Jun 23, 2025Updated last year
- Online Product Reviews for Affordances☆24Dec 12, 2018Updated 7 years ago
- ☆18Jun 30, 2019Updated 7 years ago
- [ICML 2023] A Closer Look at Few-shot Classification Again☆60Jun 5, 2023Updated 3 years ago
- [ICML 2025] OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction☆119Apr 14, 2025Updated last year
- Official repository of the "Ego3DPose: Capturing 3D Cues from Binocular Egocentric Views" (SIGGRAPH Asia 2023)☆10Dec 24, 2024Updated last year