[ICLR 2026] Official implemetation of the paper "Policy Contrastive Decoding for Robotic Foundation Models"
☆29Mar 5, 2026Updated 5 months ago
Alternatives and similar repositories for PCD
Users that are interested in PCD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICRA 2026] Official implemetation of the paper "InSpire: Vision-Language-Action Models with Intrinsic Spatial Reasoning"☆51Feb 2, 2026Updated 6 months ago
- ☆17Apr 5, 2023Updated 3 years ago
- 2025 CCF BDCI DeepSearch 赛道 Top 方案☆91Apr 15, 2026Updated 3 months ago
- Recoverable Compression: A Multimodal Vision Token Recovery Mechanism Guided by Text Information☆23Apr 13, 2025Updated last year
- [CVPR 2025] Offical implementation of the paper "Skip Tuning: Pre-trained Vision-Language Models are Effective and Efficient Adapters The…☆31Mar 12, 2026Updated 4 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2024] Offical implemention of the paper "DePT: Decoupled Prompt Tuning"☆107Nov 24, 2025Updated 8 months ago
- 🔥 open-ss2: a third-party open-source implementation of Figure AI's Helix "System 1, System 2" VLA model for high-rate, dexterous humano…☆11Mar 18, 2025Updated last year
- [WIP] Code for LangToMo☆21Mar 19, 2026Updated 4 months ago
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆33Nov 2, 2025Updated 9 months ago
- ☆25Jun 1, 2021Updated 5 years ago
- Official implementation of PriorVLA.☆17May 11, 2026Updated 2 months ago
- ☆43Apr 9, 2026Updated 4 months ago
- Code for the paper Robot Data Curation with Mutual Information Estimators☆41Apr 22, 2025Updated last year
- [CVPR 2026] Cross from Left to Right Brain: Adaptive Text Dreamer for Vision-and-Language Navigation☆19May 28, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICML 2022] Channel Importance Matters in Few-shot Image Classification☆59Apr 19, 2023Updated 3 years ago
- [NeurIPS 2025 Spotlight] Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning.☆156Mar 31, 2026Updated 4 months ago
- 🦾 A Dual-System VLA with System2 Thinking☆149Aug 21, 2025Updated 11 months ago
- OpenVLA: An open-source vision-language-action model for robotic manipulation.☆376Mar 19, 2025Updated last year
- Code for Ditto in the House: Building Articulation Models of Indoor Scenes through Interactive Perception☆16Aug 25, 2023Updated 2 years ago
- Vision-Language-Action Optimization with Trajectory Ensemble Voting (ICANN2026)☆26Feb 18, 2026Updated 5 months ago
- ☆112Mar 23, 2026Updated 4 months ago
- ☆15Jun 11, 2025Updated last year
- Cross-State Transition Attention Transformer for improved robotic manipulation with better temporal modeling; https://arxiv.org/abs/2510.…☆19Mar 8, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AFUN: Towards an Affordance Foundation Model for Functionality Understanding☆38Jun 15, 2026Updated last month
- 4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.☆58Jan 10, 2026Updated 7 months ago
- [TPAMI 2023] Object Affinity Learning: Towards Annotation-free Instance Segmentation☆14Sep 14, 2023Updated 2 years ago
- Efficiently apply modification functions to RLDS/TFDS datasets.☆44Jun 5, 2024Updated 2 years ago
- [ICCV 2025] MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation☆55Oct 14, 2025Updated 9 months ago
- This repository contains the experiments conducted in the ICLR 2022 spotlight paper "On the Importance of Firth Bias Reduction in Few-Sho…☆11Apr 20, 2022Updated 4 years ago
- ☆15Mar 15, 2024Updated 2 years ago
- [RSS 2025] CLIP-RT : Learning Language-Conditioned Robotic Policies from Natural Language Supervision☆37May 13, 2025Updated last year
- VLA-GSE: Boosting Parameter Efficient Finetuning in VLA with Generalized and Specialized Experts☆21Jul 11, 2026Updated 3 weeks ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆13Jun 20, 2022Updated 4 years ago
- Coarse-to-fine Q-Network☆59Aug 6, 2024Updated 2 years ago
- ☆30Jun 30, 2026Updated last month
- Evaluating and reproducing real-world robot manipulation policies (e.g., RT-1, RT-1-X, Octo, and OpenVLA) in simulation under common setu…☆271Jun 23, 2025Updated last year
- Online Product Reviews for Affordances☆24Dec 12, 2018Updated 7 years ago
- [ICML 2025] OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction☆118Apr 14, 2025Updated last year
- Official repository of the "Ego3DPose: Capturing 3D Cues from Binocular Egocentric Views" (SIGGRAPH Asia 2023)☆10Dec 24, 2024Updated last year