Scalable annotation pipeline for action-aglined fine-grained instruciton for Visual-language-Action model
☆75Aug 3, 2026Updated this week
Alternatives and similar repositories for FineVLA
Users that are interested in FineVLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OSWorld 2.0: Benchmarking Computer Use Agents on Long-Horizon Real-World Tasks☆225Updated this week
- Code for InSight: Self-Guided Skill Acquisition via Steerable VLAs -- insight-vla.github.io/☆37Jun 24, 2026Updated last month
- Temporal Self-imitation Learning☆16Jul 3, 2026Updated last month
- The official repo of VideoAgentTrek☆58Oct 24, 2025Updated 9 months ago
- ☆59Apr 13, 2026Updated 3 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 基于开源项目 tinyrenderer 开发的一款软光栅渲染器。在保留原项目特色(0图形API)的基础上,采用了更加modern的面向对象设计模式,并借鉴了OpenGL的上下文绑定、资源池化等设计理念。有计划在未来补上AO以及PBR渲染的内容,原开源项目链接:https:/…☆19Jan 21, 2026Updated 6 months ago
- ☆23Jun 29, 2026Updated last month
- Self-hosted GPT-4V api☆27Nov 6, 2023Updated 2 years ago
- RynnWorld-Teleop: An Action-Conditioned World Model for Digital Teleoperation☆37Updated this week
- [ICLR 2026] Computer Agent Arena: Toward Human-Centric Evaluation and Analysis of Computer-Use Agents☆67Feb 26, 2026Updated 5 months ago
- Official Implementation of "Steering Vision-Language-Action Models as Anti-Exploration: A Test-Time Scaling Approach"☆39Apr 6, 2026Updated 3 months ago
- ☆51Mar 31, 2026Updated 4 months ago
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆28Jun 1, 2026Updated 2 months ago
- [NeurIPS 2025 Spotlight] Scaling Computer-Use Grounding via UI Decomposition and Synthesis☆174Jun 18, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- a lightweight detect-and-correction inference for vla☆62Jul 6, 2026Updated 3 weeks ago
- A verified version of the WebArena Benchmark☆48Mar 8, 2026Updated 4 months ago
- SimX-OR: Extending Any Simulation Benchmark to Evaluate the Observational Robustness of VLA Models☆34Nov 4, 2025Updated 8 months ago
- Official repo for CP-Gen: Constraint-Preserving Data Generation for Visuomotor Policy Generalization☆23Oct 27, 2025Updated 9 months ago
- ☆23Aug 22, 2025Updated 11 months ago
- Implementation of VLM4VLA☆165Apr 22, 2026Updated 3 months ago
- VLAC: A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning☆324Jul 13, 2026Updated 3 weeks ago
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy☆418Feb 11, 2026Updated 5 months ago
- ☆15Jul 26, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆27Aug 20, 2025Updated 11 months ago
- ☆16Mar 8, 2026Updated 4 months ago
- ☆66Jul 15, 2025Updated last year
- View-Invariant Policy Learning via Zero-Shot Novel View Synthesis (CoRL 2024)☆31Sep 28, 2025Updated 10 months ago
- ☆42Feb 16, 2026Updated 5 months ago
- [CVPR 2025] Official implementation of "GenManip: LLM-driven Simulation for Generalizable Instruction-Following Manipulation"☆170Jul 6, 2026Updated 3 weeks ago
- Official Release of Multistep Quasimetric Estimation (MQE)☆19Mar 13, 2026Updated 4 months ago
- ☆17Jan 19, 2026Updated 6 months ago
- MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence☆61Mar 11, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICLR2025 Spotlight] Agent Trajectory Synthesis via Guiding Replay with Web Tutorials☆60Feb 21, 2025Updated last year
- [RSS 2026] Causal video-action world model for generalist robot control☆1,718Jul 9, 2026Updated 3 weeks ago
- Behavior Prompting Policy: Demonstrations as Prompts for Manipulation☆63Jul 17, 2026Updated 2 weeks ago
- CLAW🦀 is a web pipeline for scalable language-annotated whole-body motion generation on Unitree G1.☆56Jun 20, 2026Updated last month
- [RSS 2026] Code for RISE: Self-Improving Robot Policy with Compositional World Model☆332Jul 6, 2026Updated 3 weeks ago
- FieldGen is a semi-automatic data generation framework that enables scalable collection of diverse, high-quality real-world manipulation …☆27Oct 28, 2025Updated 9 months ago
- A real2sim evaluation framework for generalist policies☆224Jul 13, 2026Updated 3 weeks ago