FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
☆440Jul 21, 2026Updated this week
Alternatives and similar repositories for FlashRT
Users that are interested in FlashRT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Updated this week
- Running VLA at 30Hz frame rate and 480Hz trajectory frequency☆591Feb 10, 2026Updated 5 months ago
- ☆89May 27, 2026Updated last month
- High-performance, light-weight C++ LLM and VLM Inference Software for Physical AI☆481Jul 3, 2026Updated 2 weeks ago
- ☆135Mar 31, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Pragmatic VLA Foundation Model☆1,655Jun 11, 2026Updated last month
- ML kernels and benchmarking infrastructure written in TIRx☆67Updated this week
- Real-Time VLAs via Future-state-aware Asynchronous Inference.☆438Apr 22, 2026Updated 2 months ago
- Galaxea's open-source VLA repository☆689Jul 11, 2026Updated last week
- A lightweight toolkit for quantitatively scoring LeRobot episodes.☆70Mar 13, 2026Updated 4 months ago
- RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI☆4,179Updated this week
- ☆23Jun 29, 2026Updated 3 weeks ago
- FA4-based Relative Attention Kernel developed by TML and Colfax☆17Updated this week
- ☆22May 11, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- An Optimizer for Nvidia Compilers.☆107Jul 3, 2026Updated 2 weeks ago
- A performance analysis tool for VLA models☆77Feb 26, 2026Updated 4 months ago
- ☆32Jul 2, 2025Updated last year
- LeflexiTac: Giving robots a sense of touch.☆108May 6, 2026Updated 2 months ago
- ☆12,909Jun 16, 2026Updated last month
- ☆310Jun 9, 2026Updated last month
- An agent harness that compiles a model into one provably-correct, self-retargeting CUDA megakernel and self-tunes it past cuBLAS at batch…☆119Jun 29, 2026Updated 3 weeks ago
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing☆3,251Updated this week
- ☆54Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- mKernel: fast multi-node, multi-GPU fused kernels☆252Jun 21, 2026Updated last month
- ☆160Updated this week
- ☆760Jun 2, 2026Updated last month
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆21Nov 28, 2025Updated 7 months ago
- An all-in-one VLA engineering platform for embodied AI — from data to real-robot deployment.☆554Updated this week
- Official Repository for MolmoAct2☆678Updated this week
- ☆156May 24, 2026Updated last month
- Official repo for "StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation"☆29Jun 29, 2026Updated 3 weeks ago
- Some funny cute/cuteDSL code snippets☆33Mar 2, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official code of RDT 2☆795Feb 7, 2026Updated 5 months ago
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.☆534Updated this week
- FlashInfer: Kernel Library for LLM Serving☆5,994Updated this week
- FastCrest Tether: the OSS edge-to-cloud AI deploy CLI. Optimize, verify, deploy across Jetson, RTX, Apple Silicon, AMD. Hybrid edge-cloud…☆75Jul 13, 2026Updated last week
- Building General-Purpose Robots Based on Embodied Foundation Model☆1,180Updated this week
- High performance RMSNorm Implement by using SM Core Storage(Registers and Shared Memory)☆30Jan 22, 2026Updated 5 months ago
- FlashKDA: high-performance Kimi Delta Attention kernels☆464May 26, 2026Updated last month