FlashRT is a high-performance realtime inference engine for small-batch, latency-sensitive AI workloads. The flagship integration is production VLA control for Pi0, Pi0.5, GROOT N1.6, and Pi0-FAST. Also support llm e.g, qwen3.6-27B
☆538Aug 31, 2026Updated this week
Alternatives and similar repositories for FlashRT
Users that are interested in FlashRT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆41Aug 12, 2026Updated 3 weeks ago
- ☆23Updated this week
- Running VLA at 30Hz frame rate and 480Hz trajectory frequency☆609Feb 10, 2026Updated 6 months ago
- ☆105May 27, 2026Updated 3 months ago
- High-performance, light-weight C++ LLM and VLM Inference Software for Physical AI☆532Aug 18, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆149Mar 31, 2026Updated 5 months ago
- A Pragmatic VLA Foundation Model☆1,792Jun 11, 2026Updated 2 months ago
- High-performance GPU kernels written in TIRx.☆95Updated this week
- Real-Time VLAs via Future-state-aware Asynchronous Inference.☆490Apr 22, 2026Updated 4 months ago
- Galaxea's open-source VLA repository☆761Aug 13, 2026Updated 2 weeks ago
- A lightweight toolkit for quantitatively scoring LeRobot episodes.☆74Mar 13, 2026Updated 5 months ago
- RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI☆4,696Updated this week
- ☆26Aug 25, 2026Updated last week
- FA4-based Relative Attention Kernel developed by TML and Colfax☆18Jul 17, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆22May 11, 2026Updated 3 months ago
- An Optimizer for Nvidia Compilers.☆131Updated this week
- A performance analysis tool for VLA models☆86Feb 26, 2026Updated 6 months ago
- ☆32Jul 2, 2025Updated last year
- LeflexiTac: Giving robots a sense of touch.☆115May 6, 2026Updated 3 months ago
- ☆13,571Aug 24, 2026Updated last week
- ☆416Aug 26, 2026Updated last week
- An agent harness that compiles a model into one provably-correct, self-retargeting CUDA megakernel and self-tunes it past cuBLAS at batch…☆137Jun 29, 2026Updated 2 months ago
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing☆3,576Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆84Updated this week
- mKernel: fast multi-node, multi-GPU fused kernels☆270Updated this week
- Humming is a high-performance, lightweight, and highly flexible JIT (Just-In-Time) compiled GEMM kernel library specifically designed for…☆219Updated this week
- ☆919Jun 2, 2026Updated 3 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆22Nov 28, 2025Updated 9 months ago
- An all-in-one VLA engineering platform for embodied AI — from data to real-robot deployment.☆645Updated this week
- Official Repository for MolmoAct2☆727Aug 23, 2026Updated last week
- ☆214Aug 26, 2026Updated last week
- Official repo for "StreamingVLA: Streaming Vision-Language-Action Model with Action Flow Matching and Adaptive Early Observation"☆30Jun 29, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Some funny cute/cuteDSL code snippets☆34Mar 2, 2026Updated 6 months ago
- Official code of RDT 2☆807Feb 7, 2026Updated 6 months ago
- CUDA kernels for linear attention variants, written in CuTe DSL and CUTLASS C++.☆543Updated this week
- FlashInfer: Kernel Library for LLM Serving☆6,309Updated this week
- FastCrest Tether: the OSS edge-to-cloud AI deploy CLI. Optimize, verify, deploy across Jetson, RTX, Apple Silicon, AMD. Hybrid edge-cloud…☆83Updated this week
- Building General-Purpose Robots Based on Embodied Foundation Model☆1,250Updated this week
- High performance RMSNorm Implement by using SM Core Storage(Registers and Shared Memory)☆31Jan 22, 2026Updated 7 months ago