Preprint
Aug 2026
FlashVLA: Streaming Action Decoding for Fast and Asynchronous VLA Inference
FlashVLA is introduced, a streaming action decoding framework that addresses both challenges in a unified formulation and can achieve $\geq$30\,Hz control frequency on a single GPU with smooth asynchronous inference in real-world deployment.
Ze-Kai Li, Jiarui Tang, Zhijian Liu
· 4 citations