FlashDrive: Flash Vision-Language-Action Inference for Autonomous Driving

arXiv CS.AI
Robotics AI Hardware

Vision-Language-Action (VLA) models promise to bring end-to-end reasoning to autonomous driving, but their computational cost remains far too high for real-time control. Layered on system-level CUDA Graph compilation and kernel fusion, these techniques compound.

Related Articles