FlashDrive: Flash Vision-Language-Action Inference for Autonomous Driving
arXiv CS.AI
•
Robotics
AI Hardware
Vision-Language-Action (VLA) models promise to bring end-to-end reasoning to autonomous driving, but their computational cost remains far too high for real-time control. Layered on system-level CUDA Graph compilation and kernel fusion, these techniques compound.