arxivcs.RO2026-07-03
Exp2VLA: Enabling Vision-Language-Action for Drone Navigation from Expert Demonstrations
Van Huyen Dang, Kabilesh Rajendran, Erdi Sayar, Erdal Kayacan
Vision-language-action (VLA) models open a new path toward intuitive robot control by directly linking perception, language, and action in a single end-to-end framework. Yet for UAVs, practical adoption remains difficult because existing solutions are either computationally heavy…