arXiv AI By Shengtao Zheng, Kai Li, Weichen Zhang, Yu Meng, Chen Gao, Xinlei Chen, Yong Li, Xiao-Ping Zhang

WorldFly: A World-Model-Based Vision-Language-Action Model for UAV Navigation

Read the original on arXiv AI →

arXiv:2606. 06147v1 Announce Type: new Abstract: End-to-end Vision-Language-Action (VLA) models have shown promise in UAV navigation.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.