04:00
2026-08-18
arxiv.org
artificial-intelligence
AeroGround: A Comprehensive Benchmark for Aerial-Ground Collaborative Reasoning
A new benchmark, AeroGround, evaluates vision-language models (VLMs) on aerial-ground collaborative reasoning tasks, revealing a significant performance gap: the best model achieves 54.4% average accu…