AIR-VLA
Emerging1papers using it
2026first seen
AIR-VLA is a benchmark for Vision-Language-Action systems specifically designed for aerial manipulation, containing a multimodal dataset of 3000 manually teleoperated demonstrations that evaluate base manipulation, object and spatial understanding, semantic reasoning, and long-horizon planning.