← all datasets

AIR-VLA

Emerging
1papers using it
2026first seen

AIR-VLA is a benchmark for Vision-Language-Action systems specifically designed for aerial manipulation, containing a multimodal dataset of 3000 manually teleoperated demonstrations that evaluate base manipulation, object and spatial understanding, semantic reasoning, and long-horizon planning.

Papers using AIR-VLA (1)

AIR-VLA dataset β€” papers, benchmarks & downloads Β· Robotics