Fetching the paper…

DepthVLA: Enhancing Vision-Language-Action Models with Depth-Aware Spatial Reasoning · Around