Robot-centric pointmaps align VLA perception with the robot frame

Loading video
Loading videoRobot-Centric Pointmaps encode every pixel as 3D geometry in the robot frame, removing the camera-to-robot coordinate mismatch VLAs must otherwise learn for every camera movement.
Category: research
Author: @shenzhenfoundry
Date: 2026-08-12T00:00:00
Duration: 85.133s
Reference: https://arxiv.org/abs/2607.11498





