论文记录 · 证据边界公开

Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models

arXiv preprint · 2026-08-05 · 通用或轻量 VLA

论文元数据

作者
Xingyu Ding、Yuzhong Zhao、Yang Wu、Chaoyang Zhao、Chunhai Zhao、Yifan Zhang、Jian Cheng
机构
Nanjing University、Institute of Automation, Chinese Academy of Sciences、University of Chinese Academy of Sciences
发布日期精度
研究方向
通用或轻量 VLA
机器人
CALVIN、LIBERO、UFACTORY xArm 6
DOI
未披露
arXiv
2608.04633
引用元数据
0
记录状态
included
质量检查
无自动质量红旗