论文记录 · 证据边界公开

WorldBagel: Uncovering the Power of Unified Multimodal Models for Vision-Language-Action-World Modeling

arXiv manuscript · 2026-07-03 · 世界模型与动作条件预测

论文元数据

作者
Zelin Zhao、Min Shi、Bo Yuan、Haotian Xue、Jialuo Li、Lama Moukheiber、Humphrey Shi、Yongxin Chen
机构
机构尚未补齐
发布日期精度
研究方向
世界模型与动作条件预测
机器人
Franka Isaac Sim、Language Table recorded robot data
DOI
未披露
arXiv
2607.03461
引用元数据
0
记录状态
included
质量检查
无自动质量红旗