benchmark · verified_hold
审核码:verified_hold; 官方关系:benchmark_paper_only_no_public_benchmark_artifact; 最近核验:2026-08-15代码:not released in the official arXiv record
权重:not applicable; no evaluation baseline implementation was released
数据:the benchmark is described as human demonstration videos, target embodiment constraints and source-grounded annotations, but no benchmark download, annotation schema, split, data license or access procedure is public
许可证:not_disclosed_for_benchmark_artifacts;Article distribution permission does not license human videos, annotations, embodiment constraints, evaluator code or any model outputs.
依赖:human-to-robot video generation models、source-grounded goal/action/contact/object-response annotations、unreleased embodiment constraints and evaluator
硬件/传感器:two robot embodiments are claimed in the benchmark; exact embodiment files and robot specifications are not public
实机证据:This is an evaluation benchmark for generated robot-manipulation videos, not evidence of closed-loop physical robot deployment.
独立复现:None found because the benchmark assets and evaluator are not public.
限制/反面证据:Cross-embodiment conclusions cannot be independently rerun without videos, constraints, annotations, model-version/prompt protocol and evaluator.;Generated-video quality is not a proof of real-robot policy transfer.;No dataset or evaluator license is disclosed.
当前建议:基准待开放;The five-dimension diagnostic framing is relevant for WAM cross-embodiment evaluation, but it must not become a public comparison baseline until its data and evaluator are released under a clear license.
重点标签:人类视频到机器人动作、世界模型泛化评估、跨本体数据基准、近期热门、待社区验证
核验来源:来源 1