論文 / arXiv:2610.06814
TAPDreamer: Transferable Adversarial Patches for World Action Models
PLAIN SUMMARY / やさしい要約
ロボットがカメラ映像を手がかりに動く仕組みは、映像の一部が意図的に変えられると誤作動することがあります。この研究は、画面の小さな一部分に貼る「パッチ」を、相手のロボットの出力を見なくても作れる方法を提案します。しかも、そのパッチが別の作業や別の仕組みにも効いてしまうことを示します。例えるなら、窓の隅に小さなシールを貼っただけなのに、部屋全体の見え方がずれてしまう、という話です。
AIが専門用語を使わずに書いた解説です。内容の正確さは、下の原文の要旨で確認してください。
ABSTRACT / 要旨(原文)
World models learn to predict how their environment will evolve, making them an important foundation for general-purpose robotic control. Yet world action models depend on camera inputs whose manipulation can corrupt the visual representations used across tasks and action policies. Existing attacks on these models optimize against the victim's actions or predicted futures and therefore require access to target-model outputs. In this paper, we propose an attack, TAPDreamer, against world action models that instead uses a public encoder alone to construct a fixed local perturbation that transfers across tasks and action architectures. TAPDreamer requires no target-policy queries. Our key insight is that interactions between patch-induced changes in attention weights and value vectors broadcast a nearly identical representation shift far beyond the patch footprint, and this shift remains stable across task observations. Guided by this insight, TAPDreamer uses six frames from one source task to maximize the global L1 distance between clean and patched encoder representations. In closed-loop evaluation, one frozen patch per benchmark, covering about 6.5% of the input, reduces FastWAM's success rate from 97.7% to 0.0% across 40 LIBERO tasks and from 90.8% to 0.0% across 50 RoboTwin tasks; matched random patches retain 81.5% and 79.2% success. The same patches reduce success to 2.1% and 0.8% on two DreamWAM configurations and to 10.0% on Motus. These results show that protecting downstream action generation alone is insufficient: defenses for world action models must also secure shared visual encoders against persistent local perturbations.
ここに表示しているのは原論文の要旨です。AIによる要約や解釈は含みません。
