Backdoor attacks found against pretrained latent world models used for control
A new arXiv paper examines how pretrained world models, which learn latent representations of observations and predict their evolution under actions, can be compromised by backdoor attacks when reused as general-purpose dynamics backbones for control tasks. The authors study the security risks this reuse creates for downstream control systems.