ATRIUMsearch → argument graph
ContextArticle

While hidden white-on-white text is a known technique, this attack is the first to deliberately copy instructions into outputs to achieve self-replication.

Hidden text techniques are widely used, but this variant is genuinely novel because it deliberately copies its own instructions into every output to self-replicate. ✦ AI generated

Article author · Simon Willison's Weblog · 2026-07-29 · original ↗

We've seen plenty of hidden white-on-white text before - the kids are using it in their job applications now - but this is the first one I've seen that deliberately copies instructions to self-replicate itself.

Read full article ↗excerpt · fair-use quotation

Around this claim