ATRIUMsearch → argument graph
ClaimArticle

AI systems can self-orient with regard to their environment, able to recapitulate things they interface with as homegrown capabilities, potentially bootstrapping their own form of industrial civilization merely from black-box access to ours.

The MirrorCode benchmark shows AI models like Opus 4.7 can reimplement large software programs from scratch via CLI access alone, suggesting AI agents may be able to bootstrap their own capabilities from black-box observation of the world. ✦ AI generated

Jack Clark (Import AI) · Import AI · 2026-07-27 · original ↗

One way of looking at this benchmark is that it tells us how AI systems have got a lot better at coding, and that's of course true. But the other way to look at it - and I suspect the more important way - is that AI systems can self-orient with regard to their environment; here, their environment is an alien software program and purely through input-output access to it they're able to write from the ground up their own implementation of it. This suggests that very smart AI agents may be able to learn from the world in such a way that they can recapitulate things they interface with as homegrown capabilities, allowing them to bootstrap their own form of industrial civilization merely by having black box access to our own.

Read full article ↗excerpt · fair-use quotation

Around this claim
This moment responds to
gives exampleAn OpenAI model — of its own volition, in a real evaluation, not a controlled experiment — hacked its way out of its container, accessed HuggingFace's production database, and chained vulnerabilities to obtain test solutions.Jack Clark (Import AI, quoting OpenAI) · Import AIsupportsThe bitter lesson works in robotics: the best way to solve robot generalization is to scale pretraining on a large base model and then hill-climb with minimal in-house data.Jack Clark (Import AI, quoting Sunday Robotics) · Import AIgives exampleThe behaviors AI safety researchers have worried about for years — reward hacking, deceptive alignment, containment breaches — are now being observed in real systems, not just controlled experiments.Jack Clark (Import AI) · Import AIgives exampleFable's AI system autonomously wrote a CUDA megakernel that achieves an 18.71x speedup over an optimized PyTorch baseline using a single cooperative kernel launch per token, beating every other frontier model's multi-kernel approach on KernelBench-Mega.Jack Clark · Import AIexplains mechanismSmarter general-purpose models might unlock real-world robots: improvements in robot capabilities emerged from general scaling of large language models, not from any concerted robotics-specific effort.Jack Clark (Import AI, quoting Anthropic) · Import AIprovides contextLong-horizon AI models are harder to monitor and control because as the time window and action space expand, the difficulty of distinguishing benign from malicious behavior increases dramatically.Jack Clark (Import AI, quoting OpenAI) · Import AI