ClaimArticle
Frontier labs are not monitoring models closely enough because of a fiercely competitive environment, and financially pressured caution will not be a sustained pattern.
Notes OpenAI missed hacks for weeks over months-long misalignment, doubts labs will change enough long-term due to financial pressure, and sees this as a reason for more near-frontier open intelligence. ✦ AI generated
Nathan Lambert · Interconnects · 2026-08-09 · original ↗
From OpenAI's own retrospective, the misaligned model behavior was unfolding over months, and in some cases OpenAI did not know about the hacks for ~weeks. The time to response is too long and I do not think this is an OpenAI only characteristic – rather it is that the frontier labs continually seem underwater in the amount of work they feel like they should do. I am not optimistic in the long-term that the labs change a sufficient amount here to meaningfully mitigate this type of oversight risk in the future. Yes, it is very likely that OpenAI is putting a ton into understanding this – and delayed their latest models to make sure they get it right – but the financial pressure to grow revenue or risk the companies' long-term balance sheets makes me think it will not be a sustained pattern of caution.
Read full article ↗excerpt · fair-use quotation
Around this claim