ContextArticle
Dylan Castillo conducted a rigorous study of whether AI labs have been deliberately training models to draw pelicans riding bicycles.
The author highlights Dylan Castillo's methodical investigation into the 'pelicanmaxxing' question, contrasting it with the author's own ad-hoc spot-checking. ✦ AI generated
the author (via the article's title/description, the voice is an anonymous blogger or commentator) · Simon Willison's Weblog · 2026-07-22 · original ↗
Excellent piece of work by Dylan Castillo, who took a deep-dive into the frequently pondered question of whether the AI labs have been deliberately training models to draw pelicans riding bicycles in response to my deeply unscientific benchmark.
Read full article ↗excerpt · fair-use quotation
Around this claim
Context · 2
Dylan's methodology tested 8 animals × 6 vehicles across 7 different models with repeated runs and automated evaluation.the author (via the article's title/description, the voice is an anonymous blogger or commentator) · Simon Willison's Weblog · conf 90%GLM-5.2 showed the largest individual boost on the exact pelican-bicycle combination, but the effect was small and not statistically significant.the author (via the article's title/description, the voice is an anonymous blogger or commentator) · Simon Willison's Weblog · conf 85%