ATRIUMsearch → argument graph
ClaimArticle

Models have improved rapidly in coding, mathematics, search, and research, but writing well has barely advanced because it is a difficult, relatively orthogonal skill with inadequate specialized training data.

The author expected much greater progress in nonfiction writing, given the models' advances in other tasks. Specialized harnesses and prompts may yield incremental gains, but the fundamental difficulty of writing and the lack of suitable training interventions limit improvement. ✦ AI generated

the author · Interconnects · 2026-08-12 · original ↗

I would’ve expected way more progress on non-fiction writing from the models. I almost thought I would look dumb publishing a non-fiction book in 2026, given how things looked in 2024. Today, some of the most famous models on writing ability are pretty old, examples include OpenAI’s big GPT 4.5 and Moonshot’s Kimi K2. In and around these releases, the models have gone from okay to superhuman at other tasks like coding and mathematics. Maybe a closer, but still imperfect, comparison is how the models went from incapable to decent at search and research tasks. The pace of progress on most other skills is steep, but writing well feels orthogonal to most of them. I do not think writing is just ignored, but rather it’s challenging and lacks good training data to specifically intervene on it. There is certainly some low-hanging fruit for making AI models better at writing — such as specialized harnesses like Claude Code, prompts, and training environments that make models spend a lot more inference tokens on the output, but I don’t think these will have a multiplicative impact on ability. Writing well is a very hard task! It’s a shame that we haven’t unlocked inference-time scaling for one of the great intellectual pursuits. Regardless, writing seems very different than what the models are good at.

Read full article ↗excerpt · fair-use quotation

Around this claim