The value of a conceptual breakthrough like Galois theory is verified on a timescale of a century, through a chain of human judgment rather than any immediate reward — meaning current RLVR-style reward functions are structurally unable to recognize the 'Galois instinct'.
Grant walks through the history from Lagrange's questions about symmetry, to Abel's proof, to Galois's rejected papers, to Liouville and Jordan, and finally Gell-Mann's group-theoretic prediction of quarks. He argues that recognizing a great new conceptualization is a roughly century-long verification loop through many human judgments — the opposite of an RLVR feedback signal. ✦ AI generated
Grant Sanderson · Dwarkesh Podcast · 2026-06-30 · original ↗
plays this moment only · 11:32 — 26:12
“If you wanted to do a verification loop on whether group theory is an interesting concept—was something useful done here, or why is this proof better?—potentially that verification loop is a hundred years long.”
What makes Galois theory such an interesting example is that you literally have this hundred-year segment of an idea that flows through many different people's heads before it settles into something the math community agrees is good. With Einstein and GR, people could feel this was a good theory right away. You have to ask, what is the way of measuring progress that's not based on solving a problem, but that is somehow capturing the instinct inside Galois's mind when he says, 'I think there's something here'? What's the instinct inside Lagrange's mind when he says, 'I think this is the right way to think about it'? What's the instinct inside Liouville's mind when he says, 'These scattered notes from this long-dead youngster might have something to them'? It's so hard to put a finger on that.
verbatim transcript · starts at 11:32
11:32– The verification loop on conceptual breakthroughs can be a century long
26:12– Will we understand an AI proof of the Riemann hypothesis?
11:32– The verification loop on conceptual breakthroughs can be a century long
26:12– Will we understand an AI proof of the Riemann hypothesis?
- ·Galois theory's value took ~100 years to confirm
- ·Idea flowed through many minds before acceptance
- ·Chain of human judgment, not immediate reward
- ·Einstein's GR felt right immediately to peers
- ·Galois needed Lagrange→Abel→Galois→Liouville→Jordan→Gell-Mann
- ·Gell-Mann's quark prediction validated the conceptual thread
- ·No single reward signal within a century-long chain
- ·Instinct says 'there's something here' without proof
- ·Current RLVR structurally blind to this kind of insight