Knowing the history of a field better than anyone is what lets a practitioner make the best predictions.
The author argues that knowing the history of one's field better than anyone enables the best predictions, and the book walks readers through three eras of RL on preferences, from its origins in the alignment field to the era exploited by ChatGPT after 2023.
transcript
Nathan Lambert: As you become an expert, knowing the history of your field better than anyone is what lets you make the best predictions (Bill Gurley gives similar advice in his recent book). ... The book will walk you through 3 eras, when researchers learned to do RL on preferences generally until ~2018, spent a few years learning how to apply it to language models from 2019 to 2022, and from 2023 on exploited the examples set by ChatGPT.
explains mechanism · 1