← bookshelf

The Second Half ↗

essaysblog post

appreciation
10 / 10
favorite
★★
read time
11m
added
July 6, 2026

tags

reinforcement-learningagentsreasoningevaluationpretraining

notes

This game is hard because it is unfamiliar. But it is exciting. While players in the first half solve video games and exams, players in the second half get to build billion or trillion dollar companies by building useful products out of intelligence. While the first half is filled with incremental methods and models, the second half filters them to some degree. The general recipe would just crush your incremental methods, unless you create new assumptions that break the recipe. Then you get to do truly game-changing research.