A quote from gwern

16th January 2025

[...] much of the point of a model like o1 is not to deploy it, but to generate training data for the next model. Every problem that an o1 solves is now a training data point for an o3 (eg. any o1 session which finally stumbles into the right answer can be refined to drop the dead ends and produce a clean transcript to train a more refined intuition).

— gwern

Posted 16th January 2025 at 7:21 pm

Simon Willison’s Weblog

Recent articles