Post by Ryan Lowe on X
Ryan Lowe@ryan_t_lowe
Xo3 seems to hallucinate >2x more than o1, according to the system card
so hallucinations could scale *inversely* with increased reasoning (unlike for increased model size), bc outcome-based optimization incentivizes confident guessing
(the Transluce example is kinda hilarious)

481 likes19 repliesPosted Apr 16, 2025