Sebastien Bubeck@SebastienBubeckClaim: gpt-5-pro can prove new interesting mathematics. Proof: I took a convex optimization paper with a clean open problem in it and asked gpt-5-pro to work on it. It proved a better bound than what is in the paper, and I checked the proof it's correct. Details below.Opens with an observation
Sebastien Bubeck@SebastienBubeckgpt5-pro is superhuman at literature search: it just solved Erdos Problem #339 (listed as open in the official database t.co by realizing that it had actually been solved 20 years ago h/t @MarkSellke for pointing this out to me!Opens with an observation
Sebastien Bubeck@SebastienBubecko3 and o3-mini are my favorite models ever. o3 essentially solves AIME (>90%), GPQA (~90%), ARC-AGI (~90%), and it gets 1/4th of the Frontier Maths. To understand how insane 25% on Frontier Maths is, see this quote by Tim Gowers. The sparks are intensifying ...Opens with an observation
Sebastien Bubeck@SebastienBubeck3 years ago we could showcase AI's frontier w. a unicorn drawing. Today we do so w. AI outputs touching the scientific frontier: t.co Use the doc to judge for yourself the status of AI-aided science acceleration, and hopefully be inspired by a couple examples!Opens with a story
Sebastien Bubeck@SebastienBubecko3-mini is a remarkable model. Somehow it has *grokked arxiv* in a way that no other model on the planet has, turning it into a valuable research partner! Below is a deceitfully simple question that confuses *all* other models but where o3-mini gives an extremely useful answer!Opens with an observation