Post by Lisan al Gaib on X
Lisan al Gaib@scaling01
XIts so over for OpenAI and Anthropic.
Gemini 3 Pro Benchmarks
37.5% on HLE
31.1% on ARC-AGI-2
2439 Elon on LiveCodeBench Pro
85.4% on Tau-Bench
72.1% on SimpleQA Verified
SOTA everywhere except SWE-Bench Verified

1.5K likes103 repliesPosted Nov 18, 2025