andrew gao@itsandrewgaohere's your DEEP DIVE into @grok's architecture! I just went through the t.co for this 314B open source behemoth with *no strings attached*. 👇🧵Opens with an observation
andrew gao@itsandrewgaouh.... gpt2-chatbot just solved an International Math Olympiad (IMO) problem in one-shot the IMO is insanely hard. only the FOUR best math students in the USA get to compete prompt + its thoughts 🧵Opens with an observation
andrew gao@itsandrewgaodon't write this off as "fast, non-frontier-lab model == dumb & not worth my time" it's smarter than the SOTA models were this summer and also way faster (more chances to iterate/fix in same time, less waiting) 1 pt of reference: SWE-1.5 > GPT-5 (high) on SWE-Bench Pro!Opens with a "stop doing this"
andrew gao@itsandrewgaokeep in mind: the August snapshot of Opus 4.1 scored 22.71% on SWE-Bench Pro. SWE-1.5 (13x faster than sonnet 4.5) scores ~2x higher than the SOTA code model from only 2 months ago. made possible by owning the stack: model, inference, & agent harness! agent lab era is here!Opens with an observation