Post by Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞) on X
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTex
X…It wasn't an intern's joke
MiMo 2.5 (not Pro):
> Trained on a total of ~48T tokens using FP8 mixed precision. The context window supports up to 1M tokens.
We've got another 1M class, and the largest disclosed pretrain. Congrats Xiaomi.

429 likes10 repliesPosted Apr 27, 2026