Post by Max Weinbach on X
Max Weinbach@mweinbach
XKimi K3 being 2.8T parameter and 1M context is cool but show me the sparsity, show me the price
How quick can the inference providers scale this to 200 tok/s
This is what I care about!!!! Efficient HUGE models
312 likes15 repliesPosted Jul 16, 2026