Post by Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞) on X
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞)@teortaxesTex
X> We propose IndexShare, which reuses the same indexer across every four sparse attention layers, reducing per-token FLOPs by 2.9× at a 1M context length. We also improve GLM-5.2’s MTP
Good demonstration that a lot could be squeezed out of "mere" DSA
on benchmarks, too.

152 likes8 repliesPosted Jun 16, 2026