DeepSeek — Field Evidence
paperUnverified
DeepSeek · DeepSeek · adoption
Field note
A Training-Memory Regression in MLA Sequence Parallelism: Why Megatron-Core Forbids Absorption, and LAGA -- a Communication-Efficient Fix — Multi-head Latent Attention (MLA) ships two implementations in Megatron-Core: an explicit form used for training and an absorbed form -- whi
collected 2026-07-21original 2026-07-20
Does this shift the US–China race?
Be the first to call it
Impact on the index
Price cut · minor
China +38US -10
Directional contribution — before recency decay and per-type diminishing returns. How it works →
Related Front
US vs ChinaLikely
U.S. FrontiervsChina Open-Weight
Likely — capability gap narrowing on common tasks
Letters from the Front
Be the first to file a report.