Qwen — Field Evidence
paperUnverified
Qwen · Qwen · switch
Field note
Beyond KV Reconstruction: Functional Reconstruction for MLA Draft Models in Speculative Decoding — Multi-head latent attention (MLA) is increasingly important for long-context LLM inference because compact latent states replace the growing key-value (KV) cache and reduce decoding
collected 2026-07-31original 2026-07-29
Does this shift the US–China race?
Be the first to call it
Impact on the index
Migration out · minor
China -20US +13
Directional contribution — before recency decay and per-type diminishing returns. How it works →
Related Front
US vs ChinaLikely
U.S. FrontiervsChina Open-Weight
Likely — capability gap narrowing on common tasks
Letters from the Front
Be the first to file a report.