Qwen — Field Evidence
paperUnverified
Qwen · Qwen · praise
Field note
MyoCardBench: A Real-World Data Benchmark for Evaluating Large Language Models in Clinically Authentic Cardiovascular Care Scenarios — Background: Most medical large language model (LLM) benchmarks focus on examination knowledge or isolated tasks and may not reflect the longitudi
collected 2026-07-29original 2026-07-28
Does this shift the US–China race?
Be the first to call it
Impact on the index
Benchmark win · minor
China +30US -7
Directional contribution — before recency decay and per-type diminishing returns. How it works →
Related Front
US vs ChinaLikely
U.S. FrontiervsChina Open-Weight
Likely — capability gap narrowing on common tasks
Letters from the Front
Be the first to file a report.