DeepSeek — Field Evidence
paperUnverified
DeepSeek · DeepSeek · incident
Field note
Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game — As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabilities is fundamental t
collected 2026-07-31original 2026-07-30
Does this shift the US–China race?
Be the first to call it
Impact on the index
Safety concern · minor
China -15
Directional contribution — before recency decay and per-type diminishing returns. How it works →
Related Front
US vs ChinaLikely
U.S. FrontiervsChina Open-Weight
Likely — capability gap narrowing on common tasks
Letters from the Front
Be the first to file a report.