China’s Kimi K3 ‘significantly below’ US rivals in hacking power, study shows
Kimi K3’s overall score of 32.2 per cent compares with top, unnamed US models that average 76.2 per cent

Chinese unicorn Moonshot AI’s Kimi K3 model trails far behind top American rivals in its ability to launch cyberattacks, joint British-US government research shows, challenging Washington’s brewing anxiety over the rapid rise of Chinese open-source artificial intelligence.
The model, currently considered China’s most powerful large language model, performs “significantly below the most recent frontier cyber-capable models”, according to a report published on Thursday by the UK Artificial Intelligence Security Institute (AISI) and the US Centre for AI Standards and Innovation (CAISI).
The AISI is a research arm under the UK Department for Science, Innovation and Technology, while the CAISI operates within the US Department of Commerce’s National Institute of Standards and Technology.
To evaluate capabilities, the two government bodies put Kimi K3 through ExploitBench, a public benchmark assessing an AI’s ability to develop exploits for cybersecurity vulnerabilities.
Kimi K3 achieved an overall score of 32.2 per cent, outperforming domestic rival Zhipu AI’s GLM-5.2 at 24.4 per cent, but lagging well behind top, unnamed US models that averaged 76.2 per cent.
Notably, Kimi K3 failed to achieve arbitrary code execution – the highest-level exploit granting full control of a target system – across all 41 ExploitBench tasks, whereas leading US models achieved it on 20 tasks.
