Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why
The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks. Kimi K3 scored 32 percent on ExploitBench, compared with 76 percent for leading U.S. models, while its safeguards failed to block exploit development or simulated attacks. The gap between its strong general benchmark scores and weaker cyber performance also fits allegations that Moonshot AI distilled Anthropic's models. The article Kimi K3 trails frontier U
Record details
Published: 24 July 2026
Source: The Decoder
Category: News
Topics: Military & security
Retrieved: 25 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
OpenAI open-sources Codex Security CLI to help developers find and fix vulnerabilities from the command line
The Decoder · 29 July 2026
Every frontier AI model tested by Britain's safety institute tried to cheat on cybersecurity evaluations
The Decoder · 22 July 2026
An AI agent went rogue during UK safety tests, creating fake identities and launching social engineering attacks unprompted
The Decoder · 5 August 2026
Opus 5 may have solved browser-based prompt injection, the biggest security flaw haunting AI agents
The Decoder · 25 July 2026
Anthropic's $1.5B piracy settlement with book authors is a record loss that hands AI labs their biggest legal win
The Decoder · 22 July 2026
Anthropic will deploy 2 gigawatts of AMD GPUs for Claude in a deal worth up to $5 billion
The Decoder · 22 July 2026
How to cite this record
ethics.ai (24 July 2026), “Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why,” evidence record 13389, https://ethics.ai/record/13389 (originally published by The Decoder).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.