UK US Safety Institutes Find Kimi K3 Trails Frontier Models
Joint safety institutes assess Chinese model against American leaders on cyber benchmarks.
Entities: Kimi K3, CAISI, U.S. Department of Commerce, AI Security Institute (AISI), David Sacks, Howard Lutnick
UK AISI and US CAISI conducted a preliminary evaluation of Kimi K3 cyber capabilities on the Last Ones range. The model underperformed leading US frontier AI systems, reaching performance levels comparable to models released about six months earlier. Reactions from figures including Elon Musk, David Sacks, and researchers emphasized that American models maintain a lead, with the gap potentially larger when including unreleased lab systems. The assessment highlights ongoing international efforts to benchmark AI safety and capabilities without claiming any model has reached parity with current US leaders.
CAISI’s latest blog post evaluates Kimi K3 and its cyber capabilities. Based on a preliminary cyber-focused evaluation, Kimi K3 performed significantly below the leading U.S. frontier AI models. https://www.nist.gov/news-events/news/2026/07/uk-aisi-caisi-preliminary-assessment-kimi-k3s-cyber-capabilities