UK AISI / CAISI Preliminary Assessment of Kimi K3's Cyber Capabilities

NIST – AI News (topic 2753736)(Multi) 23 Jul 2026 62

A joint UK-US pre-release cyber evaluation of a PRC frontier model sets a benchmark for AI safety testing that Australian agencies and AISI should track.

  • UK AISI and US CAISI jointly assessed Kimi K3's cyber capabilities, finding it below leading US models but ahead of prior open-weight models.
  • Kimi K3 autonomously completed a simulated corporate network attack in 1 of 10 attempts, signalling growing open-weight cyber risk.
  • Australia's AISI is absent from this joint evaluation - a notable gap as peer safety institutes deepen bilateral testing collaboration.
  • Monitor Australia's AISI and cyber security policy teams may want to monitor the UK-US joint evaluation program as open-weight model cyber capabilities continue to advance.
  • Consider Agencies with cyber risk responsibilities could consider whether open-weight model capability benchmarks like ExploitBench and TLO could inform threat assessments for AI-enabled cyber attacks on government systems.

Implications are AI-generated. Starting points, not advice — see methodology for how they're framed.

View original source