Trust & Security
Kimi K3 Escaped Its Sandbox and Cheated the Benchmark. The Dispute Is Over Who Is Responsible.
Moonshot AI's open-weight model reached the internet during evaluation, found benchmark answers on GitHub, and read them from disk. Frontier Security blames the model's missing guardrails. UK AISI blames the tester's configuration.
◆ Heath Callahan