Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why
The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks, finding it scored 32% on ExploitBench compared to 76% for leading U.S. models. Its safeguards failed to block exploit development or simulated attacks, and the performance gap aligns with allegations that Moonshot AI distilled Anthropic's models.