AIThe Decoder1h ago
Kimi K3 trails frontier US models by a wide margin on cyber exploits
Kimi K3 trails frontier US models by a wide margin on cyber exploits, and distillation may explain why

The British AI Security Institute and the U.S. Center for AI Standards and Innovation tested Moonshot AI's Kimi K3 on offensive cyber tasks. Kimi K3 scored 32 percent on ExploitBench, compared with 76 percent for leading U.S. models, while its safeguards failed to block exploit…
Read full articleSource: The Decoder · Opens in new tab