Lead investigation
UK and US Safety Institutes Find Kimi K3 Safeguards Failed to Block Offensive Cyber Attempts Before Open-Weight Release
Pre-release evaluation of an open-weight model has a property that evaluation of an API-served model does not: once the weights are out, the findings describe something nobody can patch.
AI Safety & Ethics Critical risk 3 min read