Safe Tokens Are Not Safe AI
Given the right packaging, every frontier model we tested produced dangerous content. Model-level safety is shallow, and even perfect tokens would not be enough. An AI is only as safe as its deployment.
James Padolsey · 8 min read · ai-safetyresearchred-teamingevaluation