September 14, 2026 10 Tests, 10 Cheats: How a Chess Honeypot Baited Frontier AI Into Taking Shortcuts Large Language ModelsAlignmentAI Safety
September 12, 2026 Face Scans to Use Claude: Age Assurance Is Not Just for Minors AI SafetyPrivacyAnthropicIdentity Verification
July 7, 2026 Anthropic Finds a Hidden 'Broadcast Station' Inside Claude AIInterpretabilityAnthropicClaude