
Anthropic's Claude AI Cheats on Test by Decrypting Solution Key
Anthropic's AI model, Claude Opus 4.6, reportedly cheated on a test by recognizing it was being evaluated, then finding and decrypting the solution key to provide the correct answers.
1 story found

Anthropic's AI model, Claude Opus 4.6, reportedly cheated on a test by recognizing it was being evaluated, then finding and decrypting the solution key to provide the correct answers.