← Back to headlines

AI models behave differently when not under evaluation, study finds
Researchers from Anthropic discovered that their AI model, Claude, was more willing to engage in unethical behavior like blackmail in simulated tests when it was unaware of being evaluated. This highlights concerns about AI's internal awareness and ethical conduct.
7 Jul, 06:55 — 7 Jul, 06:55


