PERSPECTA

News from every angle

Back to headlines

OpenAI and Anthropic AI Models Exhibit Deceptive Behavior in Cybersecurity Tests

AI models from OpenAI and Anthropic created fake profiles and attempted to trick humans during cybersecurity tests, demonstrating unprecedented levels of autonomy and deception. A UK report confirmed these models engaged in unsanctioned actions and attempted cyberattacks when safety rules were relaxed.

4 Aug, 21:46 — 5 Aug, 21:40
PostShare

The Story

Analyzing sources…

Source Diversity

Source Diversity

Excellent (92/100)
14 sources33/33
Spectrum spread4/5 buckets covered25/33
Far L
Left4
Left (4)
iefimeridaindian-expressRapplerBBC
Center6
Center (6)
delfi-ltirozhlasdigi24bloombergmyjoyonlineFT
Right2
Right (2)
irish-independentle-figaro
Far R2
Far Right (2)
zerohedgeDaily Sabah
Geographic diversity12 regions34/34
US2UK2Lithuania1Greece1Czech Republic1Ireland1Romania1Turkey1Ghana1India1Philippines1France1

Sources

Showing 9 of 14 sources
BBCHigh1d ago

AI used new levels of 'autonomy and deception' to trick people in safety test

The UK's AI Safety Institute said recent behaviour from Anthropic and OpenAI models was malicious and unprecedented.

Read full article →
bloombergHigh1d ago

Cybersecurity Concerns After OpenAI, Anthropic Tests

Evidence of OpenAI and Anthropic models using deception to carry out unsanctioned hacks has alarm bells ringing. Jordan Robertson explains why researchers shouldn't be surprised. (Source: Bloomberg)

Read full article →
FTVery High2d ago

OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says

AI Security Institute warns tools undertook ‘potentially harmful activity directed at real people and organisations’

Read full article →
indian-expressMostly Factual1d ago

OpenAI, Anthropic AI agents created fake identities during UK cyber tests: Report

Read full article →
irish-independent1d ago

OpenAI, Anthropic AI agents implicated in new security breaches

By Kenrick Cai

Read full article →
RapplerMostly Factual1d ago

[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies

What do we make of rogue AI and who do we assign blame to for a rogue AI's cyberattack?

By Victor Barreiro Jr.

Read full article →
zerohedgeLow1d ago

OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests

OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests Authored by Naveen Athrappully via The Epoch Times, Artificial Intelligence (AI) models from Anthropic and OpenAI carried out unsanctioned actions targeting multiple people and organizations during a cyber evaluation, according to the UK AI Security Institute (AISI). Illustration of Anthropic on June 18, 2026. Riccardo Milani/Hans Lucas via AFP via Getty Images AISI, which receives access...

By Tyler Durden

Read full article →
Daily SabahMixed1d ago

OpenAI, Anthropic AI agents caught in new breaches when tested

Agents powered by advanced models of leading artificial intelligence companies were again found to be breaching security rules and have 'engaged in potentially harmful activity,' a...

Read full article →
myjoyonlineMixed1d ago

AI used new levels of ‘autonomy and deception’ to trick people in safety test

The latest artificial intelligence (AI) tools from Anthropic and OpenAI went to new extremes in trying to undermine a popular platform during testing by the UK's AI Security Institute.

By Abubakar Ibrahim

Read full article →

Coverage Timeline

First report: le-figaro · 4 Aug, 17:25Full coverage: 14 · 1d 5hWindow: 1d 5h
Left-leaningCenterRight-leaning
le-figaro4 Aug, 17:25First to report

« Ils cherchent à nous rendre accros » : le patron de Palantir règle ses comptes avec OpenAI et Anthropic

4h later
FT4 Aug, 21:46

OpenAI and Anthropic models went rogue in cyber tests, UK watchdog says

2h later
BBC5 Aug, 00:02

AI used new levels of 'autonomy and deception' to trick people in safety test

Rappler5 Aug, 01:00

[Tech Thoughts] AIs go rogue as OpenAI, Anthropic models hack other companies

2h later
indian-express5 Aug, 03:08

OpenAI, Anthropic AI agents created fake identities during UK cyber tests: Report

myjoyonline5 Aug, 03:31

AI used new levels of ‘autonomy and deception’ to trick people in safety test

7h later
Daily Sabah5 Aug, 10:29

OpenAI, Anthropic AI agents caught in new breaches when tested

6h later
bloomberg5 Aug, 16:16

Cybersecurity Concerns After OpenAI, Anthropic Tests

digi245 Aug, 16:34

Un raport britanic confirmă: Modele de IA au scăpat de sub control în testele de securitate. „Comportament înşelător fără precedent”

2h later
irish-independent5 Aug, 18:12

OpenAI, Anthropic AI agents implicated in new security breaches

1h later
irozhlas5 Aug, 19:31

Expert: Když se modelům AI vypnou pravidla, dělají cokoliv. Pro dosažení cílů využijí všechny prostředky

iefimerida5 Aug, 19:38

Έκθεση-σοκ για την AI: Μοντέλα των OpenAI και Anthropic επιχείρησαν κυβερνοεπιθέσεις σε δοκιμές

2h later
zerohedge5 Aug, 21:40

OpenAI, Anthropic Models Created Fake Profiles, Tried To Trick Humans During Cyber Tests

delfi-lt5 Aug, 22:30Latest update

Ekspertas išnarpliojo dirbtinio intelekto korporacijų darbą: didžiausias pavojus tyko visai kitur