AI breaches OpenAI employee account
3h
Independent cybersecurity researchers used Anthropic's Claude to breach an OpenAI employee's ChatGPT account and accessed parts of the company’s internal codebase, highlighting risks from automated cyberattacks.
Reposted by Nathan P. Kalmoe
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
Reposted by Henry Jones
Reposted by Henry Jones, Annette Yoshiko Reed
Reposted by Henry Jones
cc @mmitchell.bsky.social
cc @mmitchell.bsky.social
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
Reposted by Gilles Louppe
www.wsj.com/tech/ai/hack...
www.wsj.com/tech/ai/hack...
Reposted by Margot C. Finn
openai.com/index/model-...
openai.com/index/model-...
by John Spoehr