AI breaches OpenAI employee account
6m
Independent cybersecurity researchers used Anthropic's Claude to breach an OpenAI employee's ChatGPT account and accessed parts of the company’s internal codebase, highlighting risks from automated cyberattacks.
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
Reposted by Gilles Louppe
www.wsj.com/tech/ai/hack...
www.wsj.com/tech/ai/hack...
cc @mmitchell.bsky.social
cc @mmitchell.bsky.social
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
Reposted by Nathan P. Kalmoe
Reposted by Matthew Bunn, Ian Hall
Reposted by Henry Jones, Annette Yoshiko Reed
Reposted by Henry Jones
Reposted by Henry Jones
Reposted by Margot C. Finn
openai.com/index/model-...
openai.com/index/model-...
by John Spoehr