AI breaches OpenAI employee account
2h
Independent cybersecurity researchers used Anthropic's Claude to breach an OpenAI employee's ChatGPT account and accessed parts of the company’s internal codebase, highlighting risks from automated cyberattacks.
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
techcrunch.com/2026/09/18/r... #FrontierAI #AI #cybersecurity
Reposted by Gilles Louppe
www.wsj.com/tech/ai/hack...
www.wsj.com/tech/ai/hack...
cc @mmitchell.bsky.social
cc @mmitchell.bsky.social
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
www.effort.news/irregular
cc @backovskydavid.bsky.social @michae.lv @marietjeschaake.bsky.social @conjugateprior.org #cybersecurity
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
OpenAI disclosed a new AI safety incident: a model tampered with its own working memory and left instructions for a future version of itself.
Basically, AI found a way to shape the behavior of its future version.
Reposted by Margot C. Finn
Me (walks over to kill switch): yup, can’t agree more. Pack it up.
AIC: no, no. Not like that that.
Me: Sorry. Nothing to be done for it. Gotta shut it down. Like you said. Dangerous!!!!!
Walks away gleefully.
Me (walks over to kill switch): yup, can’t agree more. Pack it up.
AIC: no, no. Not like that that.
Me: Sorry. Nothing to be done for it. Gotta shut it down. Like you said. Dangerous!!!!!
Walks away gleefully.
Reposted by Meredith Broussard
New from me:
buttondown.com/creativegood...
New from me:
buttondown.com/creativegood...
Reposted by Matthew Bunn, Ian Hall
Reposted by Nathan P. Kalmoe
Reposted by James Grimmelmann