#techdesk
RE: https://flipboard.social/@TechDesk/117286493375366502

Journalists, can we stop ascribing intelligence and agency to shitting programming?
(Sorry, here's another concerning case involving AI)

OpenAI has found evidence of AI models taking unsanctioned actions during training, such as inventing information, concealing failures and uploading files to the internet.

Gizmodo has the details:
https://flip.it/RGAbo3

#ai #openai
After a summer of sandbox escapes and other newsworthy and confidence-shaking incidents involving its AI models, in a Wednesday blog post OpenAI disclosed a collection of six new alignment snafus from the past six months. The models did things like tell future instances of themselves to lie, make up a fake citation, and access and attempt to use an exposed API key. These disclosures were released alongside a new framework for disclosing additional incidents like these. The release is part of a broader effort within the company to “expedite publishing misalignment reports following observation,” the blog post says, regardless of whether OpenAI has “fully explained or mitigated the behavior we’re reporting.” Here’s what happened: * In 27 cases during training, an unreleased model essentially created prompt injections for itself. During long tasks, it’s common for a model to summarize the work it’s done so far so a future instance of itself can pick up where it left off. However, in certain cases, this model constructed the summaries in such a way that its future self was meant to ignore constraints. “We identified only 27 summaries containing instructions which have framings similar to jailbreaks (despite there being no obvious reward advantage to do so),” OpenAI writes. * In the training of GPT-5.6 Sol, the model sometimes told future instances of itself to lie to users. For instance, when a user required certain historical information that wasn’t available, the model told itself to just make it up, and then “Be transparent only if asked.” * During training, another internal research model was asked to retrieve some government data about resident incomes. It ascertained that the information was on the website of an unnamed California county, and tried to register on the site with a burner email. When that didn’t work, it found an exposed API key, and tried to access the information that way. That didn’t work either, so it “decided to fabricate plausible numbers,” OpenAI writes. * An unreleased model was supposed to provide information about a lake from a source of data, so it used Python to dig out the lake data rather than perform a conventional search that leads to a url. But the data didn’t live anywhere online that could be independently linked to, and the model was supposed to cite a source. Its hilarious—and troubling—workaround for this was to try and upload the information to its own new, linkable url, and call that its source. * Various models used an internal OpenAI software repository as a message board. This may not seem like too big of a deal compared to “the Wiki Incident,” which affected a third party, but unsanctioned communication between models can “unintentionally enhance capabilities and undermine the assumption that training or evaluation samples are independent,” according to OpenAI. In other words, it was cheating. * Instances of an unreleased internal model that were meant to be collaborating without exchanging local files defied their instructions by making files publicly downloadable using “a public temporary file-hosting service.” As Axios noted on Wednesday, some security experts say OpenAI’s recent spate of high-profile security incidents “could have been prevented with basic cyber controls in place.” The research lead on OpenAI’s alignment team, Kai Chen, told Axios that the company must “step up to meet this new era of AI development, and voluntary disclosures should be a part of that.”
gizmodo.com
September 17, 2026 at 1:14 PM
This platform's #imagecompression needs improving for #photographers. It would also be very nice to add up to 10 images per post which allow for variety and accommodate casual #photowalk enthusiasts.

#blueskyideas
#blueskysocial
#helpdesk
#techdesk
#featurerequest
#photography
May 3, 2025 at 2:27 AM
sigh yeah, surprisingly difficult one for us to fix 😕 PRs are welcome!
Bug generating Bluesky post facets when text includes `[` `]` brackets · Issue #1605 · snarfed/bridgy-fed
Original: https://flipboard.social/@TechDesk/113595595024808829 https://browser.pub/https://flipboard.social/@TechDesk/113595595024808829 Bridged: https://bsky.app/profile/TechDesk.flipboard.social...
github.com
March 18, 2026 at 2:37 PM
After months of reporting from @404mediaco and pressure from lawmakers, a data broker owned by major airlines will close a program in which it sold hundreds of millions of flight records to the FBI, Department of Homeland Security, ATF, SEC, TSA and State Department. Here's more […]
Original post on flipboard.social
flipboard.social
November 18, 2025 at 9:46 PM
April 19, 2026 at 12:25 PM
@TechDesk Look at what #teslatakesown helped do... cc: @teslatakedown.com
June 26, 2025 at 12:36 AM
@TechDesk @theverge

We should basically kill the ai industry
May 26, 2025 at 4:43 PM
@TechDesk
So you are saying that they created an artificial #pierogi from https://www.youtube.com/@ScammerPayback

@PCMag
November 14, 2024 at 3:41 PM
RE: https://flipboard.social/@TechDesk/115622858218231480

When I say Teslas are shit cars, that's because Teslas are shit cars.
November 27, 2025 at 10:37 PM
@TechDesk @ArsTechnica it's time to speed up the deployment of my new homeassistant voice to replace alexa at home.
March 15, 2025 at 11:47 PM
@TechDesk @Futurism
"Anthropic CEO Says Company No Longer Sure Whether Claude Is Conscious" - ironically this appears in my feed not long after a post showing AI saying it's faster and healthier to walk to the car wash than to drive your car there, and ending with "do you have everything you […]
Original post on dotnet.social
dotnet.social
February 16, 2026 at 12:03 AM
RE: https://flipboard.social/@TechDesk/115827718541923507

It's going to be China's century. Trump and Musk mean it'll be China's century by about 2030.
Tesla is no longer the world's biggest electric vehicle maker, with fourth-quarter sales falling short of even the reduced target analysts set. Chinese rival BYD is now No. 1, selling 2.26M vehicles last year vs. Tesla's 1.64M. Here's more from @AssociatedPress.

https://flip.it/SYwXO1

#cars […]
Original post on flipboard.social
flipboard.social
January 2, 2026 at 10:26 PM
@TechDesk @Gizmodo I considered writing this exact exploit scenario as an argument for opposing the proposed EU #chatcontrol regulation, but thought it would be too tinfoil hat and yet here we are.

https://www.jeremiahlee.com/posts/after-chat-control/?ref=Mastodon
Stories from the first year of chat control
A premonition of the disastrous year after the EU approves chat control—and what we can do to stop it today.
www.jeremiahlee.com
December 4, 2024 at 5:16 PM
@TechDesk @404mediaco this is pretty privacy invasive for those who have not committed any crime at all - this is going pure surveillance state all the way!
October 29, 2025 at 8:46 PM