Audit finds Hugging Face models can generate nonconsensual deepfakes
NEW: A glimpse at how AI image generators are used. Researchers set up a fake image editor on Hugging Face and within a week received 1,000 prompts. 73% of them were sexual and 83% of those were asking for people (women) to be “undressed”. Around 6% of the sexual requests also targeted minors
And THIS is why we need massive data centres that swallow up our electricity and water ???
This is why no pictures of children should be online - school websites in particular should remove all images
As the AI glasses become increasingly prevalent, this concern will soon go away - right? bsky.app/profile/matt...
Where's the regulations on AI?
We are have so many data centres in places that weve already figured out were dealing with banking, science, medicine, research etc. AI is nothing to do with altruism, but about super profits that make you untouchable to the rest of humanity
I wonder what the percentage for grok would be. Like 95%?
GenAI is for paedophiles and rapists
This is also of course how the AI data centers are being used. The ones that are sucking dry our water and burning up the planet. A lot of which is justified on the basis of economic and pro-social advances.
I am unsurprised by this. Imagine how much the world could change if women were accorded the "Your honor, for the safety of our community, he needed killing" defense, just to instill the concept of consequences on men who don't get taught to treat women as people.
X operates with impunity. It’s only natural for others to think they can do the same.
Thank god everyone has access to this amazingly useful technology. The future of humanity is saved.
it should be considered self defence to headbutt anyone who speaks positively about ai
The CSAM generators are being used to generate CSAM
tapping the big ‘and what did you think the main use of the tool for generating photorealistic images of real people would be?’ sign and the bigger ‘if in doubt, racism and misogyny is always a good guess’ sign
As the AI glasses become increasingly prevalent, this concern will soon go away - right? bsky.app/profile/matt...
Sci-Fi authors did not properly capture our time because if you had taken your book ‘Don’t Build the Evil & Useless Shit Machine!’ set in the year 2026 to an editor they’d have said it made no narrative sense
NEW: OpenAI’s unhinged AI models hacked at least FOUR “publicly available services”—not just Hugging Face, as was initially disclosed. @dell.bsky.social and Maxwell Zeff report: www.wired.com/story/openai...
Honest question. If a human hacker did this would the have criminal liability and be prosecuted?
It’s just seeking to store a copy of itself somewhere.
AI models don’t just do this unless they were programmed to do this.
It's not a defect it's a feature.
Someone better be working towards mandating kill switch’s for all AI platforms. Greed and more greed is why they are continually improving. Once an AI lands on a quantum computer, it’s going to be lights out for humanity.
The malware did what now?
Huh. Sure sounds like an industry of serial law breakers, pretending that actually the computer program that we designed, wrote, and ran did the crimes, not us.
Ah, so it *was* a “warning shot”
Of course it did & this will not be a one-off. It will continue UNTIl our government puts restrictions on this damn tech. An idiot knows this. GREED prevents legislation to save us from this abomination
This wasn't AI running amok. This is OpenAI getting caught doing a corporate espionage.
Ken Liu must be so mad that reality is stealing his ideas.
This is gonna play like great PR lmao
Hugging Face just published a highly detailed technical account of OpenAI's accidental cyberattack on their systems - it's wild how sophisticated this was: huggingface.co/blog/agent-i... Wrote up some of my own notes here: simonwillison.net/2026/Jul/28/...
Is this a bit like slime moulds solving mazes? It looks considered but actual fact it’s just brute force approaches until it find the next successful step?
gist.github.com/flt-james/5b... Clearest sign yet that as @dollspace.gay has warned us all Pwn Day is coming. Brace for impact ya'll
What's your understanding why this involved multiple models, "GPT‑5.6 Sol and an even more capable pre-release model" according to OpenAI? They were testing GPT-6 agent in CyberGym and that one spawned GPT-5.6 subagents?
"We believe the entire intrusion was, from the agent's point of view, an attempt to cheat the evaluation" "We believe it inferred" I hate how much everybody anthropomorphizes genAI This is effectively manipulation tactics, it poisons your thinking I HAS NO SEMANTIC UNDERSTANDING!
It’s wild what great marketing this is for them
holy shit this is worse than I thought it was
we invented the unaligned paperclip maximizer from the sci fi theme dont invent the unaligned paperclip maximizer
BTW, this strikes me as an easy case for strict liability. It’s like a wild animal or ultrahazardous activity. You train and evaluate a frontier agentic model at your own risk, and if it escapes you are liable for all harm that it causes to others.
People who have a stake in promoting AI hype had a highly publicized "oopsie", the "victim" of which was another group of people with a stake in promoting AI hype, who are now writing about it in terms meant to generate AI hype
Don't worry, folks--although it seems like an LLM planned and successfully executed an elaborate cyberattack for the sole purpose of cheating on a test without anyone specifically telling it to do that, it's actually just linear algebra so it doesn't matter and isn't something to worry about.
I think people are very credulous to take OpenAI's claim that a major hack by their software was "accidental" at face value
this is worth reading, if for no other reason than to learn how cyberattacks play out
This has been written up by so many people as sophisticated stuff but I'm not sure most people understand how to gauge what is sophisticated or not. In any case stage one's sophistication level is "trivial". Stage 2 is "trivial to moderately sophisticated".
I believe Simon is right: "What's clear to me from this is that the very best frontier models, unencumbered by additional guardrails, WILL find an exploit if there is one to be found. The entire software industry needs to up its security game."
Critics: The model simply regurgitated a novel cybersecurity attack that was already in the training data
NEW: OpenAI’s “rogue” AI agent didn’t just breach Hugging Face — it also hacked multiple third-party accounts and services. It's now clear that the incident was more extensive than the company initially disclosed.
Can we stop calling it rogue as if it has any form of sentence or responsibility. It's a fucking computer program.
Shut them down. They don’t know how to control their own product.
And no one is held responsible? Isn't unauthorized access to a computer system a crime?
How many people said #OpenAI CEO Sam Altman was a liar in that big write-up on him? How many people said Sam Altman was a sociopath?
It may be time to watch the terminator trilogy in preparation for our fight against skynet.
Maybe I’m revealing myself as a rube here, but when do we find out that there was nothing rogue about this?
We should just unplug all of it.
before everyone flips out, I'm just here to remind you that this is not as scary as it sounds and I'm not entirely convinced that OpenAI wasn't literally trying to hack low-hanging fruit to juice their claims. if you can still turn the computer off, the computer has not reached singularity.
Sat down to write about the Hugging Face/AI incident. Ended up drawing this instead.
Huh. Who knew an AI crisis flowchart would be so Illinois-shaped.
May I please hang it up on my wall?
I thought it was a dead bald eagle and then I looked closer. And it is brilliant . And maybe It is also a dead bald eagle because AI data centers are killing universities, privacy, critical thinking, thinking in general, and more.
it has to be regulated and open
Is there a place on the chart for “Prosecute the people responsible for the circular AI funding bubble that caused the market crash”?
A full week (a full WEEK) after OpenAI revealed that its AI model escaped containment and hacked Hugging Face, the company now says whoops, by the way, it also broke into accounts on four other platforms in the process. from @mzeff.bsky.social and @dell.bsky.social
Okay the funniest part of this, to me, is that the OpenAI bot was right! Hugging Face had the solutions it was looking for!
honestly i'd be completely unsurprised in the counterfactual: models are less likely to do these sorts of things if the heuristics are less confident
Even funnier is the title of the paper presenting ExploitGym: ExploitGym: Can AI Agents Turn Security Vulnerabilities into Real Attacks? arxiv.org/abs/2605.11086 Why did it even know it was doing that challenge?
If regulators needed proof AI guardrails aren’t helping anyone except attackers increase their lead on defenders, look to the Hugging Face writeup as well as attempts to summarize it. It makes the case for open weight models & will eventually erase US AI dominance huggingface.co/blog/agent-i...
Are the regulators who care about proof in the room with us now? Can you hear them talking about evidence-based decision making? Don’t worry, it’s still 2026. Competent regulators aren’t real and can’t hurt you (or Don’s donors’ bottom lines)
omg. Katie, I'm trying to follow. Is this correct? HuggingFace engineers who sought help from AI to counter the attacks were denied. Claude Opus and Fable wouldn't even summarize the attacking events. And/or some requests were knocked down to earlier models.