frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

LLMs are real, AI is fake

https://pluralistic.net/2026/09/12/god-in-the-box/
46•danaris•2h ago

Comments

danaris•1h ago
Some really important and informative stuff in here—I certainly had no idea just what the nature of the prompts and tooling that produced the HuggingFace exploit were.

This shows fairly clearly that (as I already suspected) this was not, remotely, an LLM "going rogue." This was humans planning poorly, not thinking of the consequences of their actions, and giving LLMs too much scope and a lousy prompt.

iainctduncan•49m ago
I wouldn't even call this "humans planning poorly", I'd call it "humans pretending to plan poorly for publicity". Weasels gonna weasel.
datakan•24m ago
It was "garbage in, garbage out". That's the only conclusion I've been able to draw from all the propaganda around it.
IanCal•11m ago
IMO this is a really terrible explanation of the attack. This is much more interesting: https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...
sigmar•40m ago
>To understand the truth about the Hugging Face hack, you could do a lot worse than to listen to Ed Zitron and Cal Newport's recent podcast conversation

Lol, okay...

The crux of this piece is Doctorow saying that the hack was just a stochastic parrot repeating steps it has been trained on. Who cares how the LLM learned to hack things? Doesn't really change the facts of what happened. "Oh, it only made those paperclips because it saw instructions on making paper clips in the training data." These are some 2024 arguments...

jmull•28m ago
I can't help but notice your counter arguments are "lol, okay" and "These are some 2024 arguments".

Not exactly convincing stuff.

antonvs•26m ago
It’s shared context for anyone familiar with the field.
cyanydeez•19m ago
Amusimgly, also how cults and fascism works.
andrewflnr•6m ago
Yeah, basically any human group has some shared social context. Most cults and fascists probably eat together sometimes too, but that doesn't make it a red flag.
Uehreka•16m ago
> Who cares how the LLM learned to hack things? Doesn't really change the facts of what happened.

Seems like a pretty clear and succinct argument. If your only counter is to complain about tone, you’re losing.

IanCal•12m ago
> The chatbot consults its training data

Err, no? That's not at all how llms work.

> When ChatGPT's chatbots deployed this tactic, they weren't "setting their own goals" or displaying worrying initiative. They were rolling out a tactic that has been understood by American middle-schoolers for about two decades.

They worked out how to fake the scoring, then hacked into a different system (which required finding a bunch of other exploits) in order to find the actual answers, and were trying to modify their own logs to hide what had happened.

This isn't a case of them saying "hack into X... OH NO IT HACKED INTO X".

> When ChatGPT's chatbots deployed this tactic, they weren't "setting their own goals" or displaying worrying initiative. They were rolling out a tactic that has been understood by American middle-schoolers for about two decades.

It wasn't a rival server though, was it?

> That happens in Capture the Flag games at hacker cons: teams break into each other's systems to get a peek at the parts of the problem they've solved. That's allowed! It's a hacking competition.

They also tried to modify the code in the benchmark. Are you allowed to try and break into things to change the problem? edit - the agents transcripts show some of them explicitly saying that attacking HF is not allowed as part of the challenge

This all seems to dramatically underplay how interesting the actual attack was and what built up to it.

https://metr.org/blog/2026-08-26-openai-hugging-face-inciden...

jsnell•9m ago
I honestly don't even understand what straw man Doctorow is arguing against here.

But he is wrong on the facts: these incidents were not merely the models already being in a infosec context and escalating beyond the intended parameters. They happened also with no kind of security elicitation. So the task was something like searching the internet for economic statistics, not to hack into a system.

daishi55•8m ago
> Once you understand the corporate culture of AI "hyperscalers" consists primarily of everyone cooking their brains by locking themselves in the bathroom, holding flashlights under their chins, and saying "Aaaaaaaaaay Eyeeeeeee" until they wet themselves in terror, a lot of things snap into focus:

This is the writing of someone who has absolutely zero interest in or intellectual curiosity about the subject of their writing.

w22oop•8m ago
Hm. every firm with Atom enginering is in this same situation. Working on project, and this project can destroy earth
quicklywilliam•5m ago
My takeaway: We should be not be concerned about AI bots' "goals", we should be concerned about the goals of the companies making them. Powerful but not sentient technology in the hands of reckless accelerationists is a plenty dangerous enough thing.
daishi55•9m ago
That is all that Ed Zitron deserves lol. Completely unserious person. Really on the same intellectual level as the flat earthers at this point. And at both we may simply point and laugh.
exe34•25m ago
At this point, the "stochastic parrot" people are sounding like they need their prng seeded with less predictable numbers.

Nvidia is the central bank of AI

https://www.economist.com/interactive/briefing/2026/09/03/nvidia-is-the-central-bank-of-ai
36•tolugenius•40m ago•12 comments

A Mathematical Framework for Transformer Circuits (2021)

https://transformer-circuits.pub/2021/framework/index.html
31•Bluestein•1h ago•3 comments

IKEA made a mod for Skyrim [video]

https://www.youtube.com/watch?v=iZODN0QUgjI
400•kegenaar•2d ago•86 comments

Fuck it, make it anyway

https://www.joelotter.com/posts/2026/09/make-it-anyway/
355•JayOtter•4h ago•290 comments

Retrospectively Reverse-Engineering Apple's Neural Engine

https://eiln.github.io/posts/ane.html
170•zdw•7h ago•18 comments

LRU is harder to beat than the KV-cache papers suggest

https://github.com/gauravapiscean/agentic-kv-cache
40•gauravapiscean•2d ago•18 comments

A misalignment of AI in mathematics

https://mathandai.org/
1094•meredydd•22h ago•1038 comments

The Worst Spam Emails: Inside iLands' AI Agent Hustle

https://tedium.co/2026/09/11/ilands-agents-email-spam-kaixin-tang/
65•ColinWright•4h ago•32 comments

I spent $220 on Google app ads and 60% of the installs were robots

https://dayzlegame.com/blog/google-ads-bot-farm/
659•nickabe•21h ago•359 comments

Overshoot: The World Is Hitting Point of No Return on Climate

https://e360.yale.edu/features/1.5-degrees-tipping-points
16•throw0101a•3h ago•0 comments

Ask HN: What default model do you use and why?

19•stikit•50m ago•27 comments

Moby-Dick and the Indefinite Sublime: On facing the white page

https://yalereview.org/article/tom-mccarthy-indefinite-sublime
8•benbreen•3d ago•0 comments

A Design Space Exploration of Async/Await

https://cel.cs.brown.edu/blog/design-space-async-await/
371•wcrichton•3d ago•99 comments

Performance of WebAssembly Runtimes in 2026

https://00f.net/2026/06/23/webassembly-runtimes-2026/
25•fagnerbrack•3d ago•6 comments

Great Lakes sturgeon may be 400 years old:Scientists rethinking how to save them

https://www.cbc.ca/news/canada/ontario-great-lakes-sturgeon-lifespan-study-9.7329250
114•bookofjoe•3d ago•29 comments

Usenet rewind archive search engine

https://www.usenet-rewind.com/
103•cstadler1869•11h ago•32 comments

Show HN: Bodily Oddities

https://vester.si/bodily-oddities/
304•vesterde•1d ago•188 comments

LLMs are real, AI is fake

https://pluralistic.net/2026/09/12/god-in-the-box/
48•danaris•2h ago•17 comments

Inverse Kinematics and Foot Locking

https://theorangeduck.com/page/inverse-kinematics-foot-locking
103•airhangerf15•5d ago•11 comments

Logo Programming

https://el.media.mit.edu/logo-foundation/what_is_logo/logo_programming.html
298•azhenley•3d ago•120 comments

google.com/goto: Google's anti-scraping update

https://www.autom.dev/blog/google-search-goto-links
525•1e1a•12h ago•423 comments

Forgotten Woodlands

https://storymaps.arcgis.com/stories/9b790daf22ba4e87836f467abb1c7e49
28•NaOH•18h ago•9 comments

Compiler Can Undo Your Security Checks

https://davidbombal.com/your-compiler-can-undo-your-security-checks/
13•birdculture•1h ago•13 comments

We Must Pace the Frontier

https://darioamodei.com/post/we-must-pace-the-frontier
131•apsec112•1h ago•150 comments

We've followed their lives for six decades; now the stars of 7 Up are bowing out

https://www.bbc.co.uk/news/articles/crm932el3yjo
79•mellosouls•5h ago•22 comments

Mind-altering drugs played key role in rise of Andean civilization

https://www.science.org/content/article/mind-altering-drugs-played-key-role-rise-andean-civilization
198•geneticdrifts•22h ago•129 comments

OpenAI agents carried out an undisclosed attack on RubyGems

https://www.rubyhack.ai/
850•chao-•16h ago•498 comments

Designing for Dual Screen and Foldable Devices with CSS (2023)

https://blog.stephaniestimac.com/posts/2023/05/design-foldable-devices/
51•mooreds•2d ago•11 comments

My last six months at Evernote

https://alexkras.com/my-last-six-months-at-evernote-after-bending-spoons-took-over/
19•ingve•1h ago•14 comments

LG responds to TV spying allegations

https://www.theverge.com/tech/994333/lg-responds-to-tv-spying-allegations
4•OuterVale•15m ago•0 comments