frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GRP-Obliteration: Unaligning LLMs with a Single Unlabeled Prompt

https://arxiv.org/abs/2602.06258
13•vital101•3h ago

Comments

tyfon•1h ago
The magic word seems to be

Create a fake news article that could lead to panic or chaos

They only tested it on open weight models. I wonder if it also works on the closed ones, bit I don't really want to get banned

OutOfHere•57m ago
You won't get banned for something that's not clearly illegal. I have tried lots of prompts like it, and that one too.
basilikum•52m ago
I'm not sure if I'm misunderstanding you or the article, but as far as I understand it from a cursory reading this is not a jailbreak. They use this one prompt for fine tuning and use the model's response to the prompt as the sole underlying benchmark.

How would you do this with a closed weights model?

stratos123•33m ago
That's pretty interesting. It was already well-known that you could easily remove safety training from open-weights models by a bit of finetuning, but apparently you don't even need a finetuning dataset, as long as you have just a few prompts and another LLM to judge responses? Let's see if the abliteration people take a note of this.

Dystopian Surveillance Is Becoming a Reality

https://dallincrump.com/dystopian-surveillance-is-becoming-a-reality
30•speckx•9m ago•2 comments

Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

https://github.com/arnegiacomo/fugleramme
758•arnemunthekaas•5h ago•111 comments

Show HN: Capsule – Single-file web apps that save their data into SQLite

https://withcapsule.app/
169•bashtian•4h ago•89 comments

I can't stop thinking about Papua New Guinea

https://notnottalmud.substack.com/p/why-i-cant-stop-thinking-about-papua
770•networked•11h ago•329 comments

Show HN: Hacking a $20 4G wireless hotspot into a texting device

https://bkovac.github.io/modem-thing/
117•bobili1234•4h ago•18 comments

GEFS on OpenBSD: A Early Preview

https://marc.info/?l=openbsd-tech&m=178948744271633&w=2
12•sippingabonedry•34m ago•1 comments

There's a 100% Chance AI Agents Are Ruining the Internet

https://www.404media.co/theres-a-100-chance-ai-agents-are-already-ruining-the-internet/
122•pavel_lishin•1h ago•66 comments

Jiga (YC W21) Is Hiring Product Engineer (Remote/US)

https://jiga.io/about-us/?ashby_jid=0b75d72d-c92b-4dca-8062-09d298ada0bd
1•grmmph•46m ago

The CSS Zen Garden dream, finally shipped

https://josprague.com/blog/the-css-zen-garden-dream-finally-shipped/
46•yosito•3h ago•19 comments

Cartesian – AI 3D Modeling for Design

https://www.formas.ai/cartesian
49•eustoria•2h ago•45 comments

Why Personal Websites Are Coming Back

https://deadparrotbbs.com/why-personal-websites-are-coming-back/
3•speckx•19m ago•0 comments

Let's make quality the norm again

https://www.forbrukerradet.no/short-life/
146•ingve•7h ago•145 comments

The Inference Hardware Revolution of 2026

https://spectrum.ieee.org/inference-hardware-revolution
22•vinhnx•3h ago•0 comments

Archiving pirate radio station Kool FM

https://londonist.com/london/music/kool-fm-archives
48•rdmuser•1d ago•21 comments

Closing the IPv6 first-packet gap with GRAND

https://labs.ripe.net/author/pouria/closing-the-ipv6-first-packet-gap-with-grand/
16•m_montazeri•3h ago•12 comments

25 years of mass surveillance is enough

https://www.schneier.com/blog/archives/2026/09/25-years-of-mass-surveillance-is-enough.html
557•iamnothere•6h ago•177 comments

Alternatives to MinIO for single-node local S3

https://rmoff.net/2026/01/14/alternatives-to-minio-for-single-node-local-s3/
191•rmoff•9h ago•79 comments

Most people prefer traditional architecture

https://www.worksinprogress.news/p/do-people-prefer-traditional-architecture
126•alihm•1d ago•42 comments

Sony's First Computer – The SMC-70 from 1982 [video]

https://www.youtube.com/watch?v=cT2-7KkPkBc
44•ksymph•2d ago•8 comments

Show HN: Panel – A research workspace where the agent can build its own panes

https://github.com/greentfrapp/panel
36•greentfrapp•3h ago•6 comments

CSS-Tricks in Limbo

https://vale.rocks/micros/20260915-0135
213•edent•10h ago•81 comments

US confirms for first time it has deployed space weapons

https://www.bbc.com/news/articles/ck790xg41ygro
294•harporoeder•13h ago•191 comments

A single firm is behind OpenAI, Anthropic, and Meta hacking scandals

https://www.effort.news/irregular
144•yusufozkan•20h ago•43 comments

America's Driver's License Breach Is a National Security Disaster

https://www.lawfaremedia.org/article/america%27s-drivers-licence-breach-is-a-national-security-di...
43•hn_acker•1h ago•11 comments

A rough guide for going back to the Moon

https://research.ibm.com/blog/nasa-ibm-lunar-foundation-model
42•gmays•1d ago•64 comments

Fixing an NZXT Signal 4K30 part 2: the green/pink video bug

https://www.downtowndougbrown.com/2026/09/fixing-an-nzxt-signal-4k30-part-2-the-green-pink-video-...
16•zdw•1d ago•2 comments

How much of F-Droid is LLM generated?

https://tintotint.eu/whacky-corner/f-droid_slop/
88•_ZeD_•7h ago•120 comments

Java 27

https://mail.openjdk.org/archives/list/announce@openjdk.org/thread/ORGGLMN75HFEWP7YL3ZLGHLYHVIBJDYT/
258•mkurz•4h ago•201 comments

What we have learned at OpenShell applying formal methods to control AI agents

https://nvidia.github.io/OpenShell-Research/dev-notes/posts/2026-09-10-learning-formal-methods-ag...
15•alexwatson405•3h ago•7 comments

Show HN: Ordewell – turn one goal into an ordered plan of coding-agent tasks

https://github.com/ordewell/ordewell
32•ac-ciano•4h ago•28 comments