frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Malicious Use of Artificial Intelligence

https://arxiv.org/abs/1802.07228
11•rasengan0•1h ago

Comments

voidhorse•9m ago
Should probably have a (2018) or (2024) (latest revision) on the title, especially given the current buzz surrounding AI and security/existential threats.
EGreg•8m ago
This paper diagnosed the disease in 2018. Eight years later none of the recommendations happened. Norms, collaboration, responsible disclosure - none of it materialized in any structural way.

I think the reason is that the paper frames malicious AI use as a policy problem, recommending social solutions. It's actually an architecture problem.

Every generation of computing has hit a version of this. Programs could write anywhere in memory - we added protected memory. Programs could hog the CPU - we added preemptive multitasking. Desktop apps could call any OS function - the iPhone sandboxed them. Nobody asked programs to please behave, the containment actually went into the infrastructure.

AI skipped that step entirely. We went straight to open-ended agents with broad permissions and tried to make them safe through alignment and prompting. I've been researching this for the past year and I think alignment is necessary but not sufficient, because the intelligence increasingly isn't in the model. It's in the substrate - the harness, the domain knowledge, the tooling around the model. I actually measured this on real coding tasks: Sonnet with a code-derived index outperformed the frontier model (Opus 5.8) exploring on its own, and the top-tier model (Fable) refused the real work entirely! The cheap model with the right rig beat the expensive model without one. https://safebots.ai/matchup.html

If that's true then aligning the model doesn't solve the problem. A bad actor who can't get the best model just uses Sonnet. Or Llama. Or Kimi. The weights have already leaked and bits don't degrade - you can't recall them the way you can stop manufacturing CFCs.

So what do you actually do? Same thing that worked for CFCs. You gotta first build the safe version — in this case, declarative workflows running in sealed compute environments — and prove it handles 99% of actual use cases at lower cost. Let it win commercially. Then regulate the dangerous version. DuPont developed HFC refrigerants first. The Montreal Protocol became possible BECAUSE of that. The ban became politically viable because the alternative already existed.

I've been building this alternative for the past 8 months: https://safebots.ai/about

esafak•3m ago
Alignment is the architectural solution. Make it so the model can't misbehave. Yet many people here deride it as tainting the model. "Who's values is the aligned to?" people say. Sandboxes are a last ditch layer. They fail, as we see.

Fable 5.1 Solves the Cyphral Distich, a 370-year-old cipher

https://www.vals.ai/blogs/fable-solves-cyphral-distich
533•u1hcw9nx•6h ago•241 comments

Why is Google still serving dodgy ads?

https://www.atomic14.com/2026/09/13/why-is-google-still-serving-dodgy-ads
639•iamflimflam1•9h ago•299 comments

Spaceships (Reverse Asteroid)

https://spaceships.treybastian.com/
46•zdw•3d ago•15 comments

Registration without a phone number on Signal will use zero-knowledge proofs

https://community.signalusers.org/t/registration-without-a-phone-number/2222?page=10
110•Cider9986•5h ago•51 comments

Open-Source AI and Open Models Reading List

https://www.interconnects.ai/p/open-source-ai-reading-list
31•simonpure•2h ago•2 comments

Julia 1.13 highlights

https://julialang.org/blog/2026/09/julia-1.13-highlights/
169•eigenspace•3d ago•12 comments

Astra and Fable still hack on simple variants of alignment evals from 2025

https://www.lesswrong.com/posts/munJKF7iWMsWJLAH2/astra-and-fable-still-hack-on-simple-variants-o...
390•Levitating•12h ago•178 comments

Show HN: Is It Greg?

https://github.com/antoineleclair/is-it-greg
27•antoineleclair•1h ago•9 comments

Data collected by cars and sold to third parties

https://www.theverge.com/column/994172/your-car-is-selling-your-data
325•bookofjoe•13h ago•168 comments

The Malicious Use of Artificial Intelligence

https://arxiv.org/abs/1802.07228
11•rasengan0•1h ago•4 comments

Ask HN: What are you working on? (September 2026)

99•david927•9h ago•219 comments

Mullenweg has returned as CEO after attempted board ouster

https://techcrunch.com/2026/09/12/automattic-confirms-mullenweg-has-returned-as-ceo-after-attempt...
59•ilamont•6h ago•114 comments

Flawed routers flood University of Wisconsin internet time server (2003)

https://pages.cs.wisc.edu/~plonka/netgear-sntp/
67•walrus01•6h ago•8 comments

AI is not a normal technology

https://12gramsofcarbon.com/p/ai-is-not-a-normal-technology
15•theahura•2h ago•25 comments

The GDR and Vietnam: From Fake Coffee to Coffee Empire

https://www.katjahoyer.uk/p/the-gdr-and-vietnam-from-fake-coffee
34•NaOH•3d ago•7 comments

Writing a better reality: The case for optimistic sci-fi

https://honisoit.com/2022/03/writing-a-better-reality-the-case-for-optimistic-sci-fi/
10•andsoitis•2h ago•5 comments

Apple's Dimensional Drawings

https://developer.apple.com/accessories/dimensional-drawings/
39•herbertl•2h ago•5 comments

Reverse engineering my e-scooter and rewriting the firmware in Rust

https://bensimms.moe/reverse-engineering-scooter/
395•vinhnx•3d ago•91 comments

The Coming War on General Computation (2011)

https://en.wikisource.org/wiki/The_Coming_War_on_General_Computation
55•gregsadetsky•3h ago•14 comments

Rope, twine and thread: Invisible technologies of the Stone Age

https://knowablemagazine.org/content/article/society/2026/prehistory-lost-threads
4•knowablemag•2d ago•0 comments

Making Startups Powerful

https://paulgraham.com/powerful.html
172•tosh•12h ago•76 comments

Why is the x86 undefined instruction called ud2? Why 2?

https://devblogs.microsoft.com/oldnewthing/20260910-00/?p=112689
216•ibobev•14h ago•50 comments

I build a mechanical watch face: a real gear train for a watch with no gears

https://myday24.com/blog/how-a-mechanical-watch-face-is-built/
16•tomasslavicek•3d ago•7 comments

CUDA for AMD on Windows

https://github.com/Speedstu/CUDA-for-AMD-Windows
142•chiassedu80•12h ago•73 comments

AI Robots – When will they be in our homes

https://spectrum.ieee.org/ai-robots
5•andsoitis•2h ago•2 comments

"Chilling" warning or overreaction? AI bioweapons report divides experts

https://www.science.org/content/article/chilling-warning-or-overreaction-ai-bioweapons-report-div...
11•sbulaev•3h ago•0 comments

Why is privacy so hard?

https://cacm.acm.org/blogcacm/why-is-privacy-so-hard/
33•andsoitis•5h ago•37 comments

Reverse-Engineering Claude Web's MicroVM: Uncovering Anthropic's Hidden Antspace

https://aprilnea.me/en/blog/reverse-engineering-claude-code-antspace
63•rzk•2d ago•14 comments

Show HN: Exploring the intersection of prediction markets and social media

https://www.thevidmarket.com/
4•ryanpere2298•1h ago•0 comments

The contagion of fear

https://bcantrill.dtrace.org/2026/09/13/the-contagion-of-fear/
160•elffjs•4h ago•125 comments