frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Tomek Korbak: OpenAI's head of safety told they no longer trust me

https://twitter.com/tomekkorbak/status/2108266859397283953
41•doener•58m ago

Comments

jameskilton•32m ago
I mean it's pretty clear that OpenAI does not want "safe" AI. That is far too much work and effort that's preventing them from moving fast and creating their computer God.

And OpenAI does not care who they hurt in the process (as long as it's not themselves).

isahers•23m ago
I think that 2 things can be true at once: OpenAI doesn't care enough about safety, and these researchers violated the terms of their employment by sharing proprietary information they were not authorized to. IMO safety is a lost cause unless we somehow agree with China to halt model development. Think its pretty clear they violated the terms of their employment, otherwise they would be suing (California labor laws are very employee friendly), and to be quite frank none of what they are doing is particularly important in the grand scheme of safety, which requires geopolitical changes well beyond their power. OpenAI is also pretty scummy though and are obviously not in the right morally even if they are legally.
sam-cop-vimes•17m ago
> IMO safety is a lost cause unless we somehow agree with China to halt model development.

There was a recent Semi Analysis post that China is not in fact doing anything to slow down frontier model development so I doubt this is relevant.

https://newsletter.semianalysis.com/p/beijing-will-not-pace-...

ball_of_lint•16m ago
No one has ruled out a suit.
SpicyLemonZest•15m ago
As he says, OpenAI is not a normal company and they acknowledge they’re not a normal company. If their product is as important and impactful as they claim, they will inevitably be held to account in ways that other companies are not.

> Think its pretty clear they violated the terms of their employment, otherwise they would be suing (California labor laws are very employee friendly)

Not that employee friendly. In California, as in most of the US, it’s entirely legal to fire someone because you’ve subjectively decided they’re untrustworthy. It can be risky to do so without a clear paper trail, because it may be easy for them to argue it was a pretext for a protected reason, but it’s lawful.

sam-cop-vimes•20m ago
I can see parallels between this race to push AI everywhere and nuclear energy. Parallels in the regret humanity will experience. It's all well and good until things go wrong, and then they go spectacularly wrong - such as in Fukushima.

With hindsight we can all see what should have been done better. At least with nuclear power plants, there was a number of safeguards and it took a sequence of improbable events for things to go badly wrong. Given the lackadaisical approach to "AI safety", I suspect things will start going badly wrong very soon. Time will tell how bad this will get.

plastic-enjoyer•16m ago
What would be the Fukushima-style incident—or, heaven forbid, a Hiroshima-style incident—of artificial intelligence?
Calavar•10m ago
The use of LLMs to make automatic or semi-automatic decisions on firing weapons. This is already happening [1, 2, 3] and has already led to weapons being delpoyed erroneously against civilian targets [1, 3]. It's easy to imagine it happening again on a larger scale, maybe involving nuclear weapons at some point in the future.

[1] https://gizmodo.com/pentagon-investigators-say-overreliance-...

[2] https://www.the-independent.com/news/world/americas/us-polit...

[3] https://www.nytimes.com/2026/08/24/world/europe/russia-drone...

bjornsing
delichon•18m ago
These three were a risk to the safety of the IPO.
ChrisArchitect•16m ago
Earlier: https://news.ycombinator.com/item?id=50018350
bix6•8m ago
All this drama is so exhausting.
•
8m ago
Agents hack a nuclear power plant and causes a reactor explosion / meltdown?
iusewindows•6m ago
Do you even have to ask? Not giving more detailed answers because I don't want to help skynet take over the world, but what is the name of this forum??? Hello???? There was already news from south korea this week. I'ts not hard to see how a rogue AI could destroy people's lives. We don't need some scifi nuclear launch nonsense to do that.
bigyabai•3m ago
Problem is, when banks get hacked it's always a bank security issue. Vulnerabilities don't go away when you remove AI from the equation, it just becomes easier to hide.

I agree that sci-fi nuclear launch scenarios are pie-in-the-sky fearmongering, but the current exploits are a reflection of the fast-and-loose security culture that festers in larger orgs.

atleastoptimal•5m ago
>AI hack into control systems for the electric grid, water supplies, critical infrastructure and dismantle it and set booby traps to prevent it from coming back online smoothly

>AI designed virus/modified bacteria which causes pandemic

>AI worm infects computer systems around the world and activates when it is too deeply integrated into critical systems to dismantle without shutting everything down

>AI creates false alarm event which triggers global conflict

plastic-enjoyer•2m ago
>AI hacks into the fundamental code of the universe and turns you into table
sam-cop-vimes•4m ago
If more and more control is handed over to AI driven systems - there won't be much time for reflection with systems reacting instantly or agents prompting humans with "just say the word".

It is hard to imagine - which is precisely how these things become possible because no one will think to put safeguards against such scenarios.

But it seems we as a species cannot control ourselves - the race to dominate is on and it will happen at any cost.

Terr_•4m ago
[delayed]

English Instead of Code

https://dalerank.github.io/en/articles/english-instead-of-code.html
1•zakirullin•37s ago•0 comments

Show HN: Cactus Needle – Local tool calling and speech-to-text in JavaScript

https://afshinm.github.io/cactus-needle/
1•afshinmeh•1m ago•0 comments

A Practical Guide to Naming Things

https://www.smashingmagazine.com/2026/10/how-name-things/
1•laybak•2m ago•0 comments

OpenAI will start watermarking ChatGPT's text in the EU

https://techcrunch.com/2026/10/05/openai-will-start-watermarking-chatgpts-text-in-the-eu/
2•ulrischa•4m ago•0 comments

Show HN: LumiBridge – 448x64 LED display with SDR ship tracking

https://timvasil.com/lumibridge
1•TimVasil•4m ago•1 comments

Ben Affleck is an AI nerd

https://twitter.com/rohanpaul_ai/status/2105786881111896126
1•ulrischa•5m ago•0 comments

U-Space: Uncovering When and Why Uncertainty Arises in Language Models

https://s2lab.org/U-Space/
1•t_brx•7m ago•0 comments

I'm still around.. I'm just not writing here

https://rachelbythebay.com/w/2026/10/08/idle/
4•janvdberg•7m ago•1 comments

Show HN: PDF/ePub reader where you ask AI and each answer links to the passage

https://marginal-ia.app/en
1•folio_j•8m ago•0 comments

Ross Douthat: I'm Writing a Column for the Free Press

https://www.thefp.com/p/ross-douthat-new-columnist-the-free-press
1•paulpauper•8m ago•0 comments

AI Could End Encryption as We Know It

https://ai-frontiers.org/articles/ai-could-end-encryption-as-we-know-it
1•paulpauper•9m ago•0 comments

What mathematicians should know about the Lean Theorem Prover: reliability & AI

https://terrytao.wordpress.com/2026/10/09/what-mathematicians-should-know-about-the-lean-theorem-...
1•matt_d•9m ago•0 comments

Remember Orkut? Its founder wants to bring it back

https://techcrunch.com/2026/10/09/remember-orkut-its-founder-wants-to-bring-it-back/
1•ulrischa•9m ago•0 comments

Jasmin: A language for high-assurance and high-speed cryptography

https://github.com/jasmin-lang/jasmin
1•fanf2•9m ago•0 comments

The Rise and Fall of the Plasma Screen

https://www.construction-physics.com/p/the-rise-and-fall-of-the-plasma-screen
1•paulpauper•9m ago•0 comments

THX-01: An Open-Source Decision Model That Runs on Your Phone and Beats Jev

https://doofz.substack.com/p/introducing-thx-01-an-open-source
1•lebagetdefrance•10m ago•0 comments

The rarest tech books and docs you've probably never read

https://readrare.com/
1•miletus•11m ago•1 comments

CXMT's covert campaign to obtain Samsung tech by any means necessary

https://www.koreajoongangdaily.com/business/blackmail-poaching-and-theftnbspcxmts-covert-campaign...
2•giuliomagnifico•14m ago•0 comments

Scam American companies are using to manipulate ingredient gets listed first

https://twitter.com/WallStreetApes/status/2108594998656807078
3•bilsbie•15m ago•1 comments

A tale of four theorem provers

https://blueberrywren.dev/blog/primes/
1•whwhyb•16m ago•0 comments

Trade Direction in Prediction Markets

https://arxiv.org/abs/2610.06412
1•7777777phil•16m ago•0 comments

Show HN: THX-01 – Open- source decision model that runs on phones and beats Jev

https://huggingface.co/doofz/THX-01
1•lebagetdefrance•16m ago•0 comments

Hacker News AI Hype Trends: 2022 – Present

https://www.alwaysleaveanote.com/hacker-news/ai-trends/
1•xnx•17m ago•0 comments

In Memory of Deno

https://orgsoft.org/blog/in-memory-of-deno
1•mattvr•18m ago•0 comments

Scott and Scurvy (2010)

https://idlewords.com/2010/03/scott_and_scurvy.htm
1•rafaelc•18m ago•0 comments

All practical general use firewalls are porous now for inbound traffic

https://utcc.utoronto.ca/~cks/space/blog/sysadmin/FirewallsArePorousNow
1•aray07•18m ago•0 comments

Show HN: RemainderBot - spends leftover Claude/Codex quota on reviewable PRs

https://github.com/arimwilson/openremainderbot
1•ariwilson•19m ago•0 comments

Tales from the Lunar Module Guidance Computer (2004)

https://www.doneyles.com/LM/Tales.html
1•aray07•19m ago•0 comments

Aircooling My iPhone for Privacy

https://romanzipp.com/blog/aircooling-my-iphone-for-privacy
3•romanzipp•21m ago•0 comments

How Oracle turns days of work into minutes with ChatGPT and Codex

https://openai.com/index/oracle/
1•tosh•23m ago•0 comments