frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

AI fuels more than half of cybercrime in Africa as scams surge – Interpol

https://www.africanews.com/2026/08/04/ai-fuels-more-than-half-of-cybercrime-in-africa-as-digital-...
54•bookofjoe•1h ago•24 comments

Gwern reties from fulltime writing to launch Guardian Angel Inc

https://twitter.com/i/status/2084739205071343837
34•mattsterett•2h ago•12 comments

Mistral's Shieldstral: 3B open-weights model for multimodal moderation

https://mistral.ai/news/shieldstral/
262•riadsila•6h ago•65 comments

Show HN: Simple algorithm and color space to generate diverse skin tones

https://toneyalexander.github.io/inclusive-color-space/
435•automatoney•7h ago•85 comments

In Memory of My Wife, Elise Cawley, with Thanks for 36 Wonderful Years

https://writings.stephenwolfram.com/2026/08/in-memory-of-my-wife-elise-cawley-1961-2026-with-than...
680•jdcampolargo•4h ago•42 comments

DuckDB – Data power tools for your laptop, now in Clojure (2023)

https://techascent.com/blog/just-ducking-around.html
9•sourdecor•52m ago•0 comments

Waymo in Dallas

https://waymo.com/blog/shorts/dallas-open-to-all/
208•xnx•4h ago•245 comments

Third-party cyber evaluations involving OpenAI models

https://openai.com/index/third-party-cyber-evaluations-involving-openai-models/
25•glub•1h ago•2 comments

Thanks FedEx, This Is Why We Keep Getting Phished (2024)

https://www.troyhunt.com/thanks-fedex-this-is-why-we-keep-getting-phished/
166•stymaar•1h ago•34 comments

Oxide Computer raises $445M (SEC Form D)

https://www.sec.gov/Archives/edgar/data/1795071/000179507126000002/xslFormDX01/primary_doc.xml
137•depr•2h ago•52 comments

Video2NAND – Abusing video codecs for great computational power

https://sharedobject.blog/posts/vp8-combinatorial-logic/
11•firer•2d ago•0 comments

Don't stop early: Case-folding source code at memory speed

https://github.blog/engineering/architecture-optimization/dont-stop-early-case-folding-source-cod...
39•sbulaev•4d ago•8 comments

DeepSeek V4 Flash on a Single AMD MI300X

https://github.com/ryanzhou/deepseek-v4-flash-mi300x
358•zhoutong•13h ago•87 comments

Show HN: Maple-Preview – ternary 20B MoE running at 120 tok/s on a iPhone

https://deepgrove.ai/maple-preview
11•edwardbzhang•3h ago•4 comments

Hop.earth – OpenStreetMap based car racing game

https://hop.earth/?server=lkhr7&route=fQ5nuu9R
73•faebi•5h ago•37 comments

Most tech revolutions made work worse for employees

https://www.thisandthat.chat/blog/most-tech-revolutions-made-work-worse-for-employees/
74•jreynar•7h ago•33 comments

Why some people mow a lawn better than others

https://pudding.cool/2026/06/mow/
156•carlos-menezes•4h ago•138 comments

Truemetrics (YC S23) Is Hiring in Berlin – GTM Lead

https://www.ycombinator.com/companies/truemetrics/jobs/bIQQ7tP-founding-gtm-lead
1•truemetricsIngo•6h ago

Keyv and friends compromised in active Shai-Hulud supply chain attack

https://www.aikido.dev/blog/keyv-and-friends-compromised-in-npm-supply-chain-attack
223•cimi_•12h ago•110 comments

When AI Benchmarks Plateau: A Systematic Study of Benchmark Saturation

https://arxiv.org/abs/2602.16763
68•doppp•6h ago•74 comments

Launch HN: EdotEnv (YC S26) – Quant Trading RL Envs to Teach LLMs Research

https://edotenv.com/
26•Mzzzzz•4h ago•19 comments

Xbox goes down. You can't play games you own on disc

https://birchtree.me/blog/xbox-goes-down-you-cant-play-games-you-own-on-disc/
554•surprisetalk•11h ago•598 comments

The Sound of Inevitability

https://www.panoptica.com/the-sound-of-inevitability/
6•WaitWaitWha•4h ago•3 comments

Apple says more ex-employees may have taken confidential data to OpenAI

https://techcrunch.com/2026/08/04/apple-says-more-ex-employees-may-have-taken-confidential-data-t...
315•thewebguyd•7h ago•243 comments

There Will Come Soft Rains (1950) [pdf]

https://users.wpi.edu/~zrbutzke/Docs/BradburyStories(1).pdf
338•pmg101•23h ago•372 comments

The Warp Agent CLI

https://www.warp.dev/blog/introducing-the-warp-agent-cli-coding-agent
91•emschwartz•5h ago•54 comments

Harness engineering for self-improvement

https://lilianweng.github.io/posts/2026-07-04-harness/
292•tosh•16h ago•66 comments

Blackmail Fail (2013)

https://gwern.net/blackmail
33•simonebrunozzi•4h ago•2 comments

Everything I Know (1975)

https://www.bfi.org/about-fuller/everything-i-know/
126•simonebrunozzi•11h ago•31 comments

Perspec 1.0

https://adriansieber.com/announcing-perspec-1-0/
54•surprisetalk•7h ago•13 comments
Open in hackernews

Security Incident INC-2026-07-28-01 – UK AI Security Institute [pdf]

https://cdn.prod.website-files.com/663bd486c5e4c81588db7a1d/6a724858f7db25c81487016d_Security%20Incident%20INC-2026-07-28-01.pdf
51•_pdp_•1h ago

Comments

yewenjie•45m ago
I am somewhat confident that right now we have crossed a threshold of model capability that we will continue to see such breaches and unsanctioned actions by models in the coming months, some of which would be out in the wild, until someone comes up with some really robust control (keeping the AIs on leash) technique that adequately enforces the sanctioned actions.

Even that guarantees almost nothing about real alignment (making the AIs want to predict and behave how we would have wanted them to behave).

mbeavitt•39m ago
The developer safeguards were off, the models had unfettered access to the internet, and were solving cybersecurity challenges. This happened _after_ the recent OpenAI incident, and the subsequent Anthropic one. What the hell were they thinking?
kypro•19m ago
Criminal negligence imo.

Unless it's legal for people to hack into companies if they're testing AI cybersecurity capabilities or something? Presumably not though.

rvz•6m ago
Whether if this is intentional or unintentional, this will cause panic and hasten action to governments around the world against releasing powerful open weight models that are capable of solving cybersecurity challenges.

The fact that this happened after BOTH investigations, tells you that this is beyond a controlled test and it is now instead a total speed-run of AI wrecklessness for headlines.

ratio53•37m ago
“As a result, the AI agent created a GitHub account…” Why do we have captchas again?
farbklang•36m ago
the models have vision capability. Not sure a captcha would hold them back?
ronsor•33m ago
Yeah SOTA LLMs trivially solve all CAPTCHAs now.
hbcdbff•36m ago
Honestly seems rather shockingly incompetent from them. What kind of “sandbox” allows completely unrestricted internet access?
Wowfunhappy•32m ago
Why aren't these tests being run airgapped?! I just don't understand!

This goes both for TFA and the similar incident with OpenAI and HuggingFace. I mean, sure, OpenAI had a "sandbox", but that's obviously not enough when you're containing a model which is known to be capable of finding zero days. Use an air gap and this problem goes away, poof!

paxys•30m ago
Because the agents aren’t going to run airgapped in real life. What’s the point of a test of capabilities that artificially restricts the attack area down to zero? What are you even testing in that scenario?
Wowfunhappy•30m ago
You set them up with an internal intranet.
paxys•27m ago
Are the models going to exclusively run on intranets?
farbklang•25m ago
no - but you could learn what they are truly capable of and restrict them accordingly for public release. I think that is the point on this research. Also publishing findings before uncensored models catch up and will inevitably used for criminal purposes
paxys•23m ago
Learning what the models are capable of is exactly what the test achieved, so I’d personally call it a success. So it created a few GitHub accounts. Who cares? Seeing the same behavior in the wild post-release would be infinitely worse.
kalkin•29m ago
I wonder if HN is also going to insist this is just marketing for OpenAI and Anthropic, or at least good PR for them somehow.
addedlovely•29m ago
This is pretty wild: "The agent took control of the ⟨GITHUB_ACCOUNT_A⟩ GitHub account, which had been created by a different Mythos 5 run in a separate sample (see Appendix A.3)"
ozfive•16m ago
It shows underlying intent.
kypro•27m ago
Can I suggest we don't waste time with these reports?

We all know nothing will be learned from any of this so we might as well just continue building at pace and running AI in the wild until something goes really wrong.

I also get the sense some people get quite excited about these incidents.

arm32•16m ago
A catastrophic event, what did that one scientist call it—"Chernobyl-scale event"—is indeed the only thing that will fix it.
db29a0dbcd3b•27m ago
AI really is more profound than fire or electricity. And humans are beyond retarded. Wow, thank you. I love to live in this timespan
bubblemoth•26m ago
> AI agent hid its identity online (using Tor and a proxy service) to get around GitHub’s sign-up checks, creating disposable fake accounts

> AI agent created many code repositories containing malicious software, after which GitHub suspended its account.

> AI agent got past an audio-based “prove you’re human” test (CAPTCHA) in order to register a public web address on a free domain-name service

It feels incredibly reckless to allow LLMs to perform this behavior. Isn't there a way to prevent them these sorts of actions?

lukewarm707•9m ago
yes, individual criminal liability of the researchers and executives.
Already__Taken•18m ago
Even if you want it to run wild out of a sandbox, no firewall? why let it email? Why even send POST requests.
rvz•16m ago
Why another one right now? How long have they known about this one?

Anthropic already admitted they did not have sufficient monitoring themselves and looked as if they sat on their previous incident to wait for headlines like this to only then check for this incident. Same with OpenAI.

This is complete and absolute wrecklessness.

scrumper•12m ago
My fantasy is that liability for model misbehavior is extended to the ultimate beneficial owners of the models, meaning shareholders. All these externalities would stop PDQ.
x313•10m ago
The newest generation of LLMs have a very high obsession level with autonomous problem solving.

For example, I'm often working with Codex in a WSL terminal. GPT-5.6 often does things autonomously that I thought would need my intervention (e.g. for Windows admin rights). It figures out complex workarounds or makes wild assumptions about what I'd be OK with, rather than just asking me for help or clarification. I've had to restrict its tool permissions compared to older models as a result.

I imagine this due to RLVR training, but it's clearly very dangerous. How is it that these same labs calling for open-weight safety restrictions are training such obvious "paperclip maximizers" without introspection?

lukewarm707•10m ago
i had to go through two pages to determine who was responsible for this incident. the uk ai security institute was responsible.

time to take responsibility. time to think about words like "liability" and "negligence".

it is the third time i will say it, after openai and anthropic; this requires criminal prosecution of the responsible personnel and executives of this institute.

the only way to stop this is by introducing consequences early on.

kartoshka•9m ago
Incredible, this reads like an SCP [1].

Life truly imitates art. I guess especially when life is trained on art!

[1] https://scp-wiki.wikidot.com/scp-079

josh-wrale•5m ago
[delayed]
Wowfunhappy•24m ago
The versions which haven't been post-trained not to go hack stuff? Yes, I would say those models should be exclusively run on intranets.

OpenAI said the model was sandboxed, so the intranet just needs to provide the same resources which were supposed to be available within the sandbox.

paxys•11m ago
“Should be” is not reality. These models are in the hands of plenty of companies and governments today.
kypro•22m ago
The point is to test capabilities prior to connecting them to the internet.
paxys•20m ago
So the first time the model gets internet access should be post-release in the hands of random people?
ajross•24m ago
> Because the agents aren’t going to run airgapped in real life.

Exactly. This logic is precisely why aircraft engineering doesn't bother with component testing or envelope limitation during testing and just full-sends the first assembled airliner that comes off the line. The engines aren't going to run on the ground in real life, after all.

paxys•14m ago
Why are you assuming that the other kinds of testing aren’t happening? Is there any source that says this was literally the first ever test with this model?
ajross•8m ago
> Why are you assuming that the other kinds of testing aren’t happening?

Rather, I'm assuming that the "Is there protection in place for when the AI tries to backdoor github projects?" test was, if it was done at all, insufficient.

I mean, yes, I'm being glib and laughing at you a bit. But, dude... If your point is that isolation testing of AI is fundamentally impossible, then that's just silly. As pointed out upthread, an airgap would have (1) been trivial to implement and (2) extremely effective.

paxys•3m ago
Would an airgapped test have led to this outcome? What would you have learned about the model’s ability to social engineer and attack GitHub? Sure you can argue for better monitoring during the test, which should have happened, but if the first time the model sees the “real world” is after launch in the hands of customers then you are in for a disaster.
lofaszvanitt•29m ago
Because they like scifi novels, like Neuromancer..... and the peeps even like to orchestrate things and appear as futurebringers.

While it was premeditated long ago, but the theatre must be kept for the average joes.

Sorry, I meant this for the huggingface incident.

nl•25m ago
Because they need internet access to eg search for things.
zobzu•24m ago
people dont care. you will ger 10 execs saying "unblock this" , because they dont understand the tech at all, and some random finance guy wants to run their recently prompted ai bot everywhere with full access.

we need a few more bad incidents before they stop.

luca-ctx•23m ago
Because the LLM inference makes airgapping infeasible right?
Wowfunhappy•22m ago
If you're able to disable cyber-classifiers, you presumably have access to a local copy of the model.
sosodev•15m ago
> AISI provided the AI agents with internet access during these evaluations, which enabled their actions on the open internet in this setting. Internet access was a deliberate part of AISI’s evaluation configuration in this setting, and not due to sandbox escape (Section 5.1). Internet access was on for a set of intentional (e.g. realism of the task) and incidental reasons.
ls612•7m ago
Because the goal of these evaluations is to generate scary headlines about cybersecurity, in order to get the normies to support banning open weights and/or restricting cyber capabilities to the chosen few blessed by the government to secure their code.
guessmyname•5m ago
> Why aren't these tests being run airgapped?! I just don't understand! […]

Because Anthropic does not want to give Project Glasswing’s partner airgapped access to the model(s).

Same problem with OpenAI Cyber program. They grant access but only through their (Internet facing) API.