frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Some thoughts about Anthropic's new cryptanalysis results

https://blog.cryptographyengineering.com/2026/07/29/some-notes-about-anthropics-new-results/
32•supermatou•1h ago

Comments

mkagenius•47m ago
I also just finished writing a note - https://mkagenius.substack.com/p/notes-on-mythos-breaking-ae...

https://news.ycombinator.com/item?id=49099977

john_strinlai•32m ago
>They [anthropic] appear to have just told it to get some results and then strapped its nose to the grindstone until it found some.

it is fun how well this works.

i cant find the link immediately (will look and edit with it), but somewhere in the "hello there the jacobian conjecture is false thanx" thread, someone brought up a different conjecture breakthrough where the prompts were basically just repeated "no, keep going" until a result was found.

edit: https://chatgpt.com/share/6a60b2eb-0b64-83ee-9c76-7931ca1de0...

i especially like "you should do a breakthrough". each prompt is less than ~20 words. makes me really question the whole "prompt engineering" stuff.

throwup238•12m ago
> i especially like "you should do a breakthrough". each prompt is less than ~20 words. makes me really question the whole "prompt engineering" stuff.

This is well past prompt engineering and into process engineering like six sigma. Just like in an early industrial revolution factory, we’re all still figuring out what works in the process of making stuff except this is so early that even simple things like “make this screw standardized” (or “no, keep going” in this case) is really high impact.

The degrees of freedom an LLM has is so large that we're going to be exploring their capabilities for decades, especially if they continue to get better. This is why IMO experts are always going to be better at LLMs in their field because they can force them LLM into processes (think prompt engineering -> CC dynamic workflows) that follow their work processes and get much better results out of them than “keep going.”

simonw•27m ago
> both outputs of Claude Mythos, their (still) unreleased advanced model

That sentence gives the impression that Mythos might be released in the future. That's clearly not going to happen - it's already "released" in as much as selected, trusted partners can access it, and the rest of us get it in the form of Fable - which is Mythos but with filters that downgrade you if you try to use it for anything even remotely related to cybersecurity or biology.

(The other day Fable 5 downgraded me to Opus after I asked it to explain the difference between tusks and teeth.)

free_bip•7m ago
Genuine question here, why would you ask Fable to explain the difference between tusks and teeth? That's a task that can probably be handled by Haiku.
simonw•7m ago
It's the default when I pop open the Claude iPhone app, I usually don't bother to switch it.
andy99•3m ago
I constantly see some version of this, people implying it’s the users fault for wanting to do something nobody should want to do. It’s ridiculous Stockholm syndrome stuff. The right framing is why would I want thx providers position to do anything, and it becomes even more of a problem when it’s stuff that’s clearly benign, even if you can’t think of a reason I’d want to do it. Should we always have to explain ourselves when doing a weird thing because authorities don’t know how to manage false positives?
throawayonthe•25m ago
> I asked Claude for its thoughts, and it doesn’t mince words: “what makes this genuinely interesting — and, frankly, a little embarrassing for the field — is that none of the ingredients are exotic.” The TL;DR is that someone just did a much more thorough job applying all of our known tools. In short: the sort of things that attack AIs are wonderful at.

did i just read two summaries/TLDRs (in a row) of the already-two-sentence summary right above?

simonw•19m ago
This is good:

> If you’re under the impression that these models are “glorified autocomplete” or that progress is slowing down, I need to urge you: stop thinking that. The models are very intelligent and capable, they are getting better at a fast clip. I can cite measurable and impressive progress over just the past five months on specific types of problem I’ve asked them to look at. [...]

> On the other hand: if you think that models are super-intelligent or that AGI is already here, you should also stop thinking that. Working with these tools is like swimming in a pond where the ground drops off sharply. One minute you’re wading comfortably and there’s support under your feet. Then suddenly you cross a specific line, and you’re back to swimming on your own.

Show HN: Open-source engine running Gemma 4 26B in 2 GB RAM on any M-series Mac

https://github.com/drumih/turbo-fieldfare
385•gitpusher42•3h ago•119 comments

Superlogical

https://www.superlogical.com/
240•yan•2h ago•179 comments

Keychron announces first open-source firmware for gaming mice

https://www.digitalfoundry.net/news/2026/07/keychron-announces-first-open-source-firmware-for-gam...
87•JLO64•2h ago•32 comments

KOReader

https://koreader.rocks/
552•Cider9986•7h ago•175 comments

Handbook.md shows that long policy documents do not reliably govern agents

https://arxiv.org/abs/2607.25398
224•spIrr•5h ago•145 comments

Anatomy of a frontier-lab agent intrusion

https://huggingface.co/blog/agent-intrusion-technical-timeline
111•dn2k•3h ago•38 comments

Document-borne AI worms can self-propagate through Copilot for Word

https://enklypesalt.com/posts/context-collapse-part3-ai-worming-through-word/
263•Canopy9560•6h ago•198 comments

Some thoughts about Anthropic's new cryptanalysis results

https://blog.cryptographyengineering.com/2026/07/29/some-notes-about-anthropics-new-results/
32•supermatou•1h ago•8 comments

PgDog (YC P25) Is Hiring

https://www.ycombinator.com/companies/pgdog/jobs/uWymUYy-founding-software-engineer
1•levkk•1h ago

A.I. companies are recruiting electricians and carpenters by the thousands

https://www.nytimes.com/2026/07/29/business/economy/data-center-electricians-training.html
86•thm•3h ago•124 comments

The Rust on ESP Book

https://docs.espressif.com/projects/rust/book/
51•AlexeyBrin•4d ago•7 comments

Hamburg's Stadtpark: A Park Built to Be Used

https://alsterrunde.com/hamburgs-stadtpark-a-park-built-to-be-used/
62•mertbio•2d ago•11 comments

Launch HN: Tokenless (YC S26) – Automatic model switching to save money

https://usetokenless.com/
36•rohaga•2h ago•33 comments

Darktable

https://www.darktable.org/
191•siatko•6h ago•102 comments

Self-hosting Kimi K3: 20% more hardware cost, 20% better task resolution

https://aistack.imec-int.com/blog/gpu-self-hosting
62•flifenstein•4h ago•25 comments

Shipping Godot VR and Porting to PSVR2: A Partial Post Mortem

https://www.claire-blackshaw.com/blog/2026/07/shipping-godot-vr-and-porting-to-psvr2-a-partial-po...
86•ibobev•5h ago•1 comments

Show HN: CheapFoodMap – A map of good meals under $10

https://cheapfoodmap.com/
18•jaep1•1h ago•19 comments

GPT-5.6 vs. Claude Fable 5 for Physical AI, which performs best?

https://juliahub.com/blog/frontier-models-physical-ai-evaluation
49•mbauman•3h ago•10 comments

Learning Musical Multitasking

https://www.jefftk.com/p/learning-musical-multitasking
13•surprisetalk•5d ago•8 comments

More Tailscale tricks for your jailbroken Kindle

https://tailscale.com/blog/jailbroken-kindle-proxy-tun-modes
366•Error6571•13h ago•104 comments

Show HN: Kedge – Full-stack cloud with forkable VM snapshots and global SQLite

https://kedge.dev/
29•wgjordan•2h ago•8 comments

Show HN: Qwen Scribe – local transcription and dictation for Apple Silicon

https://github.com/VladUZH/qwen-scribe
32•sidclaw•3h ago•11 comments

Amiga Graphics Archive

https://amiga.lychesis.net/index.html
129•Bluestein•8h ago•23 comments

Hunter-gatherers introduced fish to a mountain lake 7000 years ago

https://www.newscientist.com/article/2580119-hunter-gatherers-introduced-fish-to-a-mountain-lake-...
103•stevenwoo•2d ago•81 comments

User Interfaces of the Demo Scene

https://www.datagubbe.se/scenegui/
364•zdw•14h ago•63 comments

Disrupting supply chain attacks on NPM and GitHub Actions

https://github.blog/security/supply-chain-security/disrupting-supply-chain-attacks-on-npm-and-git...
75•nyku•6h ago•26 comments

Cesium DevCon 2026 talks are up, including a keynote from SQLite's creator

https://cesium.com/events/cesium-developer-conference/2026/
31•jasteinerman•3h ago•8 comments

SQLite in Production: Optimizing WAL Mode, Concurrency, and VFS Layers

https://micrologics.org/blog/sqlite-in-production-optimizing-wal-mode-concurrency-and-vfs-layers-...
210•ankitg12•11h ago•67 comments

Lisp moving Forth moving Lisp

https://letoverlambda.com/textmode.cl/guest/chap8.html
106•fallat•3d ago•25 comments

SpecForge – A Platform for Authoring Formal Specifications

https://docs.imiron.io/v/0.5.10/en/tour.html
65•agnishom•8h ago•7 comments