frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Too much efficiency makes everything worse

https://sohl-dickstein.github.io/2022/11/06/strong-Goodhart.html
1•yurivish•1m ago•0 comments

Liveness Proofs in Veil, Part I: The First Step

https://proofsandintuitions.net/2026/06/24/liveness-proofs-in-veil-part-1/
2•abiro•1m ago•0 comments

Show HN: XHFS starting from 0.6.x is now fully CoW

https://codeberg.org/futureg-lab/xhfs
2•michael-0acf4•3m ago•0 comments

Acer Relaunches Packard Bell with Full New Lifestyle Collection: Tech It Easy

https://news.acer.com/acer-relaunches-packard-bell-with-full-new-lifestyle-collection-to-tech-it-...
1•taubek•5m ago•0 comments

Research acceleration: The view inside OpenAI

https://openai.com/index/research-acceleration-view-inside-openai
2•iamsyr•5m ago•0 comments

Unlimited Free DeepSeek API

https://y-api.bestvirtualgoods.com/
1•freeourdays•5m ago•0 comments

A recruiter who placed creative talent told me the truth

https://twitter.com/Marcellinoposts/status/2096583042315825244
1•bilsbie•6m ago•0 comments

Was Huawei's rise built on crime? A Brooklyn jury will decide

https://www.ft.com/content/9422805f-519c-4b22-839c-cbd7a6464f07
1•sbulaev•6m ago•0 comments

The universe may have been building rocky planets almost from the start

https://www.sciencenews.org/article/universe-rocky-planets-formed-early
2•KinetiNode•12m ago•0 comments

Show HN: A local, Apple Intelligence meeting assistant

https://github.com/rcarmo/swift-smart-prompter
3•rcarmo•12m ago•0 comments

I Feel about AI

https://beza1e1.tuxen.de/ai_feelings.html
8•qznc•13m ago•0 comments

Douglas Hofstadter: Analogy as the Core of Cognition [video]

https://www.youtube.com/watch?v=n8m7lFQ3njk
1•tosh•14m ago•0 comments

Feather Atlas

https://www.fws.gov/lab/featheratlas/identify.html
2•bookofjoe•14m ago•0 comments

AI made it worth building games for our game nights

https://mmazzarolo.com/blog/2026-09-05-ai-has-completely-changed-our-game-nights/
5•mmazzarolo•15m ago•1 comments

Show HN: 188M Hindi encoder, 28B tokens, 8K context, 1× RTX 4090

https://github.com/kkkamur07/indic-modernBERT
1•kkkamur•16m ago•0 comments

Release 0.12.10 · astral-sh/uv

https://github.com/astral-sh/uv/releases/tag/0.12.10
1•kyisaiah47•16m ago•0 comments

The Garage Where Silicon Valley Began

https://www.coolidgereview.com/articles/garage-silicon-valley
1•RickJWagner•17m ago•1 comments

Dirty Coding Tricks, part 3

https://dalerank.github.io/en/articles/dirty-coding-tricks-3.html
2•birdculture•18m ago•0 comments

Recreating Minecraft Is Not a Benchmark

https://kuber.studio/blog/Reflections/Recreating-Minecraft-is-Not-a-Benchmark
1•kuberwastaken•21m ago•0 comments

Mirror of your own self – on AI sycophancy and getting what we asked for

https://substack.com/app-link/post
2•yenniejun111•23m ago•0 comments

Artificial Intelligence as a Judge Advocate's Force Multiplier

https://www.lineofdeparture.army.mil/Journals/Army-Lawyer/Archive/Issue-2-2026/AI-as-Force-Multip...
1•everybodyknows•24m ago•0 comments

Computer Science Achievement and Writing Skills Predict Vibe Coding Proficiency

https://arxiv.org/abs/2603.14133
1•iamflimflam1•25m ago•0 comments

Speculative Programmatic Tool Calling

https://alexzhang13.github.io/blog/2026/spec-ptc/
1•gmays•28m ago•0 comments

Over 60k arrested for communication offences as trivial as viewing TikTok videos

https://www.gbnews.com/news/free-speech-row-more-than-60000-arrested-communications-offences
2•josephcsible•30m ago•0 comments

Americans are vibe-coding trading algorithms

https://www.wsj.com/tech/ai/the-ai-shift-turning-everyday-investors-into-mini-quant-funds-ebe4d45f
4•simonpure•31m ago•0 comments

A cross-chain DEX for atomic swaps, without middleman

https://basicswapdex.com/
2•dRK-U•31m ago•0 comments

Nutrient-supplemented carbonized aerogels for oil adsorption and biodegradation

https://www.sciencedirect.com/science/article/pii/S2666821126003583
1•wslh•35m ago•1 comments

Ryoku – A hand-built Omarchy alternative

https://ryoku.dev/
3•hkalbasi•35m ago•0 comments

Show HN: Kadō – open-source habit tracker, with non-binary habit score, for iOS

https://github.com/scastiel/kado
13•scastiel•39m ago•5 comments

A/I shuts down – Stay human

https://keepitfree.ai/announcements/a/i-shuts-down-stay-human/
123•captainmuon•39m ago•16 comments
Open in hackernews

GPT-6 Astra soundly defeats Fable 5.1 on recognizing handwritten corrections

https://dorrit.pairsys.ai/
3•svcrunch•49m ago

Comments

svcrunch•49m ago
The [Pelican Benchmark](https://github.com/simonw/pelican-bicycle) is in the LLM's training data and probably not a useful indicator of improving LLM capabilities any more.

For the past two years, I've run the Little Dorrit Editor Benchmark. Typesetting is my hobby and I wanted to see how well LLMs could extract editorial marks from a printed page.

The initial results were not encouraging, but performance has risen rapidly since the summer of 2025. From F1 scores in the low 0.2s in 2024, we are now at 0.78 with GPT 6 Astra! Fable 5.1 scores 0.73 (high thinking mode for both).

The best thing about this benchmark is that there is still plenty of room to climb.