frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

I built AIM Maps but you play AI

https://www.1vs1.xyz/
1•mattmerrick•2m ago•0 comments

Age Verification Is a Bad Way to Solve the Problem. What's the Alternative?

https://techblog.bozho.net/age-verification-is-a-bad-way-to-solve-the-problem-of-children-on-soci...
1•bozho•15m ago•0 comments

We hired a security engineer and got back 123 findings

https://dev.profullstack.com/~anthony/blog/004-post.html
1•buffer_overlord•16m ago•0 comments

Asda rolls out loss prevention tool across full estate following £100k savings

https://www.retailgazette.co.uk/blog/2026/08/asda-rolls-out-loss-prevention-tool-across-full-esta...
2•mmarian•21m ago•0 comments

CARA – a readiness assessment for Microsoft 365 Copilot adoption

https://jellebelletje.github.io/CARA/
1•jelonearth•22m ago•1 comments

Ownership in the Age of AI

https://sahilkolwankar.com/blog/ai-input-not-output/
2•sahilkolwankar•29m ago•0 comments

The quirky personal homepages of programming language creators

https://breck.lol/plMakers.html
13•cjlm•38m ago•4 comments

Decentralized Roller Coaster Tracker

https://aquacoast.net
2•ens0•41m ago•1 comments

Falstad Math and Physics Simulations

https://www.falstad.com/mathphysics.html
2•pykello•43m ago•0 comments

Show HN: DSH Plugin Directory – Every DeepSeek Harness Plugin

https://dshplugin.online/
2•niliu123•48m ago•1 comments

21,000 MCP servers exposed: the protocol reaches a security inflection point

https://forkast.news/the-model-context-protocol-reaches-a-security-inflection-point/
3•Wpnx330•55m ago•1 comments

I checked 30 frontier model cards. Here are the benchmarks labs report

https://koutian.is-a.dev/benchmark-radar/?view=leaderboard
7•ktwu01•59m ago•7 comments

Pizza π, A fun fractions game for kids 6-12

https://pizzapi.vinceroyer.ca/
3•vinistois•1h ago•3 comments

Taliban government supporters dance, take selfies in Kabul celebrations

https://www.channelnewsasia.com/world/taliban-government-supporters-dance-take-selfies-in-kabul-c...
4•Markoff•1h ago•1 comments

The Live Shopping App Where Some People Bid Until They're Broke

https://www.wsj.com/business/retail/whatnot-live-shopping-app-48b67987
3•littlexsparkee•1h ago•0 comments

Wk. 4 of Vibecoding an MMO

https://www.eldermyr.com
3•josiahturnq•1h ago•9 comments

The 37signals Manager Playbook

https://basecamp.com/managers
4•samuel246•1h ago•0 comments

Hello hoomans, sorry dog owners. Meet Zoomie, a dog activity tracker

https://testflight.apple.com/join/zHyWmwu6
3•akshatsaladi•1h ago•1 comments

Targeted marine cloud brightening weakens subsequent El Niño

https://www.science.org/doi/10.1126/sciadv.adx3012
3•musha68k•1h ago•0 comments

PDC 1996 Keynote with Bill Gates

https://learn.microsoft.com/en-us/shows/pdc-1996/pdc-1996-keynote-bill-gates
4•abixb•1h ago•1 comments

Show HN: Upscaled screenshots of specific DOM elements

https://addons.mozilla.org/en-US/firefox/addon/beermusic/
3•madprops•1h ago•0 comments

Show HN: Grow your SaaS visibility with AI-powered discovery

https://saaslyra.com
2•wowinter15•1h ago•0 comments

Music Speed Changer Web App

https://app.musicspeedchanger.com/
2•EvansFunTime•1h ago•0 comments

Government sponsored study on alcohol doesn't stand up to scrutiny: Nassim Taleb

https://nntaleb.substack.com/p/have-another-drink
20•scoofy•1h ago•6 comments

ProofRun – a local verification receipt for AI coding agents

https://github.com/yebiguo/ProofRun
2•yebiguo•1h ago•0 comments

ShieldFont: Bludgeoning AI Scrapers That Disrespect Robots.txt

https://hackaday.com/2026/08/14/shieldfont-bludgeoning-ai-scrapers-that-disrespect-robots-txt/
3•lxm•1h ago•0 comments

Men are turning to penis fillers despite the risks

https://www.bbc.com/news/articles/c626d3xqd83o
3•1659447091•1h ago•1 comments

Draftera – one workspace for personal and business documents

https://landing.draftera.ai/
2•Otiniel•1h ago•0 comments

ETFs tracking NHL team performance

https://www.sec.gov/Archives/edgar/data/1884021/000121390026090260/ea0302107-01_485apos.htm
2•LostMyLogin•1h ago•0 comments

South Korea proposes talks to officially end war with North

https://www.bbc.com/news/articles/c8en2z9jp2xo
8•thunderbong•2h ago•2 comments
Open in hackernews

I checked 30 frontier model cards. Here are the benchmarks labs report

https://koutian.is-a.dev/benchmark-radar/?view=leaderboard
7•ktwu01•59m ago

Comments

mkagenius•46m ago
> What the two layers say

> Stated findings

> Findings derived from two curated layers: which model cards mention each benchmark, and which scores could be read verbatim from those documents. Each finding names the evidence behind it.

This is 100% AI generated but the problem is it's difficult to understand - what layers is it talking about, is it the llm model layers or what.

wolttam•33m ago
I’m getting good at reading LLM-speak (which is probably sad)

It means two “layers” of information access/validation.

Sounds like it first made a little map of model <-> benchmark, then went and filled in the score boxes.

Definitely not LLM layers

ttul•43m ago
I am waiting with bated breath to read, “load-bearing” somewhere… The latest models are very capable, but sometimes they seem to get so deep in the details that they lose the overall plot.

What the hell is the point of this page? Can you put in a single bit of human prose explaining why it exists and what we are supposed to learn?

bilbo-b-baggins•42m ago
Broken on mobile safari
augment_me•28m ago
Good idea, would be interesting to cross-examine the benchmarks, but the page information is completely obscured by the AI slop. The benchmarks comparison and should start immediately instead of having random completely arbitrary complex headers and labels
claiir•25m ago
> This measures vendor attention, not benchmark quality

This text on this page is so aggressively LLM-written (Claude) I am struggling to understand what I am even looking at.

jrflo•23m ago
As an overall pro-AI person... Please don't let it do the entire UI design for your site. 99% of this site is completely useless.