frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Intent, Not Sophistication: The AI Attacker Is on the Record

https://simonroses.com/2026/09/intent-not-sophistication-the-ai-attacker-is-on-the-record/
1•speckx•1m ago•0 comments

Ask HN: How do you manage the volume of generated slop?

1•throwaway1183•1m ago•0 comments

America's biggest convenience-store chains overcharge consumers

https://www.theguardian.com/us-news/ng-interactive/2026/sep/24/7-eleven-circle-k-convenience-stor...
1•toomuchtodo•1m ago•1 comments

Evidence: The Missing Runtime Stream (a.k.a. Not Logs or Telemetry)

https://gluufederation.medium.com/evidence-the-missing-runtime-stream-57cdd76661d8
1•idm_guru•2m ago•0 comments

Robot Basketball

https://kevinkelly.substack.com/p/robot-basketball
1•surprisetalk•5m ago•1 comments

Fedora Discussion Raises the Idea of Replacing LibreOffice with Collabora Office

https://www.phoronix.com/news/Fedora-Discuss-Collabora-Office
1•cf100clunk•6m ago•1 comments

Proof a Human Sent an Email

https://github.com/Inkline-Verify/ProofOfHuman
2•georginaalcaraz•8m ago•0 comments

The Zig Journey

https://kristoff.it/blog/the-zig-journey/
3•kristoff_it•8m ago•1 comments

Entrepreneurship and Industry in Europe

https://eudata.vercel.app/
2•tosh•12m ago•1 comments

Artificial Symbiotic Intelligence

https://institute.deepmind.com/essays/artificial-symbiotic-intelligence/
2•kristianpaul•12m ago•1 comments

Natural-Fiber Clothing Is in Demand–But Scientists Say It Needs More Scrutiny

https://www.wsj.com/style/fashion/cotton-fiber-pfas-47042bd2
1•geneticdrifts•13m ago•0 comments

Japanese used bookstores see 5x sales surge as books are being bought by the ton

https://www.tomshardware.com/tech-industry/artificial-intelligence/japanese-used-bookstores-see-5...
1•speckx•14m ago•0 comments

Breaking Up with Google Play: Why Conversations Is Now Free

https://gultsch.de/posts/breaking-up-with-google-play/
4•inputmice•16m ago•0 comments

$5,999 Ryzen AI MAX+ 395 and Radeon R9700 workstation

https://videocardz.com/newz/lucebox-opens-direct-orders-for-5999-ryzen-ai-max-395-and-radeon-r970...
2•GreenGames•17m ago•0 comments

Map of where crypto and privacy services are sanctioned or blocked

https://dontkyc.me/sanktionsradar
1•paco4747•17m ago•0 comments

Analyst Calls for AMD Investigation over Restricted Chips in China

https://www.tomshardware.com/tech-industry/leading-semiconductor-analyst-accuses-amd-of-treason-o...
2•rndsignals•18m ago•0 comments

Eutetic: Codex for Data Scientist and Data Engineers

https://eutetic.com
1•FrancescoMassa•21m ago•1 comments

The worlds top cities for 2026 – and why they stand out

https://www.bbc.com/travel/article/20260923-the-worlds-top-cities-for-2026-and-why-they-stand-out
1•bookofjoe•22m ago•1 comments

Show HN: Hermes, Kokoro and Parakeet, behind a face you can talk to

https://memoe.app/
1•yisraelgottz•22m ago•0 comments

Show HN: Resume Claude Code subagents killed by a usage limit, don't redo them

https://github.com/error0702/agent-limit-retry
2•autorunfun•24m ago•0 comments

ask HN:Are we locking ourselves in the Google ecosystem?

3•KinetiNode•24m ago•6 comments

The most-cited vector DB benchmark is not publishing results from competitors

https://medium.com/@glitcher00/the-most-cited-vector-database-benchmark-is-not-publishing-results...
1•austin2020•25m ago•0 comments

Show HN: ContextDebt – finds code whose stated reason to exist has expired

https://github.com/ContextDebt/contextdebt
1•wittachai•26m ago•0 comments

Submersion AI Debuts Basin

https://submersion.ai/news/basin-cybergym-ranking/
1•Bluestein•26m ago•0 comments

AI Safety Is Mostly a Sex Cult in Berkeley

https://www.verysane.ai/p/ai-safety-is-mostly-a-sex-cult-in
10•martythemaniak•27m ago•0 comments

When to Use an LLM

https://extralongdivision.com/articles/when-to-use-an-llm/
1•extralongdivisi•28m ago•0 comments

Humans Are Reading Your ChatGPT Chats, Lawsuit Claims

https://openclassactions.com/lawsuits/privacy/openai-chatgpt-human-review-project-lily-class-acti...
4•nreece•29m ago•1 comments

Show HN: PreFlight – Local citation verifier and claim auditor for LaTeX

https://preflightai.tech/workbench
1•stackfrost•29m ago•0 comments

Show HN: PlaceCall (YC W26) – agentic API to call businesses and get things done

6•ymarkov•29m ago•0 comments

The Antennagate Q&A, which has never been available, has been posted today [video]

https://www.youtube.com/watch?v=BiN5ERktXz0
2•philistine•30m ago•0 comments
Open in hackernews

Best LLM for every budget, updated daily

https://bestmodelforyourbudget.terrydjony.com/
51•terryds•55m ago

Comments

radial_symmetry•32m ago
The coding and math tabs seem to be missing the latest models...
SkyBelow•19m ago
Math in particular is quite far behind. GPT 5.2 is recommended as the best for highest cost. Really?
pelagicAustral•30m ago
Is there a cheaper model than Gemini 3.8 Flash (High) that maybe/kind-of is on-par with it? For me it works really good but hit the limit in two hours tops... last week was the first time I hit the weekly limit and had to wait 4 days... Claude patches OK, but that is also getting drained really fast these days...
TomGarden•29m ago
is this on the $20-ish sub? I hear ultra lasts for a very long time
pelagicAustral•27m ago
Yeah, 20... I was going to try Deepseek later on, just chuck 20 in there and see how good it does, and how far I can go as well...
TomGarden•25m ago
yeah I enjoy the speed of Gemini, but I also just have the low tier one (I use the 20x Claude and Codex subs for most of my work). For iteration, Gemini is so much fun, but Opus 5.5 is quite fast, as is sol. If you're on a budget, deepseek does look great
pelagicAustral•21m ago
Yeah, I mean, I think in theory I could push it to the next tier on Gemini, but at the same time I wouldn't mind trying something else, since maybe my workloads are not really that smart and I am wasting a lot of computational power on something a cheaper model with similar capabilities can do.
montroser•28m ago
DeepSeek 4.1 Flash is your answer.
pelagicAustral•26m ago
Ah OK, yeah, I was actually going to go for that earlier, I'll check it out once I get home, thanks!
Tepix•30m ago
1. In real life, most of us use token packages like OpenCode Go etc.

It would be handy to have a site like this one that takes into account the various deals and attempts to calculate the number of tokens per monthly fee for a chosen model. I realize this makes the task a lot more difficult.

2. It would be handy to have a chart like that for the AI hardware that people own. It helps you decide which model to run (resulting in different levels of intelligence and speed). Also difficult to please everyone (preprocessing vs token generation for example) and to keep updated!

I found https://llm-list.com/ yesterday and when I had a detailed look, I quickly found outdated entries, for example looking at GLM 5.3 flash it listed several providers as "free" that weren't free any longer.

TomGarden•30m ago
If you maintain this over time, maybe include other sources than just AA and update the design to look less like zero-shot claude styling (I know that font! I know that color! Lol) it's genuinely useful :)
jwolfe•29m ago
This definition of cost is not particularly useful. You want cost per task, not cost per 1m tokens. Artificial Analysis does a good job of this.
hsnewman•28m ago
I'm sure that local LLM will be far cheaper
mrngld•7m ago
Depends on what level of intelligence you're wanting to use. A vanishingly small number of people can or would want to go to the hardware expense of running something like GLM 5.3 Flash, much less something like K3.

And if you want Astra/Fable/Opus frontier level, then there's no option at all.

But if you don't need that, or you don't need speed... That opens up the discussion. I've been impressed even with how Siri's been doing with the Apple Foundation Models in MacOS/iOS 27 given how small they are.

Edit: I can't even fully spec the M5 Ultra Mac Studio you'd need for GLM5.3 Flash since 512GB isn't available yet, but it's already at $9500 for 256GB RAM.

aslkalska•25m ago
all I ever wanted is an updated website where I can see the best models I can run on my different devices locally, I don't get why people are throwing money at these companies
jeremysalwen•24m ago
It's missing Opus 5.5 which was released over a day ago (and also is clearly on the pareto frontier).
shakow•21m ago
Don't know if they updated it since your comment, but for me it's on the graph – and on the Pareto frontier indeed.
SJMG•13m ago
For intelligence only as of 8:51 MT
noumenon1111•3m ago
It was expert timing on the commenter's and the developer's part. Opus 5.5 still doesn't show under coding though.
npongratz•3m ago
Opus 5.5 is indeed on the Intelligence tab, but I don't see it on the Coding or Math tabs.
floppyd•20m ago
As the time goes on it only becomes harder to differentiate between model capabilities with just one or two numbers. I would love to see some kind of multi-axis placement of all the models on less objective attributes, like wordiness, willingness to give up, an ability to "think ahead" and pre-solve possible problems in code, for example, that I didn't think of or didn't think of talking about, etc etc etc.

For example I've been really enjoying Deepseek v4.1 Flash, it's very "straightforward" to the point of being almost dumb sometimes, but it's absolutely relentless and would solve almost any problem no matter how inefficient the solution is.

No idea how to measure all that, just average CoT length per task is probably a good approximation for some things, but not others.

jrflo•17m ago
Does anyone actually pay API costs out of their own pocket? It's about 10x cheaper to just get a codex or chat gpt subscription, it's so heavily subsidized compared to the API that I'm sure it would be cheaper to use frontier models on a subscription plan rather than paying API prices for deepseek flash.
h46u5jytyhtg•14m ago
Large enterprises pay per token even through a ChatGPT “membership”, for example.
jrflo•12m ago
Which is why I said "out of their own pocket"
lowercased•14m ago
I do. For local dev work, I'm mostly using jetbrains' Junie, I can swap between a collection of models from google, openai, anthrophic.

I've had more than a few people tell me "oh, it's so much cheaper to use a $20 claude account" or "i've never hit a limit ever using my openai". Inevitably.. I end up reading/hearing "oh, I need to give it another couple hours to start using it again"... I've never hit that with my approach, even if it's costing me a bit more. Being able to work when I want when I have time has some value.

I also have openai and anthropic direct API billing set up for hosted and client projects that need to call out to an LLM service.

the__alchemist•12m ago
I thought Air was JB's multi-model interface? What is Junie? (I see the buttons, but am very confused by JB's AI offerings in general)

Is it worth the ~10x extra cost over the subscriptions? (This is obviously a leading question). Also, I think you can use OpenAI's subcription login with Air, but not Claude's.

ford•11m ago
You're better off going directly to artificial analysis, this is a feature-poor/misleading/outdated repackaging

Ex. this type of price estimation is quite naive - some models can require 2-3x the number of tokens to achieve the same level of intelligence. Artificial Analysis' own cost per task is a more fair estimation of cost.

Xeoncross•7m ago
If you have a 24-64GB mac, consider running Qwen3.8 27B locally at night. It's a bit slower to run locally, but if you're sleeping it's less of a problem.

Depending on your memory, you'll need to use the weaker Q4 versions but they still perform well.

It ranks higher than GPT-5.3 Codex (xhigh) or Claude Opus 4.6 (max) so is great for pairing with https://github.com/kunchenguid/gnhf for nightly experimentation, cleanup, or recommendation lists for in the morning.

newsy-combi•5m ago
Coding and math graphs are very interesting. Extremely cheap models make it into the upper echelon, delivering 90% of the performance for 1% of the price compared to the #1.
verytrivial•5m ago
I find it interesting that with the given metric comparison, for coding at min 50 strength, every frontier model brand is from a distinct vendor: Ling, Qwen, Gemini, Muse, Grok, GPT, and Claude in increasing value.
jrflo•10m ago
Why not use a codex or claude subscription? If you use the entire usage allotment on the $200 plan it's about $2,000 in equivalent API costs. Switching providers may be valuable but it's quite literally an order of magnitude cheaper.
mmmattt•7m ago
I do, 3-4$ a month of deep seek is enough for my usage
LeBit•6m ago
I use local LLMs on my Mac Mini.

Otherwise DeepSeek Flash 4.1 is dirt cheap (other "Flash" models are not that expensive either). I pay (very few dollars) out of my own pocket.

There are many things where having an API Key is necessary.

Maybe I’ve missed the boat though: is there now a method to use an api key to access a subscription?