frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Oracle bans AI-generated code from OpenJDK

https://app.dealroom.co/news/feed/oracle-bans-ai-generated-code-from-openjdk-despite-ellison-s-cl...
160•delduca•1h ago•95 comments

DeepSeek V4 Flash 0731

https://arcprize.org/results/deepseek-v4-flash-0731
88•tosh•1h ago•53 comments

Assembly Hall of Shame

https://github.com/xoreaxeaxeax/asm-hall-of-shame
33•piotrgrabowski•56m ago•8 comments

An all-sky map of half a million supermassive black holes

https://www.sdss.org/black-hole-mapper-release-20/
80•MarcoDewey•3h ago•28 comments

New Mexico court orders Meta to pay $567m over harms to children’s mental health

https://www.theguardian.com/technology/2026/aug/06/new-mexico-court-meta
637•boplicity•18h ago•355 comments

Making Postgres 300x faster for analytics: batching, operator fusion, and SIMD

https://malisper.me/how-we-made-postgres-hundreds-of-times-faster-the-query-engine/
158•poly2it•7h ago•75 comments

AMD acquires Taalas to boost inference performance by etching models in silicon

https://www.theregister.com/systems/2026/08/06/amd-acquires-ai-chip-startup-taalas-to-boost-infer...
841•itvision•22h ago•637 comments

Show HN: Wyzer Programming Language

https://github.com/Wyzer-Lang/wyzer
142•v0id_isgood•6h ago•78 comments

Möbius-Strip Crosswords

https://quuxplusone.github.io/blog/2026/08/04/mobius-crossword/
25•ibobev•3h ago•4 comments

Kitesurf: Agent-first browser that runs in V8 isolates

https://blog.cloudflare.com/kitesurf/
104•m3h•8h ago•26 comments

Show HN: textlog – A quiet, text-only microblogging platform, open-source, no JS

https://textlog.cc/about
92•stagas•8h ago•38 comments

Petri Nets as a Music Sequencer

https://blog.stackdump.com/posts/petri-net-sequencer
46•m_kos•4d ago•17 comments

Why Are There Statues of Beavers on Top of This Oxford Street Shop?

https://londonist.com/london/history/oxford-street-beavers
33•bookofjoe•4d ago•13 comments

Responding to the next frontier of critical cyber capabilities

https://openai.com/index/responding-next-frontier-critical-cyber-capabilities/
87•artninja1988•2h ago•89 comments

Building Community Out of Strangers

https://tracydurnell.com/2023/11/30/building-community-out-of-strangers/
18•surprisetalk•3d ago•1 comments

Taste Is All That's Left

https://notashelf.dev/posts/taste-is-all-thats-left
623•tsak•1d ago•491 comments

Energizing a vacuum-tube flip-flop module from a 1948 IBM system

https://www.righto.com/2026/07/ibm-604-trigger-tube-module.html
5•geerlingguy•4d ago•1 comments

São Paulo resident transforms degraded area into urban forest

https://saopaulosecreto.com/en/tiquatira-linear-park-en/
291•rmason•5d ago•113 comments

This Mine Predicts Major Wars. It's Opening Again

https://www.bloomberg.com/graphics/2026-opinion-australia-tungsten-mine-us-war-defense-china/
47•mooreds•2h ago•14 comments

Bioengineered chewing gum may offer a way to fight HPV and other microbes

https://www.sciencedaily.com/releases/2026/08/260803080917.htm
195•Audiophilip•21h ago•59 comments

TypeStax: A type scale generator with vintage audio hardware interface

https://www.typestax.com/
44•jasim•1w ago•11 comments

Scientists discover Kelvin-Helmholtz Instability on the surface of the Sun

https://nso.edu/press-release/nsf-inouye-solar-telescope-enables-major-discovery-of-a-hidden-sola...
296•neversaydie•2d ago•60 comments

GitHub Actions and Pages are experiencing degraded availability

https://www.githubstatus.com/incidents/qcvjkzcs7j74
481•Footkerchief•1d ago•400 comments

USA Today Co., partners with Palantir to analyze audience data

https://www.niemanlab.org/2026/08/americas-largest-newspaper-chain-usa-today-co-partners-with-pal...
161•cdrnsf•4h ago•67 comments

Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users

https://openai.com/index/improving-gpt-5-6-sol-in-chatgpt/
300•tedsanders•1d ago•245 comments

A quine in Piet – a GIF image that prints itself [video]

https://www.youtube.com/watch?v=GwMtzhjCzyc
69•surprisetalk•4d ago•8 comments

99% of My Website Traffic Is Bots

https://patronview.com/news/99-percent-of-my-website-traffic-is-bots/
285•petercooper•4h ago•273 comments

Welcoming the Nepalese Government to Have I Been Pwned

https://www.troyhunt.com/welcoming-the-nepalese-government-to-have-i-been-pwned/
203•gnabgib•21h ago•33 comments

The BBC Tetris Companion

https://www.leadedsolder.com/2026/07/28/bbc-bridge-companion-part-1-overview.html
67•zdw•1w ago•10 comments

Launch HN: ProvenMetal (YC S26) delivers circuit boards in days instead of weeks

https://provenmetal.com
221•willcarkner•1d ago•152 comments
Open in hackernews

DeepSeek V4 Flash 0731

https://arcprize.org/results/deepseek-v4-flash-0731
84•tosh•1h ago

Comments

tosh•59m ago
results comparable to gpt 5.6 luna but cheaper

promising!

minimaxir•49m ago
Since the x-axis is log-scaled, DeepSeek is much cheaper than visually implied (mousing over the raw values, it's 1/4th the cost of Luna).
literallyroy•40m ago
Is this pricing from Deepseek with training on usage?
minimaxir•38m ago
Per the announcement tweet, BaseTen was the inference provider which has 20% cache cost that is typical: https://www.baseten.co/library/deepseek-v4-flash-0731/
literallyroy•7m ago
Ah thanks. That looks like 10x cost on cache reads vs Deepseek as the provider: https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...
LUmBULtERA•32m ago
Is it still cheaper than Luna if using an OpenAI subscription? My gut is no, but I have not done the math.
minimaxir•27m ago
Everything is cheaper if using a subscription, but some applications require API usage.
swiftcoder•7m ago
You'd have to compare against something like the OpenCode Go subscription, and I'm fairly sure deepseek napkins out cheaper in that scenario
minimaxir•56m ago
It's always fun when Max reasoning is cheaper than High reasoning.
542458•49m ago
Kimi K3 was an interesting model only a month ago, and now we're looking at the same performance for 1/20th of the price. Wild how fast this is advancing.
whinvik•47m ago
Yeah either the benchmark isn't very useful anymore or V4 Flash is a really, really good model.
ignoramous•29m ago
In my use, DeepSeek v4 Flash (which replaced the quite excellent MiniMax M3) lags behind GLM 5.2 & Muse Spark 1.2 (let alone Kimi K3). Also, K3 is a much bigger multi-modal model, while Flash is text-only and likely optimised for coding tasks.
thehamkercat•29m ago
And now nobody seems interested in it because the price hasn't gone down

it's still $3/$15 for all providers on openrouter

because of some Kimi license

https://openrouter.ai/moonshotai/kimi-k3#providers

muricula•45m ago
Price is confounded by VC subsidies, economies of scale, and inference optimizations. I think a more interesting chart would be ARC AGI vs forwards pass flops or ARC AGI vs training tokens. Of course we don't have those numbers for the closed source models or even some of the open weight ones.
minimaxir•41m ago
DeepSeek V4 Flash 0731 is an open-weights model which means price is determined by competition/invisible hand of the marketplace: https://openrouter.ai/deepseek/deepseek-v4-flash-0731

With the exception of cache costs, all providers have similar input/output costs.

kennywinker•34m ago
Not counting the cost of making the model, which is subsidized by… someone? The chinese gov i think?
npn•39m ago
weak argument. deepseek v4 flash is open weight, you can easily find other providers with competitive price with Deepseek (except for input caching), some even half as cheap.
luyu_wu•44m ago
It is wild that this a log scale of cost to me!
LaurensBER•44m ago
I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day.

OpenCode Go even has double limits temporarily so for 10 USD you effectively get 140 USD of tokens to spend. It would impress me if someone could burn that amount with "normal" usage. Even when running multiple sessions.

I have a Claude Max subscription but I've barely touched it, it just feels like a step back to have to think about limits and usage even though the models are stronger.

The beauty of intelligence at this cost (even if it's not SOTA) is that it opens a whole bunch of new use cases. Test failure in CI? Have the bot automatically propose a fix, its cheap enough that you can discard it w/h issues. Test coverage too low? Auto generate tests on CI for every pull-requests! Monitoring server logs, continuous security audits and investigating every received exception now becomes possible.

I'm thinking about having it automatically filter and re-rank my social media feeds so I can steer the algorithm instead of the other way around.

Perhaps other people (with enormous budgets) were already doing all of the above but for us this is a really exciting release!

Aeolun•41m ago
But DeepSeek now has a warning they’re going to sharply increase their API pricing sometime in the future.
metadat•40m ago
Source?
striking•38m ago
https://www.reddit.com/r/DeepSeek/comments/1vgpysh/deepseek_...
CharlesW•43m ago
Last weeks's discussion (591 points): https://news.ycombinator.com/item?id=49120299
antirez•43m ago
Price is not a good meter. Active parameters per token are. Joule would be even better.
minimaxir•26m ago
Price accounts for computational/architectural efficiency improvements whereas active parameters does not.
orbital-decay•14m ago
It's an excellent metric, the amount of applications not viable now due to cost/latency/throughput is vastly bigger than the amount of current use cases. Even current ones do benefit, e.g. it's a great executor subagent.

Energy and intelligence are good too, sure.

surprisetalk•41m ago
This reminds me of those pareto-style speedrun record charts when a new glitch is discovered.

[0] https://taylor.town/silver-landmines

When I see dramatic leaps like this, it tells me that the important hacks haven't yet been discovered.

clayhacks•41m ago
Why wasn’t this run against ARC-AGI-3? Or did it fail to solve anything?
esafak•41m ago
It's serviceable but, like many Chinese models, it uses a lot of tokens to get work done.
Havoc•36m ago
They did recently announce they're increasing prices though (got a mail yesterday I think), so not sure this analysis showing it as price outlier will last
minimaxir•34m ago
That is only when using the DeepSeek API directly. OpenRouter has 24 different providers serving it at existing prices.
dcchambers•34m ago
This latest DeepSeek is almost at the "too cheap to meter" level. That's going to be a larger unlock than models like Fable/Mythos that are way too expensive to justify, IMO.

What secret sauce do they have?

throwaway_95283•13m ago
limited resources, no modern GPUs, no $10 billion dev budgets.

pair it with codewhale, 50 agents, 200 MB of ram.

gentlewater•31m ago
I’ve been refreshing hacker news constantly for a week now waiting for v4 pro, after they stated it would follow «soon». I have learnt «soon» is a matter of definition.
mosura•17m ago
I strongly recommend trying this for programming tasks.

It is strong (not Fable strong though) with a much better “persona” than Opus, and very different blindspots. If you flip between Claude and this you will find both catch the mistakes of the other before they get out of control.

On balance I actually prefer DeepSeek for programming now, because of the way it talks.

dolebirchwood•37m ago
If you're on the DeepSeek Platform, you'd see this:

"We plan to raise the overall pricing for DeepSeek API services in the near future, with a significant increase expected. Please plan your usage accordingly. The specific pricing plan will be subject to official notice."

LaurensBER•38m ago
Dax (from Opencode) has tweeted that they can replicate or beat the price with rented GPUs. Deepseeks secret sauce is the incredibly cheap caching (magnitude cheaper than other providers).

vLLM has recently released a similar approach. It's not as effective as what DeepSeek does but still an interesting development.

I have no doubt that in due time other providers will match or perhaps even beat the current DeepSeek prices.

NorwegianDude•31m ago
Eh, what are you guys even talking about? Deepseek is not cheapest provider as is, and it's MIT. So deepseek making it more expensive to use is just nonsense, they can only change their own pricing. It's the beauty of MIT license and open weights. If anything, these models are some of the safest in the world to use if you worry about a rug pull.
akman•24m ago
90%+ cache hit rate is common, and so you'll see on places like openrouter that Deepseek cache cost is indeed a magnitude cheaper than the rest.
LaurensBER•17m ago
There's more to inference than just the input/output token cost. Caching has a massive impact.

Deepseek charges $0.0028 per cache read on Openrouter. The next cheapest is $0.018.

That's a massive difference and quickly adds up on coding sessions (which often hit 95%+ cached tokens).

retinaros•23m ago
any link to this caching tech?
LaurensBER•10m ago
[Feat][Core] Add disk offloading support to SimpleCPUOffloadConnector — #49644 https://github.com/vllm-project/vllm/pull/49644

This adds disk as a tier in the HBM → CPU → Disk KV cache hierarchy.

There's also a cluster of related KV-offload FS PRs: #49225 (read/write batching, still open) and #49152 (batch store/load in C, merged Jul 28).

It's hard to say if these are similar to the approach DeepSeek takes but they definitely seem very interesting.

minraws•15m ago
As someone who recently tried it on some blackwell cards, it's possible to match the prices especially the input can be even cheaper and output can match the costs so you can easily build a net 20-30% margin business even at current GPU prices.

The entire issue is caching, I tried to write some custom to dump to disk kv-caching using some ideas from their papers and my experience with snapshots and vm checkpoint systems, I must say they must have really squeezed that lemon it's hard.

Atleast me with Sol couldn't figure it out over a couple days, a few hours each day, which isn't much but I did feel a bit stuck with existing solutions and felt like I might have to write something from scratch. But if you are willing to put in the effort into the infra I do think it's doable. But it will be really hard to pull it off.

My congrats to anyone who manages to pull it off, they might be able to kill off most AI labs. Assuming they can find the compute, Deepseek really has killed all models for me other than Sol/Fable/Opus/K3 tier stuff.

onlyrealcuzzo•10m ago
> Deepseeks secret sauce is the incredibly cheap caching (magnitude cheaper than other providers).

Can anyone working at one of the main US labs (Google, OpenAI, Anthropic) comment on WTF they haven't even tried MLA - despite the obvious massive advantages?

I know enough to know they aren't completely incompetent. So there must be a quite good reason.

But it remains a mystery to me.

DeepSeek's MLA is like almost 2 years old at this time. They've got thousands of people working on this stuff. They clearly have the ability to at least try it...

ms8•30m ago
Yes, there is warning, but also there are many providers on OpenRouter[0], hosting open weight model with similar pricing. The question is Will they go up as well?

[0] https://openrouter.ai/deepseek/deepseek-v4-flash-0731#provid...

HSO•13m ago
even if they double it it`s from such a low base it is still supercheap
eli•5m ago
I assume/hope this is about prices going up for the next release of Pro
anramon•38m ago
>even if it's not SOTA

And, probably 99.99% of people using LLM probably don't even need SOTA anyway.

swiftcoder•8m ago
At least on these benchmarks, it seems to be pretty handily scoring up with the SOTA from 6 months ago?