frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Whistle: Speech to Text in 16.9 MB

https://cactuscompute.com/blog/whistle
288•gmays•3h ago•73 comments

The value of not getting to the point (2015)

https://ken.arneson.name/2015/11/the-value-of-not-getting-to-the-point/
40•NaOH•1h ago•10 comments

Theranos.World

https://www.theranos.world/
42•kbyatnal•2h ago•13 comments

Show HN: Making a flexible "neon" t-shirt with LED filaments

http://scottbezek.blogspot.com/2026/10/making-flexible-neon-t-shirt-with-leds.html
32•scottbez1•3h ago•7 comments

Show HN: K10s – A Clickable Kubernetes TUI (Go, Bubble Tea)

https://github.com/p10node/k10s
48•pierreneter•1h ago•30 comments

Why isn't the industry freaking out about DeepSeek 4.1 Flash?

https://www.dgt.is/blog/2026-10-07-deepseek-freek-out/
86•jonotime•20h ago•68 comments

A Terminal Protocol for Program Status (OSC 7501)

https://mitchellh.com/writing/program-status-osc7501
24•mfiguiere•1d ago•3 comments

Step 5 Preview, a 1M-context MoE from StepFun, shows up on OpenRouter

https://openrouter.ai/stepfun/step-5-preview
56•AnneWodell•4h ago•20 comments

Beauty in DVD Menus

https://vale.rocks/posts/dvd-menus
208•speckx•7h ago•131 comments

Man discovers his parents' coffee machine used 1TB of data in 10 days

https://www.dexerto.com/entertainment/man-discovers-his-parents-coffee-machine-used-1tb-of-data-i...
135•ck2•1d ago•65 comments

OpenAI annualised revenues $20B less than previously signalled

https://www.cnbc.com/2026/10/08/open-ai-revenue-nvidia-oracle-coreweave.html
268•mfiguiere•3h ago•164 comments

The dawn of the age of the exoskeleton

https://theconversation.com/the-dawn-of-the-age-of-the-exoskeleton-281804
21•speckx•1d ago•3 comments

I hired an illustrator to draw my house. Now it's my Home Assistant dashboard

https://antonfrolov.substack.com/p/i-hired-an-illustrator-to-draw-my
121•soheilpro•1d ago•8 comments

ETH-68: Ethernet Audio Interface for Linux

https://naturalsystems.io/eth68
23•chabad360•1d ago•4 comments

A 5.3M-year-old deep-sea whale necropolis in the Diamantina Zone

https://www.nature.com/articles/s41586-026-10546-z
29•bryanrasmussen•1d ago•0 comments

“Math 2.0” will need to value mathematical progress more holistically

https://mathstodon.xyz/@tao/117395269325940185
577•ent101•15h ago•598 comments

DuckDB Ducklake

https://github.com/duckdb/ducklake
63•saikatsg•1d ago•5 comments

OLED burn-in test: 30-month update

https://www.techspot.com/article/3178-oled-burn-in-test/
45•baal80spam•9h ago•24 comments

Trump administration is suspending Microsoft from a green card program

https://apnews.com/article/h1b-visa-program-vance-microsoft-e7b3a407f822702b269ee277d21343ea
700•alephnerd•5h ago•1196 comments

Archaeologists Are Reconstructing the 'Invisible' Technologies of the Stone Age

https://www.smithsonianmag.com/science-nature/archaeologists-are-reconstructing-the-invisible-tec...
78•Hooke•23h ago•39 comments

I gave Opus 5.5 one prompt and six hours to visualize Invisible Cities

https://quesma.com/blog/invisible-cities-one-shot/
308•stared•8h ago•159 comments

License update: AI derivation prohibited on all my art, lore, stories, comics

https://www.davidrevoy.com/article1178/license-update-ai-derivation-prohibited-on-all-my-art-lore...
27•frizlab•44m ago•12 comments

Show HN: I Put an AI Agent on a Nokia 110

https://github.com/anupray95/AI-Agent-on-a-NOKIA
20•anupray•6h ago•5 comments

Show HN: TerrainSR – fast, realistic heightmap upscaling model

https://huggingface.co/joe-gibbs/terrainsr
11•joegibbs•1d ago•2 comments

Vitalik Buterin backs crypto ‘bunker mode’ amid rapid AI math advances

https://cointelegraph.com/news/justin-drake-urges-crypto-bunker-mode-as-ai-could-break-wallet-sec...
32•firstcomm•1h ago•14 comments

Show HN: Rgpu – a PyTorch device whose tensors live on a remote GPU

https://github.com/ymcrcat/rgpu
12•boxstream•1d ago•1 comments

Yes, and

https://htmx.org/essays/yes-and/
12•Michelangelo11•10h ago•3 comments

Tell HN: I've been paying for a rural Tanzanian's education for 10 years

599•lukehandcool•5h ago•167 comments

New CRAM method offers giant boost to compressed memory reads

https://www.tomshardware.com/software/linux/new-linux-tech-compresses-memory-in-ram-as-ram-for-45...
37•danny00•7h ago•16 comments

Orkut.com

https://orkut.com/
99•andreynering•6h ago•71 comments
Open in hackernews

Why isn't the industry freaking out about DeepSeek 4.1 Flash?

https://www.dgt.is/blog/2026-10-07-deepseek-freek-out/
82•jonotime•20h ago

Comments

verdverm•18h ago
Why would we freak out? The systems we use have always gotten better, faster, cheaper with time
giancarlostoro•45m ago
Call me crazy but:

VRAM & Memory Requirements by Precision

• FP16 (Full Precision): Requires ~1,664 GB of VRAM (e.g., an 8x B300 288GB cluster).

• INT8 Quantization: Requires ~832 GB of VRAM (e.g., 8x H200 141GB).

• INT4 Quantization: Requires ~416 GB of VRAM (e.g., 8x A100 80GB)

VRAM aint cheap, Sam Altman ruined the cost of memory, Nvidia doesnt make enough consumer GPUs letting the market go insane over them, I still have friends on 1070s or 1070 TIs because GPUs have been severely overpriced for too long. I remember when a gaming PC was only $1000.

Even so why would anyone not sleep on a model they cannot run?

kristopolous•33m ago
Seriously, if a single politician stepped forward and said "i'll bring down ram prices" they could then shoot a puppy and call me a slur and I'd still go out and doorknock for them.

Memory companies have price fixed multiple times. They've paid hundreds of millions in fines. wikipedia even has a page on it. https://en.wikipedia.org/wiki/DRAM_industry_price_fixing.

Look at the financials of these companies, they're all making obscene margins and do they plan to increase production? No. Micron is doing a stock buy back to pump the price of their share.

The Micron CEO just recently said this is the exact plan https://www.theregister.com/systems/2026/10/01/ram-supply-se...

There's sanctions, tarrifs, and a DOJ who doesn't give a shit. Until we can fix that the insanity will continue. Phones will be unaffordable. Laptops will be obscene. Gaming consoles will be thousands of dollars. Desktops will be dead.

If you're waiting for some David Ricardo equation to happen, tough cookies, it's not coming.

The market is legally locked down and we're in hostage pricing mode.

And what's the story? You can't afford electronics because we're using it to build robots to take your job? I mean ...

Nobody is coming to save us. That's our job.

gchamonlive•27m ago
You won't see it because this politician wouldn't survive one rally before getting shot, and even so it would be hard to bring prices down. There would have to be major investigation into ram vendors forming cartel in order to break them, and major investments to increase production, because this is also a supply shortage fulled inflation
Analemma_•12m ago
I don't think RAM vendors have formed a cartel and I think this is knee-jerk anger without any thought. RAM is a commodity product with massive upfront capex costs, and those always have boom-and-bust cycles. At various points in the 2010s and 2020s RAM vendors were getting eaten alive by a supply glut, this would not have happened if they were a cartel.

Is it really so hard to believe that RAM prices are up because demand is simply exceeding supply, especially in a market where additional supply takes years and billions of dollars to come online? There's no need to posit cartel behavior and a fair amount of evidence that there is none.

m463•14m ago
> "i'll bring down ram prices"

wonder what voting would be like?

gamer vote ++

datacenter hater vote --

datacenter lobby ++

micron lobby --

mwambua•9m ago
Wouldn’t cheaper memory make it easier to bring compute out of data centers and onto consumer hardware?
phil21•10m ago
> do they plan to increase production? No.

Micron has 3 brand new fabs currently under construction, 2 Boise, 1 in New York as the first of 4 planned for a campus.

Plus expanding other existing facilities.

These things take ~3-5 years from breaking ground to full production. You'd have had to anticipate the current demand years before it happened in order to be bringing production on-line before 2030 or so.

Samsung and HK Hynix also have fabs under construction and planned.

CXMT started 11 years ago and only now is reaching any real volume. If they decided a year ago to react to the current demand cycle they'd be 6-7 years out.

Not much you can really do to wish for more fabrication to exist on any timeline not measured in fractional decades.

Could they do more and react quicker? Probably, but everything I've read on the subject seems to point to 3 years is absolute bare minimum if you happen to have a shovel ready project with the land bought, local permitting completed, infrastructure extended to the site, and a skilled workforce already in place. They could suspend buy-backs/dividends today and dump it all into building production and there would be no material impact until around 2030.

> The Micron CEO just recently said this is the exact plan

CEO simply stated the demand pressure will not go away through 2027, and supply will not increase until around 2028 when currently under construction fabs start shipping volume. The article does not support your statement.

ByteAtATime•29m ago
I think, considering the size of this model, it's closer to a Pro than a Flash on everything other than speed
holoduke•29m ago
He doesn't ruin the cost of memory. Advances in memory size and speed are now in full speed mode. Expect drastic increase in the upcoming years. Big factories are in the making and planned. Gigalab in the US and many others in the east. Since 2010 we have computers with 16gb as being normal. Finally we are moving into a new era where the standard will be 64gb next year and 128 in 2028. Hopefully we reach 1tb in 2030.
CorrectHorseBat•20m ago
I've read the exact opposite, vendors are reducing the standard from 16GB back to 8GB
functionmouse•18m ago
one can make a fine gaming pc for ~$350

1660 ti, 4790k, 16gb ddr3

petu•16m ago
There's no BF16, original full quality weights are quantized already and 510GB.

Then good portion of those weights are n-grams (~200GB) that don't need to be in VRAM.

Then KV cache of that model is super lightweight at ~1GB per 1M tokens. If HBF succeeds, then accelerator with 16GB of VRAM and 1TB HBF/NAND is probably all you need (?).

liuliu•12m ago
What are you talking about? The model is native NVFP4, why you run it at any precision higher than that?
anvuong•9m ago
I just un-retire my pair of 1080Ti for some small models development because the current GPU prices literally make me sad.
nullc•6m ago
The bulk of its weights are natively MXFP4. And engram values don't need to be in vram.
booi•45m ago
Because GLM 5.3 Flash is even cheaper?
wg0•43m ago
Don't think so.
0123456789ABCDE•33m ago
here you go: https://artificialanalysis.ai/models/releases/comparisons?co...
ActionHank•29m ago
Nah fam, not true, also DS edges it out on coding / dev tasks.
jacquesm•28m ago
That is opposite to my experience so far, can you describe your coding tasks? Mine are systems level code, utilities, operating system code, networking and real time control stuff.
sampullman•14m ago
On design tasks too, for me.
f311a•4m ago
tengbretson•44m ago
I don't know about "freaking out", but I'd say I'm having a good time here with DS 4.1 flash.
smallmancontrov•41m ago
They might be. They would delay public admission as long as possible, because public admission would make stocks go down.
browningstreet•40m ago
What would freaking out look like, or is this just a stupid bloggish title flourish?

Is OpenAI coming in $20B under a sign of "freaking out"?

efficax•36m ago
they should be freaking out because every time the chinese labs or non "frontier" labs release a model that is only a few months behind and much cheaper than the openai/anthropic models it shows that they don't deserve their valuations
jerf•22m ago
It would look like major chaos in the markets.

People tend to conflate the question "is AI a useful technology?" with "are the AI companies going to do well?" but they're surprisingly separated in practice, with either one able to be true while the other is false. There is a lot of money tied up in a lot of hardware with a lot of loans made against that hardware as collateral all based on the assumption that AIs are going to need more and more and more and more hardware and whoever has the hardware wins. If a much better model comes out that requires vastly less hardware, or even more accurately, merely charges vastly less than the current AI companies, then to a first approximation (barring Jevon's paradox, and bearing in mind there's no timeline guarantee on that) all that hardware becomes much less valuable for being grotesquely oversupplied relative to what is necessary, and even though that would generally make AI objectively more useful than it was before, it would cause mass financial chaos in the markets.

The markets need a very particular rate of progress. It isn't entirely clear to me that it's even a possible rate of progress, it may be overconstrained, but they certainly don't have plans for the AI models to get commoditized on the timeframes of these vast, vast array of loans being made against hardware as collateral. Spend a metric shit ton of money to kill all your competition then charge monopoly rent on the one thing absolutely everyone needs doesn't work if you can't economically "kill all your competition" because the economics favor them in the spending spree.

And then, based on the fact that this is not even remotely complicated logic, there are plenty of people who are fully aware that they have a lot of money tied up in not running around telling everyone how wonderful the cheap models have become.

anguralbanish2•36m ago
I would love to get them more better, it's good not a bad thing.
wg0•35m ago
While using DeepSeek v4.1 Flash I was architecting a system and I made a mistake of drawing the RPC boundaries at a wrong place that did cost me in so many ways.

I realized that mistake and guided DeepSeek where it should be.

Next I fired Fabble 5.5 set to high to check if the hype is real about Fabble. It exhausted 89% of quota and came up with NOTHING that DeepSeek hadn't flagged itself already in its notes.

sampullman•15m ago
Do you mean Fable 5.1? Or Opus 5.5? I'm not sure what you're working on but for me DS 4.1 flash isn't nearly at their level. For the price it's obvious very impressive, though Luna 6.0 is excellent too.
hirako2000•7m ago
The problem with benchmarks and proprietary models is that one day a model is best at doing X, another day that's not so sure. And anyway, we are throwing the same X.

I've found supposedly smaller and, less performant models do better on certain tasks. I end up using several models, sticking to what my unconscious statistical observations tell me to use for the kind of task at end.

simpaticoder•35m ago
The question seems rhetorical but I think there are two reasons in some combination. First is it there is some awareness lag here. That lag can be on the producer and consumer side. Software enterprises are pretty slow to adopt new things and slow to try new things so they might only be aware of openai and Claude as options. Plus there are some scariness because deep seek is a Chinese model and therefore export restricted - never mind that there are American in European providers.

The other reason is more interesting. Maybe the frontier providers think that price performance is irrelevant in light of very powerful frontier models that can start the RSI loop and or a huge displacement of work and a winner take all economic situation. After all if frontier providers earn everyone's money then you won't have any money to spend on any model 100x cheaper or not.

agoodusername63•19m ago
I think it also has a bit to do with the AI sector of tech still moving at lightning speed.

Theres already models that outdo DS 4.1 flash in cost/performance. Luna 6 on max effort for example. Luna also doesn't care what time of the day it is for cost calculation.

And I'm sure by the time people ask why Luna 6 is being slept on there will be another cost/performance king

thefourthchime•34m ago
For non-coding tasks it may be fine. But for coding, Opus 5.5 is just a completely another level than something like Deepseek 4.1 Flash.

Opus 5.5: TIME 9.3m COST / $1.99 / SCORE 99/100 https://jonclegg.github.io/pacman-bakeoff/#claude-opus-5-5

Deepseek 4.1 Flash: TIME 2.8m / COST $1.89 / SCORE 72/100 https://jonclegg.github.io/pacman-bakeoff/dev/#deepseek-v4.1...

AIblemblio•33m ago
No they can't.

And as long as I pay as little for claude opus 5.5 i do right now, i'm using it.

But yes i'm glad that we have alternatives.

hmontazeri•33m ago
I had the same experience using ds 4.1 last couple of weeks. It’s insanely good for the price. I’m doing mostly web dev it excels at everything I throw at it. The pricing is ridiculous. I canceled my gpt subscription and haven’t looked back hope the pricing stays like that. I almost never need a better model. I still keep my Claude 20$ sub for now but I feel like one more iteration and I won’t need even that anymore I hardly use it
jacquesm•29m ago
If DS4.1 impresses you I would be really interested to see your comparison to GLM 5.3. I switched from the one to the other and even if GLM 5.3 is a bit slower I don't think I'll be going back.
ctolsen•21m ago
GLM 5.3 is very impressive and definitely better, but it also at least 4x the price.

On that note I’ve been subbing in MiMo-2.6-pro when cost is an issue, which is super cheap and also performing really well.

pdhborges•4m ago
What inference provider are you using?
swiftcoder•31m ago
I think the interesting provider to cross-check this assumption with here is Meta, who is clearly freaking out, and is currently providing Muse 1.3 even cheaper so long as you are willing to share data with them
hypfer•28m ago
Is it known why unsloth seems to not have touched DeepSeek 4.1 Flash?
jacquesm•24m ago
You can ask them directly, Daniel Han-Chen is pretty responsive.
pianopatrick•24m ago
I was just using a bunch of models in Cursor to review a project. I went looking for DeepSeek and it was not one of the options.

Would be cool if they added it.

hirako2000•5m ago
Since they adhere to the same API spec, you can hook any model. It takes one line edit in /etc/hosts

There are some quirks if your harness use unsupported features of course.

lmf4lol•22m ago
Oh man. v4.1-flash has been an sbolute game changer for us. We run all our Personal Assistants now on flash (thinking high) by default and it works incredibly well. There is really no need for basic agentic tasks that might require Kimi K.3 or GLM-5.3 levels.

Once its gets juicier, we let flash launch specialized subagents with specific models. GLM-5.3 for coding or Kimi K.3 for research and critique.

But as a main driver. I love flash. And it brought our bill down by A LOT :D

aftbit•9m ago
Have you compared it against actual SOTA models like latest Fable or Astra?
sneurlax•2m ago
Of course there's still a huge performance gap

but DS 4.1 Flash is good enough for most tasks

sroussey•20m ago
Not comparing to gpt-6-luna which seems comparable and priced well.
gsky•17m ago
America bans Chinese models sooner or later just the China banned American big tech
kristianp•16m ago
> shrank the KV cache by roughly 437X

Can't you just say "shrank to 1/437th the size"? It's not that hard.

aguilaair•15m ago
What about MiMo v2.6 Pro? It’s throughput is slower by default (UltraSpeed is faster than DS4.1F) but is above the pareto line, and cheaper.

see https://artificialanalysis.ai/models/releases/comparisons?co...

MisterMunchkin•13m ago
I had it make 25 different things today and it cost $0.70

It’s disgustingly good value. I find it capable of doing anything I want.

Obviously can’t use it at work, but for home projects it’s awesome.

doctorpangloss•11m ago
because it doesn't work very well?

if you have a legitimate coding application, it isn't very good. if you have some kind of inauthentic activity, which could be what it is trained for for all sorts of reasons...

liuliu•11m ago
DeepSeek 4.1 Flash 0910 is perfect for M5 Ultra 256GiB. Running it fully resident in RAM, prefill at ~2500 tok/s and decode at ~40 tok/s. Probably tons of room to improve from there.
sdg03ksdv0d•1m ago
You are running this now? :o
distantsounds•9m ago
because we've all figured out that AI is just a huge grift?
wewewedxfgdf•7m ago
You might also choose to pay money for a service that provides real value instead of actively choosing to support the Chinese deliberate effort to undermine this country.
elmer2•6m ago
DeepSeek isn't even on my mind. I use the frontier models and can get the best in the industry for a relatively cheap price.
RGS1811•3m ago
This model finally got me off my Claude Max subscription. I’ve found it superior to Opus 5.5 in certain use cases, and certainly faster.

I’m convinced that I’ll have good enough inference on my laptop at reasonable speeds within the next year.

sergiotapia•3m ago
In my experience it just takes so much longer to arrive at "done" state for me. It thinks for soooooo long. I guess if you're running 12 sessions at once you don't really notice.
xyzsparetimexyz•2m ago
There was a moment 3 months back where the sentiment was that cheaper models were the way to go. Since then the pendulum has swung back.
mlinsey•2m ago
I'm paying for the heavily-discounted subscriptions, not the API rates. There isn't really a cost gap for me. DeepSeek doesn't have a subscription to compare to, but when I compared the GLM 5.3 usage I got from a $100/mo Z.ai subscription compared to Opus 5.5 on a $100/mo Claude subscription, there wasn't a big gap. And GLM 5.3 is very clearly not a frontier model (deepseek v4 seemed a lot closer, but I didn't use it enough to really say for my workloads).

I don't think those subscriptions nave negative contribution margins, either. I think we're seeing a lot of price discrimination by the big labs, and huge margins on their frontier models. The fact that they have been cutting prices to their second-biggest tier of models (Opus/Sol).

Open models catching up and collapsing these margins would worry me if I were a shareholder in the big labs, but as a user, I really doubt that the western labs have bigger environmental impact just because they have higher API costs, I think they have a ton of efficiencies they aren't sharing with customers yet because demand is so high.

xyzsparetimexyz•8m ago
Neither political party cares at all about memory pieces get real lol
ashdksnndck•3m ago
RAM manufacturers are bidding against NVIDIA and everyone else for the same constrained supply of EUV machines. And it takes years to build more fabs. Micron has multiple fabs coming online in 2027 and 2028.
Opencode Go gives only 6300 requests for glm and 23 000 for deepseek. And, if I wanted to, I would be able to do all my work on $10 plan with deepseek. It’s very cheap.
hirako2000•2m ago
[delayed]