frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

DeepSeek costs OpenCode Go user $1.14/day; dual DGX breaks even in 24 years

https://twitter.com/thdxr/status/2086599224674681242
35•delduca•1h ago

Comments

jauntywundrkind•1h ago
That's almost half way to their stated 6X usage goal! Just from DeepSeek.

> With Go, you pay $10/month and we aim to give you 6x that in usage.

For most models, we make this work through bulk discounts and reserved GPU capacity. We then pass those savings on to you through the 6x multiplier.

https://opencode.ai/docs/go/#why-some-models-have-lower-usag...

I think it's back to 2X usage, meaning it's cheaper token-burn than usual to use. Which is lovely.

OpenCode Go has been so nice to have. I love having access to DeepSeek, Qwen and MiniMax M3 when doing design work, to see what different models cook up. I've been very surprised with MiniMax M3, not as a particularly good architect, but at it's very good ability to state the problem elegantly & to frame the different decision points very well. That's been a fun ongoing surprise.

delichon•47m ago
It'll be great if people start to seek out diversification in their model portfolio as in their financial portfolio, and model builders can make bank by producing idiosyncratic oddballs that think different. It could be a force against a uniform bubble of models trained to the same social guardrails, until banned for wrongthink.
notjes•21m ago
Since the announcement of DeepSeek price hike, I have been using MiniMax M3 and I am really surprised by the quality this model spits out. Subjectively speaking, the bullshit MM3M produces is waaay less than DS0731. Could it be a result of less hallucinations than DS0731? https://artificialanalysis.ai/evaluations/omniscience#aa-omn...

Maybe hallucination is good for prototyping and creative work, but maintaining and debugging code might be better done by a boring model?

nchmy•18m ago
but DS prices havent increased yet... in fact, opencode go is offering an extra 2x bonus on DS Flash right now. Why switch before anything has changed?
infecto•20m ago
Don’t forgot a number of models they have offered to you will train on your data
spwa4•1h ago
Deepseek will run fine on a single m5 max, cutting that 24 years in half.
PunchyHamster•28m ago
also worth remembering API price includes power, just dividing price of the hardware by usage doesn't
epolanski•24m ago
The post is also misleading because the m5 max 128 gb costs a lot and you can only run DS4 flash quantized, thus not a 1:1 comparison.
echelon•53m ago
Cloud > Local

I have a stack of ten or so 3090s sitting in boxes, but it's not worth the hassle to use them. You can easily run models as cheap as water in the cloud.

Sitting around 15 minutes for local Minimax is stupid when you're trying to be productive. You can spin up parallel job instances and multitask in the cloud.

If you want freedom, build open source cloud infra.

You rent your ISP line. Why isn't renting GPU compute seen the same way? You still have compete ownership over your stack, you're just letting someone else deal with the capital outlay and headache.

knicholes•47m ago
Why not rent those out on vast.ai and make a little money each month?
TacticalCoder•39m ago
> You rent your ISP line. Why isn't renting GPU compute seen the same way?

Is it mostly seen the same way. What is unacceptable is removing the freedom (which you mentioned) of people who prefer to run models locally.

Just as people have the right to tinker at home with DIY and electronics, fully knowing they won't compete with the latest ASML machine, people are free to use open-weight models at home.

BTW Minimax H3 running local can generate amazing short vids quite fast: about 50 seconds to generate a 7 seconds vids on a 4090 (depending on the settings). I've got a friend who spams my Telegram daily with such (no censorship and NSFW btw) vids.

I don't run local but I'll defend the rights of people who want the freedom to do so and I won't look down at them from my high-horse talking about "electricity" and "productivity".

trouve_search•38m ago
The value prop really depends on what you're doing.

If you're just vibe coding with giant frontier models, yes, the value will be worse. Especially now, where GPU prices have spiked another 20% last month.

For some tasks where owning the setup and full kv cache matters, the payoff calculation is ridiculously in favor of running your own deployment.

For instance for some batch classifications jobs where the prefix cache hit rate will be >95%.

The calculus also changes if you just use AI as a light tool while coding and don't need the giant models; qwen3 27B runs at 80TPS on a 5090 properly deployed.

bewareofscams•45m ago
Beware of the author of tweet, who happens to be author of OpenCode - OpenCode will leak all your data to themselves and to shady 3rd parties. Author feigned ignorance and never fixed the issue. OpenCode among other harnesses is the shadiest of all.

https://github.com/anomalyco/opencode/issues/10416

tokai•38m ago
Thats not that big of an issue. Seeing a lot of accusations against opencode, regarding all kind of things, here on HN lately.
bewareofscams•36m ago
Anything of substance besides shallow dismissal and (anti)appeal to masses?
lukewarm707•23m ago
agreed. when using local models, they did send your prompts to openai with 30 day retention to make the titles, silently.

their recent changes to the privacy policy broke their promise of zero data retention. specifically, they offered chatgpt luna under a zero data retention privacy policy. luna was later shown to be 30 days retention.

their privacy policy has never guaranteed your prompts will not be logged and when asked they have failed to revise it.

when challenged about sending data to openrouter without listing it as a 3rd party processor they offered a dismissive response. the same with running prompts through cloudflare. seems trivial, but signifies general disinterest in security.

by default if you run the harness outside your config file by accident, it will automatically run silently with a free model that sends your prompts and local data to an endpoint with training enabled. on top of that it used to dump all the prompts sent to free models into an s3 bucket, the feature was literally called 'datadumper' in the source.

perarneng•34m ago
The reason to run local models is not for coding mostly it's for learning how to deploy models and tinker with self hosting. It's also for massively crunching data 24/7. Imaging having an agent analyzing constinous log streams etc .. that could be a usescse where even deepseek could add up cost.
epolanski•26m ago
Also privacy and offline capabilities.
DanielHB•8m ago
One thing I realized is just how much offline local models can hurt mass data collection.

For example, I needed to write an invitation letter for immigration control for a relative visiting me. Previously I would have used a search engine for a template. Today I fire up my local qwen 3.5-9b for this kind of stuff and feed it all the private data I need.

Unfortunately it is unlikely the average user will known how to avoid this data collection. Even if the LLM is local you are likely feeding the prompts to remote servers if you harness/chat-interface is not properly vetted.

pu_pe•19m ago
That's only true if the value of keeping your data and code private is zero. And in that case, Anthropic and OpenAI subscription plans may be even cheaper per day.
monster_truck•10m ago
I'd really rather just pay Deepseek directly. Why wouldn't I want to support the company that trained the model?

It isn't even really worth the (minimal) ops to stand up rented MI300Xs to sell excess capacity to them even if it was minimally profitable, when I tried I was content to give API keys to friends to beat on it.

gessha•29m ago
Why are you hoarding the 3090s T.T They've jumped from $700 to almost double on eBay.
fancyfredbot•29m ago
Why do you have ten or so 3090s sitting in boxes?

I mean obviously it's worth it just so you can flex on HN. But curious whether there was any other reason? Retired scalper?

hluska•12m ago
Do you really keep ten cards in boxes just to brag on Hacker News? Geez, our industry has gotten pathetic.
infecto•21m ago
Plus 100 to this. Of all the harnesses I find it to be the worst. IMO training should always require opt in and the way they continue to run their business/framework is shady.
infecto•22m ago
Probably less of an issue but I always disliked with their paid plan/credits that they made I nonobvious that some of the Chinese models train off of your usage. It may have changed now but they would show all these great models you could use, somewhere have a bullet that everything is private adobe have an asterisk next to a handful of those models (including their own). I am sure some cost sensitive folks are ok with that but I disliked how there was not an easy way to tell and it was opt in automatically if you used those models.
7373737373•20m ago
It's not even possible to delete one's account!
tonyhart7•17m ago
I guess that's why they called OpenCode
polski-g•4m ago
I like how everything you said is false.

OpenCode does not leak "all your data". Author did not "feign ignorance" and there is no "issue to be fixed."

Title generation is handled by the small_model setting in the config. If you have that configured to use a local LLM, then that is where the your prompt will be sent to generate a session title.

Meta Muse Glimmer – open weights 30B local coding model

https://research.meta.ai/blog/introducing-muse-glimmer-open-agentic-model
370•riordan•3h ago•168 comments

Mistral Patent for "Code implemented tool calls"

https://patentsgazette.uspto.gov/week26/OG/html/1547-5/US12670045-20260630.html
29•theanonymousone•22m ago•21 comments

50k Boat Names

https://www.beautifulpublicdata.com/boat-names/
27•jonathanmkeegan•53m ago•14 comments

Squeak/Smalltalk 6.1 Release Notes

https://squeak.org/release_notes/6.1/
33•fniephaus•1h ago•7 comments

Docker Sandboxes – Disposable, isolated sandboxes for AI agents

https://www.docker.com/products/docker-sandboxes/
353•etoxin•7h ago•219 comments

Tail-call optimization in C is relatively recent

https://lwn.net/Articles/1034703/
48•prakashqwerty•2h ago•19 comments

Parametron: 50s Japanese computer that uses neither transistors nor vacuum tubes

https://ethw.org/Milestones:Parametron,_1954
41•xeonmc•3h ago•11 comments

Over 181,000 AI meeting recordings left wide open in note taking app

https://bobdahacker.com/blog/tldv-hack
80•colesantiago•1h ago•26 comments

What Happened to HackerOne?

https://blog.teknogeek.io/posts/what-happened-to-hackerone/
297•hipparchus•11h ago•153 comments

DeepSeek costs OpenCode Go user $1.14/day; dual DGX breaks even in 24 years

https://twitter.com/thdxr/status/2086599224674681242
36•delduca•1h ago•33 comments

Run Android ARM64 VR APKs on Apple Vision Pro

https://github.com/shinyquagsire23/Klepton
127•LorenDB•10h ago•23 comments

An Interesting Fourier Transform – 1/F Noise

https://www.dsprelated.com/showarticle/40.php
75•q7m•3d ago•16 comments

Show HN: Voice driven murder mystery, Interview AI suspects with your voice

https://www.whodunnitai.com/
134•MrRowTheBoat•10h ago•54 comments

Why a Raspberry Pi shouldn't be powered through its GPIO pins

https://82mhz.net/posts/2026/08/why-a-raspberry-pi-shouldn-t-be-powered-through-it-s-gpio-pins/
11•speckx•4d ago•5 comments

Tail-Call Interpreters in Rust – Jimmy Ostler

https://lordgoati.us/blog/tail-call/
52•amatheus•3d ago•17 comments

How Blackwing Pencils are Made [video]

https://www.youtube.com/watch?v=fow-LsdaH2E
40•NaOH•4d ago•16 comments

Findphone: Locate a nearby Bluetooth device by signal strength

https://github.com/ben-z/findphone
12•helsinkiandrew•5d ago•8 comments

Taxi drivers rarely die of Alzheimer's

https://theconversation.com/taxi-drivers-rarely-die-of-alzheimers-how-complex-mental-maps-and-spa...
333•jader201•22h ago•235 comments

How I use LLMs to learn complex topics

https://laurentiugabriel.github.io/blog/articles/how-i-use-llms-to-learn/
721•laurentiurad•18h ago•463 comments

Defending my own brain against enshittification

https://mrmarket.lol/how-i-feel-calmin-control-of-my-life-in-the-time-of-enshittification/
9•mrmarket•31m ago•0 comments

How We Pushed CDC into Postgres

https://www.snowflake.com/en/blog/engineering/postgres-to-snowflake-replication-mirroring/
114•craigkerstiens•12h ago•22 comments

Ask HN: What are you working on? (August 2026)

277•david927•20h ago•958 comments

Because It's Not Fun Enough: why languages fail

https://bytecode.news/posts/2026/08/because-it-s-not-fun-enough
66•jottinger•2h ago•68 comments

Cool URIs Don't Change (1998)

https://www.w3.org/Provider/Style/URI
261•Klaster_1•23h ago•63 comments

ATProto for Distributed Systems Engineers

https://atproto.com/articles/atproto-for-distsys-engineers
109•LelouBil•3d ago•21 comments

Picophysics: Single file physics for games on platforms like N64, PSX, DC

https://gitlab.com/Kazade/picophysics
72•klaussilveira•5d ago•22 comments

An alias-based formulation of the borrow checker (2018)

https://smallcultfollowing.com/babysteps/blog/2018/04/27/an-alias-based-formulation-of-the-borrow...
18•parksb•2d ago•1 comments

Tuxedo No. 2 – Cocktail recipes

https://tuxedono2.com
120•smartmic•17h ago•35 comments

Everything you do is being recorded

https://www.theatlantic.com/technology/2026/05/ai-wearable-surveillance-countermeasures/687203/
350•ike_usawa•1d ago•310 comments

Nearest Pint

https://knowwhereconsulting.co.uk/maps/pubs/
43•bookofjoe•5d ago•25 comments