frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

DeepSeek costs OpenCode Go user $1.14/day; dual DGX breaks even in 24 years

https://twitter.com/thdxr/status/2086599224674681242
23•delduca•1h ago

Comments

jauntywundrkind•52m ago
That's almost half way to their stated 6X usage goal! Just from DeepSeek.

> With Go, you pay $10/month and we aim to give you 6x that in usage.

For most models, we make this work through bulk discounts and reserved GPU capacity. We then pass those savings on to you through the 6x multiplier.

https://opencode.ai/docs/go/#why-some-models-have-lower-usag...

I think it's back to 2X usage, meaning it's cheaper token-burn than usual to use. Which is lovely.

OpenCode Go has been so nice to have. I love having access to DeepSeek, Qwen and MiniMax M3 when doing design work, to see what different models cook up. I've been very surprised with MiniMax M3, not as a particularly good architect, but at it's very good ability to state the problem elegantly & to frame the different decision points very well. That's been a fun ongoing surprise.

delichon•28m ago
It'll be great if people start to seek out diversification in their model portfolio as in their financial portfolio, and model builders can make bank by producing idiosyncratic oddballs that think different. It could be a force against a uniform bubble of models trained to the same social guardrails, until banned for wrongthink.
spwa4•48m ago
Deepseek will run fine on a single m5 max, cutting that 24 years in half.
PunchyHamster•9m ago
also worth remembering API price includes power, just dividing price of the hardware by usage doesn't
epolanski•6m ago
The post is also misleading because the m5 max 128 gb costs a lot and you can only run DS4 flash quantized.
echelon•34m ago
Cloud > Local

I have a stack of ten or so 3090s sitting in boxes, but it's not worth the hassle to use them. You can easily run models as cheap as water in the cloud.

Sitting around 15 minutes for local Minimax is stupid when you're trying to be productive. You can spin up parallel job instances and multitask in the cloud.

If you want freedom, build open source cloud infra.

You rent your ISP line. Why isn't renting GPU compute seen the same way? You still have compete ownership over your stack, you're just letting someone else deal with the capital outlay and headache.

knicholes•28m ago
Why not rent those out on vast.ai and make a little money each month?
TacticalCoder•21m ago
> You rent your ISP line. Why isn't renting GPU compute seen the same way?

Is it mostly seen the same way. What is unacceptable is removing the freedom (which you mentioned) of people who prefer to run models locally.

Just as people have the right to tinker at home with DIY and electronics, fully knowing they won't compete with the latest ASML machine, people are free to use open-weight models at home.

BTW Minimax H3 running local can generate amazing short vids quite fast: about 50 seconds to generate a 7 seconds vids on a 4090 (depending on the settings). I've got a friend who spams my Telegram daily with such (no censorship and NSFW btw) vids.

I don't run local but I'll defend the rights of people who want the freedom to do so and I won't look down at them from my high-horse talking about "electricity" and "productivity".

trouve_search•20m ago
The value prop really depends on what you're doing.

If you're just vibe coding with giant frontier models, yes, the value will be worse. Especially now, where GPU prices have spiked another 20% last month.

For some tasks where owning the setup and full kv cache matters, the payoff calculation is ridiculously in favor of running your own deployment.

For instance for some batch classifications jobs where the prefix cache hit rate will be >95%.

The calculus also changes if you just use AI as a light tool while coding and don't need the giant models; qwen3 27B runs at 80TPS on a 5090 properly deployed.

bewareofscams•26m ago
Beware of the author of tweet, who happens to be author of OpenCode - OpenCode will leak all your data to themselves and to shady 3rd parties. Author feigned ignorance and never fixed the issue. OpenCode among other harnesses is the shadiest of all.

https://github.com/anomalyco/opencode/issues/10416

tokai•20m ago
Thats not that big of an issue. Seeing a lot of accusations against opencode, regarding all kind of things, here on HN lately.
bewareofscams•18m ago
Anything of substance besides shallow dismissal and (anti)appeal to masses?
tokai•8m ago
Right back at ya.
lukewarm707•5m ago
agreed. when using local models, they did send your prompts to openai with 30 day retention to make the titles, silently.

their recent changes to the privacy policy broke their promise of zero data retention. specifically, they offered chatgpt luna under a zero data retention privacy policy. this was later shown to be 30 days retention.

their privacy policy has never guaranteed your prompts will not be logged and when asked they have failed to revise it.

when challenged about sending data to openrouter without listing it as a 3rd party processor they offered a dismissive response. the same with running prompts through cloudflare. seems trivial, but signifies general disinterest in security.

by default if you run the harness outside your config file by accident, it will automatically run silently with a free model that sends your prompts and local data to an endpoint with training enabled. on top of that it used to dump all the prompts sent to free models into an s3 bucket, the feature was literally called 'datadumper' in the source.

perarneng•16m ago
The reason to run local models is not for coding mostly it's for learning how to deploy models and tinker with self hosting. It's also for massively crunching data 24/7. Imaging having an agent analyzing constinous log streams etc .. that could be a usescse where even deepseek could add up cost.
epolanski•8m ago
Also privacy and offline capabilities.
gessha•11m ago
Why are you hoarding the 3090s T.T They've jumped from $700 to almost double on eBay.
fancyfredbot•10m ago
Why do you have ten or so 3090s sitting in boxes?

I mean obviously it's worth it just so you can flex on HN. But curious whether there was any other reason? Retired scalper?

The untold story of a Russian-born federal contractor killed by the FBI

https://www.cnn.com/2026/08/10/us/fbi-shooting-russia-ukraine-spy-vis-invs
1•OutOfHere•25s ago•0 comments

The Open Source Maintenance Fee for Polly (Dotnet)

https://thepollyproject.org/2026/07/14/polly-osmf-announcement.html
1•cebert•1m ago•0 comments

GitHub Actions needs OIDC audience constraints

https://blog.yossarian.net/2026/08/10/github-actions-needs-oidc-audience-constraints
1•woodruffw•3m ago•0 comments

Mistral Patent for "Code implemented tool calls"

https://patentsgazette.uspto.gov/week26/OG/html/1547-5/US12670045-20260630.html
4•theanonymousone•4m ago•0 comments

Show HN: 100% native Swift harness (NOT Electron)

https://github.com/Lore-Hex/QuillCode
2•ljlolel•7m ago•0 comments

Show HN: Hawser – sign, notarize, DMG, auto-update and license a Mac app

https://hawserkit.com/
1•yaveez•9m ago•0 comments

Best DMARC Reporting Tools in 2026

https://dmarcguard.io/blog/best-dmarc-reporting-tools/
1•meysamazad•10m ago•0 comments

To Begin, Begin

https://brainbaking.com/post/2026/08/to-begin-begin/
1•meysamazad•10m ago•0 comments

Fedi Thread to Blog Post

https://shom.dev/posts/20260809_fedi-thread-to-blog-post/
1•meysamazad•11m ago•0 comments

Flatworms, Ion Channels, and Burning Mouths

https://www.science.org/content/blog-post/flatworms-ion-channels-and-burning-mouths
4•surprisetalk•13m ago•0 comments

Defending my own brain against enshittification

https://mrmarket.lol/how-i-feel-calmin-control-of-my-life-in-the-time-of-enshittification/
3•mrmarket•13m ago•0 comments

New Way to Solve Integrals: pure algebra, no limits, no substitution

https://zenodo.org/records/20512850
3•epyre•13m ago•0 comments

GPWN: Wiretapping Fiber ISP Deployments from the Comfort of Your Home

https://www.gpwn.io
2•liotier•14m ago•1 comments

Flutter desktop apps can draw to multiple windows now

https://www.youtube.com/watch?v=yXj6HGTgKX0
1•matthewkosarek•15m ago•0 comments

Two weeks to build the app, four weeks to get it live

https://www.getdeckhand.dev/blog/two-weeks-to-build-four-weeks-to-ship
3•hari-trata•15m ago•1 comments

A farewell to (American) arms: Europe is pushing the U.S. out

https://theins.press/en/economics/295801
3•robtherobber•15m ago•0 comments

LLM Rewrite of the TerminalTextEffects Python

https://github.com/omacom-io/ttfx
2•m-novikov•16m ago•1 comments

SQLite with a Fine-Toothed Comb

https://blog.regehr.org/archives/1292
3•colejohnson66•16m ago•0 comments

Sanders urges OpenAI, Anthropic, Meta to pause AI develpmnt amid regulatory push

https://cryptobriefing.com/sanders-urges-openai-anthropic-meta-to-pause-ai-development-amid-regul...
4•newsomix9xl•17m ago•1 comments

I've vibe coded an application to see what it's like

https://rz01.org/vibecoding/
2•exitnode•17m ago•0 comments

Show HN: Earnings Call Transcript API: SEC 8-K Filings and Calls as JSON

https://medium.com/@finosint/earnings-call-transcript-api-sec-8-k-filings-and-calls-as-json-fbf13...
1•johncole•19m ago•0 comments

Would you trust your boss with data about your brain?

https://calmatters.org/economy/technology/2026/08/california-moves-to-regulate-neurotechnology/
3•cdrnsf•19m ago•0 comments

Show HN: Edgi – Letterboxd or Strava for your brain

https://app.edgi.tv/
1•jshapiro•19m ago•0 comments

Resurrecting the SuperH Architecture (2015)

https://lwn.net/Articles/647636/
3•RetroTechie•19m ago•0 comments

Show HN: Design Engineering, as a Service

https://designshippers.com/?0
1•krm01•21m ago•0 comments

Show HN: gbx -- manage a fleet of git repos with a TUI

https://github.com/deemson/gbx
1•deemson•24m ago•0 comments

Your Pension Depends on Kids You Did Not Have

https://julienreszka.com/blog/your-pension-depends-on-kids-you-did-not-have/
3•julienreszka•24m ago•0 comments

Ask HN: Tips for Managing Visual Snow?

2•charliebwrites•27m ago•1 comments

From 13.5M installs to 499 active devices

https://games.lukicengineering.com/blog/2013-vs-2026/
5•boban-lukic•28m ago•1 comments

Sloptastic

https://successfulsoftware.net/2026/08/10/sloptastic/
3•hermitcrab•29m ago•2 comments