AI's Economics Don't Make Sense

https://www.wheresyoured.at/ais-economics-dont-make-sense/

82•spking•1h ago

Comments

wonderwhyer•1h ago

Yeah. And weird pricing seems like it's winding down.

It's interesting to compare it to electricity. Basically Anthropic was selling a flat fee electricity subscription, and when someone started connecting expensive washing machines (OpenClaw) to their subscriptions, instead of changing the pricing model, they banned washing machines...

I wonder if we will get to "electricity" style pricing for AI. What makes electricity predictable is relatively constant average usage over time + price is manageable. I'm just not buying electrical house heating and manage my electricity spending within some bounds.

With AI the problem is that we are only now getting to useful AI, and for now it's still too expensive to be useful, so they subsidize until they can stabilize at "cheap enough and smart enough" level. But it feels like that's still 2 years away while they are stopping to subsidize now. Will be interesting.

linkregister•54m ago

OpenClaw was never banned from the Claude API, only flat-fee plans.

gruez•35m ago

>Basically Anthropic was selling a flat fee electricity subscription

No? It was flat, but with ambiguously stated limits (eg. 5x, 10x 20x). They were discriminating on how the "electricity" was used, but that's not that much different than how power companies have different rates for residential users vs industrial users.

ethin•15m ago

Even now they are insanely ambiguous with respect to their usage limits. They don't from what I know openly disclose them anywhere, so them saying "5x increase" is utterly meaningless, alongside "20x" or "10x" or whatnot, because we don't know what "x" is.

swader999•13m ago

The Uber subscription analogy works well too.

jcgrillo•58m ago

The finding out phase has begun.

asah•57m ago

meh - by this logic, every new tech and startup ever is a "scam"

The truth is that the AI companies are gambling that inference cost will continue following a hyper version of Moore's Law, e.g. Google TurboQuant.

The countervailing thesis is that frontier models are consuming more and more compute.

The deepest truth: you often don't need a frontier model to get commercially acceptable results from AI. Thus, bring on the true pricing! and I'll just switch models to something financially sustainable.

swader999•9m ago

We work comes to mind. The math is fairly easy if we know what a company like OpenAI's datacenter commitments are, what their sub and token revenue is right now and what their operation costs are. This is very basic and if you had that info you would know exactly if we are in bubble or not. Waiting for the S-1's...

wood_spirit•55m ago

The general problem the average user has with a metered instead of provisioned billing model for computer services is you can’t easily control for cost overruns. From the old days customers getting stung for hosting costs when slashdotted or DOSed, to last decades microservice shock horror of the CI retry loop that burns money overnight to today’s AI that you basically have no idea how efficient the AI will be while it ponders your question, you are just setting yourself up for disappointment and cost overruns and a feeling that you’re not getting the value for money you got last week etc.

gruez•32m ago

>The general problem the average user has with a metered instead of provisioned billing model for computer services is you can’t easily control for cost overruns.

Is this an actual issue aside from people letting their autonomous agents run overnight?

wood_spirit•28m ago

I can speak of myself. Sometimes my session starts out well and I get the AI to cruise to 80%. But then gains after that seem impossible and what was built steadily unravels and then I get the compacting conversation message and realise that I’ve just spent a lot of money on nothing.

zozbot234•5m ago

Cost overruns? What cost overruns? The HTTP API just returns 402 Payment Required once you're above your paid-for quota.

lbrito•51m ago

>At some point, the incredible, toxic burn-rate of generative AI is going to catch up with them, which in turn will lead to price increases, or companies releasing new products and features with wildly onerous rates (..) that will make even stalwart enterprise customers with budget to burn unable to justify the expense.

I pray this happens soon, but I feel I've been hearing some version of it for a while.

ambicapter•46m ago

Big ships take a while to turn.

ToucanLoucan•45m ago

The only reason it hasn't is the sheer amount of credit being thrown at this tech. Both that and the valuations of the firms in question is stratospherically over-hyped and over-valued.

This tech has uses. It has quite a lot of them in fact. However there is no usage of ChatGPT or Claude that makes OpenAI or Anthropic worth anything fucking close to what they're valued at right now, and both firms are scrambling to figure out how to get down from the top of the AI house of cards without detonating in the process.

Meanwhile DeepSeek is coming out with more capable models that run on far less onerous hardware and with far less compute requirements that does basically exactly what the vast majority of users actually want it to do.

This is going to be a financial bloodbath. Not for anyone actually responsible for it, of course, they'll be fine. It'll be everyone else getting soaked which is the only reason I give two shits.

joshjob42•44m ago

There's a few major problems with the article. The most obvious is that frontier labs are not charging remotely close to the cost of tokens; afaik most estimate north of 80% profit margins. As a reference, providers are profitably providing Kimi K2.6 for $4/1Mtok out. Is that as good as Opus? No, but it's probably at least Sonnet level, so that's ~4x cheaper than Sonnet while still being profitable to serve on the margin. So you aren't plausibly getting into actual subsidization territory until you're over 5:1 sub to nameplate token costs.

How many tokens can you realistically burn through in one chat session? Opus and many other frontier models do maybe 60tok/s, less 250k/hr out. In you can use more, but in most cases cache is 5-10:1 cheaper than new input. Say you average 500ktok in, 90% cache, per request. That amounts to 100-150ktok in new input-equivalent costs, which in most cases is ~20-30ktok in output-equivalent costs. Do a request every minute, that's a total of about 1.5-2Mtok/hr. At API prices that's $50/hr for Opus, but really it probably only costs Anthropic $10/hr to serve that.

That said, even if a developer is burning $50/hr, many, many employees at large companies cost more than $100k/yr to employ all costs considered, so making them say 20-30% more productive can easily make that worth it for most. If the labs shave their margins ultimately to more like 20-30%, you'd have ~$15/hr in costs to use the services, and nearly every white collar job is way over 30k/yr to employ. If your salary is 80k, you probably cost the company 200k all in, so making you 15% more productive offsets the $15/hr cost.

So first party providers are not in a horrifying position or anything from a subsidization standpoint. The people in bad shape are Cursor and Perplexity, who don't have frontier models and are dependent on the open source community, which is typicly 6-12 months behind the frontier. They have to pay full freight API costs at 80% margin for the big boys to serve their harnesses, which is indeed untenable, and they'll have to either force users to use open source models and/or in house models they can serve at-cost or they will have to charge vastly more.

Gemini, Claude, and ChatGPT first-party services like Antigravity, Codex, and Claude Code are not in serious trouble though.

ToucanLoucan•29m ago

> That said, even if a developer is burning $50/hr, many, many employees at large companies cost more than $100k/yr to employ all costs considered, so making them say 20-30% more productive can easily make that worth it for most. If the labs shave their margins ultimately to more like 20-30%, you'd have ~$15/hr in costs to use the services, and nearly every white collar job is way over 30k/yr to employ. If your salary is 80k, you probably cost the company 200k all in, so making you 15% more productive offsets the $15/hr cost.

Nobody including the connected article is making the argument that this cannot be profitable ever. People are saying "there is no way this admittedly quite interesting tool is going to be able to make back all of this money" and I think they are completely right to say that.

You can absolutely make money with this stuff, just not at this scale. The buildout for this shit has been certifiably crazy and a number of the involved firms are overleveraged for tens and even hundreds of billions of dollars.

How in the sweet fuck are you paying that off, plus giving investors dividends, selling this at $15/hour/user??? That math does not math. A quick google says there are between 1.5 and 4.4 million developers in the US alone, let's say it's 5 million, to be generous, and each of them is subbed to this for 8 hours per day, continuously. That's 600 million per year in revenue. If you took ALL that revenue, and put it towards paying down this debt, not leaving any for employee salaries, upkeep, ongoing development, it would take DECADES to pay down what OpenAI already owes.

And yes I'm sticking directly to code, because that's the only thing I've seen it be really good at. Are we really proposing that every knowledge worker on earth and every manager of such workers is going to have an autonomous agent running all the time!? To do what, make sure they don't have to read or write email? Which even just that example is bringing in a fucking mess of legal, compliance, and security violations because LLMs are not intelligent and are not capable of being properly secured.

Like I'm sorry, I cannot take this industry seriously when even the most basic back-of-napkin math is saying, nay, screaming from the rooftops that they are FUCKED.

vidarh•22m ago

By your numbers, it'd be $120/day per developer * 5 million = $600m per day, not per year.

Of course people don't work every day, but even with European-level holidays that number is off by a factor of 240 or so.

ToucanLoucan•14m ago

Quite right, honestly not sure how I fucked that up so bad but I'll own it. Okay so all we need is every coder + 0.6 million more or so in the United States, subscribed to this for 8 hours a day, and the business model can work.

That still feels incredibly optimistic given how split the community at large seems to be about how good this tech is, and it assumes all those developers also all work for firms large enough to pay for all of that.

However we are still very much in back of napkin math. We haven't even gone into what it costs to provide these services, how much it's going to cost yet for all these datacenters to be built, how much electricity and water they're going to rip through, and all the rest. So IMO, we've now elevated it from "hopeless" to "this could work if a whole lot of other things line up really well."

strongpigeon•21m ago

> That's 600 million per year in revenue.

According to your math, that's $600 million per day

marcosdumay•7m ago

Yes, the GP wrote the wrong unit on this place. That supports his conclusion that the pay-off would take decades, if it was actually per year, it would take several centuries.

belval•16m ago

> selling this at $15/hour/user??? That math does not math. A quick google says there are between 1.5 and 4.4 million developers in the US alone, let's say it's 5 million, to be generous, and each of them is subbed to this for 8 hours per day, continuously. That's 600 million per year in revenue

That math is not mathing. $15/hour/user, with 5M devs, 8hrs and 240 working days per year that is 144B in revenue.

loeg•8m ago

> How many tokens can you realistically burn through in one chat session?

I've used single digit billions in a couple days, FWIW.

cheeseblubber•41m ago

It make sense if you account for cost of intelligence getting cheaper every year. Most of the models per unit of intelligence is getting far cheaper. We get better hardware, architecture, training techniques, inference optimizations and caching. All those improvements add up. In in early 2022 you were getting 10x cheaper annually now is closer to 2x - 5x cheaper annually. The cost is still dropping where as Uber can only get the cost down by so much.

iooi•37m ago

The entire basis of this article is that generating tokens is a variable cost and that that cost will not decrease over time.

> On an economic basis, a monthly subscription only makes sense with relatively static costs.

Running a data center is a fixed expense. Whether or not people use that data center to it's capacity doesn't change how much the operator pays (electricity use factors into this, since a GPU running at 100% will use more watts than an idle one, but it doesn't move the needle much on other fixed and variable costs of a data center).

> They also assumed, I imagine, that the cost of tokens would come down over time, versus what actually happened — while prices for some models might have come down, newer “reasoning” models burn way more tokens, which means the cost of inference has, somehow, gotten higher over time.

This is backwards. When the cost of something goes down, people use it more. This is basic supply and demand. Inference has gotten cheaper already, and will continue to do so.

Companies subsidizing costs for growth happens all the time. Yes, switching to usage-based pricing instead of subscriptions sucks for customers, but enterprises will continue to pay.

xnx•25m ago

> it doesn't move the needle much on other fixed and variable costs of a data center

I wonder what the rough costs of a data center look like over the lifetime of one GPU generation?

10% building

60% GPU

30% power

I haven't gone looking for that information, but I haven't run across it either.

Marciplan•36m ago

I am a paying subscriber to Ed Zitron and I enjoy his writing a lot. He should at some point admit that not everything is bullshit and there is definitely a business model to it. It is fun to read, though

mediaman•24m ago

He has a fun writing style but has so many willful errors, and is so committed to one point of view regardless of the facts, that his writing seems kind of worthless.

I soured on him when he could not calculate cumulative revenue on an exponential curve, ignored everyone who showed him how to calculate it, and then kept writing that Anthropic’s revenue numbers are fake based on his inability to do math.

It’s too bad because any heavily hyped industry needs good critics (think Ida Tarbell to Rockefeller) but they should be honest critics, and he’s not, which really undermines not only his but others’ criticism of the industry.

xnx•18m ago

It's good to have contrarian viewpoints, but Ed Zitron is so blinded by his AI hate that his articles should be treated not just with skepticism, but heavy suspicion.

putzdown•32m ago

The moves from “the subscription model for AI isn’t working given these parameters” to “a subscription model for AI can never work” to “the model was deliberately deceptive” to “it’s a fucking ripoff” is not logical. AI companies are feeling the need to get hold of spiraling costs by increasing prices and limitations. Inference hasn’t gotten cheap enough fast enough, and for some reason they feel they can’t wait longer. That doesn’t mean a subscription service can’t work: only that it will be expensive, maybe vastly so, and will need tiers based on usage with some fluidity for users to move between tiers in a given month. The model is something like HP’s “instant ink” service. Sure, there’s a question whether the moves companies are making now are worth the cost in the eyes of customers. But that’s a question of economics and timing, not a fundamental blow to monthly subscriptions as a model. The article doesn’t deal with these considerations fairly. It’s too much in the direction of a rant, with conspiracy theories thrown in.

christkv•30m ago

I'm just flabbergasted at the massive inefficient usage of tokens. What are people doing to spend 500 usd/day in tokens. I just don't understand what you could possibly be doing that would be not complete spagetti at the end if you run something in an autoloop.

xnx•19m ago

> What are people doing to spend 500 usd/day in tokens

1) They're lying

2) Status signalling

milesvp•28m ago

Reading this piece, I'm reminded of a podcast I heard some years ago where they were interviewing an early google marketing employee who was talking about the economics of google search. They said they'd done some surveys and concluded that they determined that the average user would get something like $20/year of value, and so that was the most they could realistically charge for search. Meanwhile, they could make something like $500/user in Q4 alone for advertising. So, of course, advertising.

I just don't think that LLM business models can survive the allure of advertising dollars, any more than Search could, or TV, or Radio, or Movies. Ignoring the talk of copilot putting ads into pull requests, there is just no way that publicly hosted LLMs will not end up inserting ads into the output.

This looks like what I remember. https://freakonomics.com/podcast/is-google-getting-worse/

Ritewut•27m ago

It makes sense when you realize the goal is not the consumer but large gov and enterprise contracts.

throwawayajner•26m ago

Zitron misunderstands the economics of models. Inference costs have dropped 99% in less than 2 years. Models are being commoditized faster than any technology in history.

A $20 subscription 2 years ago is not providing the same level of intelligence you're getting today.

Every major lab knows open source models are 6 months behind (See Google's "We have no moat") and none of them plan to make money on inference. Companies are subsidizing users to create moats that persist when models are essentially free for most everyday use.

mNovak•24m ago

Do we know the breakdown of revenue from API vs subscriptions for OAI/Anthropic? That seems very relevant, since this entire article seems to be on the premise that users are only willing to pay for a subsidized subscription and would never pay the 'true' token cost.

The internet seems to be saying that 70%+ of Anthropic revenue is per-token metered API, which would largely invalidate the article, but I can't find a solid source.

swader999•15m ago

I don't think these companies will give this information up until their hand is forced with an S-1 when they want to IPO. So stay tuned...

feverzsj•16m ago

It makes perfect sense, if you treat it as a Ponzi scheme.

ameliaquining•15m ago

As it happens, published just this morning is an article from Kelsey Piper that explains in some detail what's wrong with Zitron's takes: https://www.theargumentmag.com/p/ais-biggest-critic-has-lost...

matchagaucho•15m ago

Same debate as the dot-com era.

Customer: “I don’t want to pay more than $100/mo for my website” Developer: “What are your goals?” Customer: “1M daily visits, 1,000 monthly signups.”

And we've spent the past 25 years offering serverless compute, auto-scaling, pay-as-you-go for AWS and Internet infrastructure. And the economics are still a hard sell.

threepts•13m ago

I thought this burning of cash was all an excuse for the exponential growth we saw in the last 6 years.

They went from GPT 2 a text only, goldfish-esque memory at a 8th grade reading level to what we have today, GPT 5, multimodality + a token window encompassing a enclyopedia and a Doctorate/Masters level of mastery in major subjects.

The economics are probably betting on this exponential growth to continue, which if it fails, the cash would burn.

Glyptodon•9m ago

I think there's another route this goes. At $7k a year or more per eng in token use, I think it's very reasonable to buy engineers machines with obscene GPUs and RAM and run models locally. And if it doesn't make sense now, someone will figure it out and save companies $10k+/eng over 3 years.

Cloudflare Q1 2026 Internet disruption summary

The Feedback Loop in AI SDLC

A Texas developer got a $2B loan to build Oracle data centers in the 'burbs

From CVS to Git: thirty years of source control, lived from inside

Chatbot Act Introduced in Senate

Claude.ai is unavailable

Canonical's approach to AI is refreshingly thoughtful-Microsoft should take note

The Americans queueing up to renounce their citizenship

Humorphism – The Human Interface

We Have to Hold the Line Against the Pseudoscience of Facilitated Communication

Mastering Celestial Navigation [video]

10Gb Ethernet: what I had to (re)learn

Commodore: Re-Introducing the C64C

Colorado's Anti-Repair Bill Is Dead

MacBook Pro M5 review: serious power, still long battery life

X Is Down - edit - was down

The Bloomberg Terminal Is Getting an AI Makeover, Like It or Not

Blueprint: A planning copilot that one-shots bigger coding tasks

A DOGE Affiliate Is Now in Charge of the US Government's ID Platform

How to Keep Your Brain Sharp: A Practical Playbook Beyond the Basics

Claude Design Is 404ing

Phaser: Create 2D games for the web – free, open source, and AI-ready

Wild GPT-image-2 use cases

Amtaitfy – Let Me Google That for You, but the AI Is Wrong on Purpose

Nvidia Nemotron 3 Nano Omni

Height hunt: a quest to find and visit every possible low bridge / height restri

Shots Fired by Google Cloud CEO Thomas Kurian

Woman's Talkspace therapy app sessions exposed in court

The Guard Act Isn't Targeting Dangerous AI–It's Blocking Everyday Internet Use

GPT-Engineer: Precursor to Lovable.dev