Of course, and so does everything in the software world. The point is getting the cost so low that it’s basically free. The new DS V4 Flash or the smaller Qwen3.6 models are still really expensive compared to what we were used to in the economics of software, but it’s not unreasonable to expect these costs to continue falling down.
Rough chatgpt estimate says 3-5 orders of magnitude of difference compared to a typical user interaction with a SPA (db/cache lookup, CDN…)
Well, no. Copying is free, or so near free it makes zero sense to charge. LLMs are just papering over the damage caused by profit.
But, it's a hard problem. The models that run locally on normal computers/phones are pretty terrible compared to the frontier, without specialization and fine-tuning. And, even with specialization and fine-tuning, often a high-end general purpose model is going to do a better job and people don't need a bunch of local tools installed to do their various tasks.
I think this is the critical point that would be interesting to see if it holds. Technology seemingly tends towards increased specialization.
Right now the "hosting" cost for inference is per-unit because it's new and expensive, but that won't last.
There is a lot of inefficiency right now keeping prices elevated. That will change very fast and soon paying for inference will likely resemble paying for hosting your app.
The bigger problem for SaaS is that the floor has risen - people can build their own solutions for things that they used to buy SaaS for. So the industry needs to level up and solve harder problems.
It’s so cheap that companies choose to spend more on AI inference (more reasoning, more capabilities, longer context), not less - see Jevons paradox.
Most SaaS already works this way. M365 or Adobe Creative Cloud are great examples. They value it like a life insurance policy and find ways to make you sticky. It’s easier to just buy it.
The first round of AI products suck because they are not well defined. Copilot only makes sense if you do shit in office and SharePoint isn’t a dumpster fire. In my large O365 environment the bottom 50% of users use less storage than the top 2%. So why would i buy copilot for my janitor?
When M365 E9 reconciles invoices automatically with Excel, I’ll pay $150/mo and fire a bunch of people.
Absolutely not. Customers want systems for sales, reservations, accounting, and taking stock. That's where almost all the SaaS money is and none of it benefits from AI - and never will.
Not really.
>Every inference call costs money.
Not really, either. If you buy your own GPU, rack it, and run an open model, there is no unit cost. This is just expensive hosting infra. You also pay unit costs for SaaS that your software uses (things like SMS etc).
No. There is economic opportunity cost (borrowing), energy cost, infra cost, depreciation / risk of failure with each unit of work, bandwidth, maintenance, and lots more. Small, but not zero, and often overlooked - especially the opportunity cost.
smalltorch•8h ago
Are there any examples of products containing ai inference that are successful? Products that are beyond just direct access to frontier LLM's, I mean.
esafak•55m ago
cwmoore•48m ago