frontpage.
newsnewestaskshowjobs

Made with ♥ by @iamnishanth

Open Source @Github

fp.

Ask HN: Anyone Using a Mac Studio for Local AI/LLM?

49•UmYeahNo•2d ago•31 comments

Ask HN: Opus 4.6 ignoring instructions, how to use 4.5 in Claude Code instead?

3•Chance-Device•10h ago•0 comments

Ask HN: Ideas for small ways to make the world a better place

22•jlmcgraw•1d ago•22 comments

Ask HN: Who wants to be hired? (February 2026)

139•whoishiring•5d ago•528 comments

Ask HN: Non AI-obsessed tech forums

35•nanocat•1d ago•28 comments

Ask HN: 10 months since the Llama-4 release: what happened to Meta AI?

45•Invictus0•2d ago•11 comments

Ask HN: Who is hiring? (February 2026)

313•whoishiring•5d ago•515 comments

LLMs are powerful, but enterprises are deterministic by nature

5•prateekdalal•19h ago•7 comments

Tell HN: Another round of Zendesk email spam

105•Philpax•3d ago•54 comments

AI Regex Scientist: A self-improving regex solver

7•PranoyP•1d ago•1 comments

Ask HN: Is Connecting via SSH Risky?

19•atrevbot•3d ago•37 comments

Ask HN: Has your whole engineering team gone big into AI coding? How's it going?

18•jchung•2d ago•14 comments

Ask HN: Non-profit, volunteers run org needs CRM. Is Odoo Community a good sol.?

3•netfortius•1d ago•1 comments

Ask HN: Is there anyone here who still uses slide rules?

123•blenderob•4d ago•122 comments

Kernighan on Programming

171•chrisjj•5d ago•62 comments

Ask HN: Mem0 stores memories, but doesn't learn user patterns

9•fliellerjulian•3d ago•6 comments

Ask HN: How does ChatGPT decide which websites to recommend?

5•nworley•2d ago•11 comments

Ask HN: Is it just me or are most businesses insane?

8•justenough•2d ago•7 comments

Ask HN: Why LLM providers sell access instead of consulting services?

5•pera•1d ago•13 comments

Ask HN: What is the most complicated Algorithm you came up with yourself?

3•meffmadd•1d ago•7 comments

We built a serverless GPU inference platform with predictable latency

5•QubridAI•2d ago•1 comments

Ask HN: Does a good "read it later" app exist?

8•buchanae•4d ago•18 comments

Ask HN: Anyone Seeing YT ads related to chats on ChatGPT?

2•guhsnamih•2d ago•4 comments

Ask HN: Have you been fired because of AI?

17•s-stude•4d ago•15 comments

Ask HN: Does global decoupling from the USA signal comeback of the desktop app?

5•wewewedxfgdf•2d ago•3 comments

Ask HN: Anyone have a "sovereign" solution for phone calls?

12•kldg•4d ago•1 comments

Ask HN: Cheap laptop for Linux without GUI (for writing)

15•locusofself•4d ago•16 comments

GitHub Actions Have "Major Outage"

53•graton•5d ago•17 comments

Ask HN: Has anybody moved their local community off of Facebook groups?

23•madsohm•5d ago•20 comments

Ask HN: OpenClaw users, what is your token spend?

15•8cvor6j844qw_d6•5d ago•6 comments
Open in hackernews

Ask HN: Is token-based pricing making AI harder to use in production?

3•Barathkanna•3w ago
Hi HN,

I’ve noticed a recurring theme in many threads here: AI is powerful, but once you move past demos, token based pricing becomes expensive and hard to reason about.

We ran into this problem ourselves while building AI powered systems. Predicting costs, budgeting usage, and experimenting safely all got harder as workloads grew. So we built a small AI API platform for inference, aimed at early developers and small teams who want to integrate AI without constantly calculating token usage. The focus is on lower and more predictable costs rather than chasing the newest model.

This is still early, and I’m mainly posting to learn from others here. For people running AI in production, what’s been the hardest part to manage so far? Cost, predictability, performance, or something else?

I’d really appreciate any insights or experiences.

Comments

iamrobertismo•3w ago
Not clear what you are pitching, if you don't control the infrastructure or have a major contract, how exactly are you lowering or stabilizing costs. Especially if you are not chasing the newest model, at this point token economics is essentially a commodity. Commodity pricing is not a engineering problem, it is a financing problem.
Barathkanna•3w ago
That’s fair, and I probably didn’t explain it clearly. We’re building an AI API as a service platform aimed at early developers and small teams who want to integrate AI without constantly thinking about tokens at all.

I agree that token economics are basically a commodity today. The problem we’re trying to address isn’t beating the market on raw token prices, but removing the mental and financial overhead of having to model usage, estimate burn, and worry about runaway costs while experimenting or shipping early features. In that sense it’s absolutely an engineering and finance problem combined, and we’re intentionally tackling it at the pricing and API layer rather than pretending the underlying models are unique.

iamrobertismo•3w ago
Would you just be... subsidizing low volume users? I am saying this isn't like a new problem in the grand scheme of things. hopefully I am not being too negative, do you have a site or something to learn more? It's not clear how you can have better token economics to provide me or someone else better token economics, rather than just burning more money lol.
Barathkanna•3w ago
Totally fair question, and you’re not being negative.

We’re not claiming better token economics in the sense of magically cheaper tokens, and we’re not just burning money to subsidize usage indefinitely. You’re right that this isn’t a new problem.

What we’re building is an AI API platform aimed at early developers and small teams who want to integrate AI without constantly reasoning about token math while they’re still experimenting or shipping early features. The value we’re trying to provide is predictability and simplicity, not beating the market on raw token prices. Some amount of cross-subsidy at low volumes is intentional and bounded, because lowering that early friction is the point.

If you want to see what we mean, the site is here: https://oxlo.ai Happy to answer questions or go deeper on how we’re thinking about this.

iamrobertismo•3w ago
Oh you're arbing! I see now. Makes sense, seems like it could be useful if you have a rock solid DX.
Barathkanna•3w ago
Thank you!! We are definitely fully focused on Developer experience. Would love some feedback if it looks interesting
elmascato•3w ago
The point about "removing the mental overhead" is underrated. It is often the cognitive load of the pricing model, rather than the absolute cost, that kills adoption.

I'm seeing a strong parallel in SaaS regarding Purchasing Power Parity (PPP). We often assume a user in India or Brazil doesn't convert because they "can't afford" $49, but the friction is often psychological. Even for high earners in those regions, paying a double-digit USD subscription feels "wrong" or predatory relative to local goods.

Just as you are abstracting away the token math to lower the barrier to entry, we need to abstract away the currency inequality. I've been working on a client-side widget to handle this (tierwise.dev) and noticed that simply aligning the price with the user's local context making it "feel" fair spikes conversion rates significantly.

Whether it's flattening token variance or localizing purchasing power, the goal is the same: stop the user from doing math and let them focus on the product value.