frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Spending on AI Is Becoming Almost Impossible for Businesses to Budget

https://www.wsj.com/tech/personal-tech/ai-token-spending-businesses-431ee94a
24•swolpers•1h ago

Comments

FinnLobsien•25m ago
You can set spending limits, but I don't feel like that helps much because everyone's accustomed to AI. Nobody would accept "We're out of usage so we have to wait until Monday" and do all of their work manually.

I believe that we're in a scenario where usage is unlikely to go down and neither are frontier AI costs.

I believe we'll see a shift to more organizations building their own harnesses with model routing logic to get central control over who can use what AI and for what.

A marketer doesn't need to default to Opus 5.5 to upload a blog article with MCP, which could be done by a model 10% of the price.

awesan•5m ago
I have some bias as a software dev/manager but at least in our org I noticed you can get a lot better value out of your tokens with some proper configuration of context and skills. Providing the latest models with good context can result in extreme savings.

After we started excessively documenting and extracting skills everyone has stopped complaining about running out of tokens, because agents stopped having to reconstruct the full context each time from scratch. Harnesses like claude code also push the model to aggressively keep this documentation in sync so there's little concern about drift.

The worst thing to do from a token usage pov is to give a model a vague open ended prompt because they are so scared to be wrong that they'll waste a ton of tokens "thinking" through the issue and verifying everything. Whereas they almost trust skills blindly and skip all this unnecessary work.

skeptic_ai•3m ago
Well at a big bank I work, if you finish in 1 week you do manual coding the rest of the month until next month. Management doesn’t give a shit. And they have daily talks about using more AI.
simianwords•20m ago
Disagree with this because we have ways to steer price use per task.

1. choose a good model

2. choose the appropriate reasoning effort

3. choose a prompt to nudge it even further

Then it comes down to understanding the intuition of what kind of task deserves what effort?

CharlieDigital•10m ago
Now get the 1000 engineers in your company to be as well-behaved and mindful as you are.

(Couldn't even get this to happen in a 30 person team...)

epistasis•3m ago
That "intuition" is not intuitive. How do I choose an appropriate reasoning effort?

I use these models all day long, experiment, and have no clue how to choose that.

It just showed up one day in the interface with no explanation or guidance. Its use is mysterious, its effects unclear except through intensive experimentation, and to this day it mostly seems "how many bad decisions will Claude go forward with when it finally dumps out screenfuls of text instead of getting better guidance early on" though it's certainly not a guarantee on anything.

These models are being released at breakneck speed even before their creators know how to use them. It's a big project of collective discovery to figure out what they are doing and how to use them.

hilariously•13m ago
I see a lot of businesses who just dump this all on their employees and then get mad at their employees effectively for not following the most recent LLM related talk on twitter.

Most employees can't tell you what database to use, what software programming framework to use, what document management framework to use, but they are expected to know which of the 25 models available to use for a task, budget appropriately, monitor efficacy, update models to the most relevant for a task, continue to manage architecture patterns??? for LLM agents, this list goes on.

This is getting stupid folks.

nathanaldensr•11m ago
Was it ever not stupid?
2OEH8eoCRo0•10m ago
I use "auto" in vscode which likely selects for cheapest. Good enough!
ndriscoll•3m ago
Assuming the employees are calling themselves "engineers" here, who else is supposed to be doing that work? If we were going through a renaissance in material science, don't you think it would be the mechanical engineers who try to figure out how to figure out what sorts of materials might be suitable for their use case? Surely you don't want the sales team or executives figuring that out.
neom•11m ago
https://archive.ph/aBXzS
rglover•9m ago
Turns out running an unattended LLM like a slot machine is expensive.

IMO, human in the loop is the only serious usage of AI (I know, I know, "software factories bro"). Everything else is a hope and a prayer and a big bill.

bentt•8m ago
I do quite a bit of coding with Claude but am perfectly fine on the $20/mo plan. You people who just let agents go for hours on end... I'm not sure you're doing it right.
cyanydeez•4m ago
they're not doing it right _or_ they're doing it correctly.

We're sorta in the age of alchemy. Lots of cranks out there, but there are real recipes.

sajithdilshan•3m ago
Why not just set a spending limit per person and extend/adjust the limit case by case. This would actually make people be more mindful about burning tokens on useless stuff
bravetraveler•2m ago
https://unwall.app/www.wsj.com/tech/personal-tech/ai-token-s...

Ada Lovelace answered the big questions about AI

https://www.nytimes.com/2026/10/05/opinion/ada-lovelace-ai.html
1•ripe•35s ago•0 comments

Nvidia's $20B licensing deal with Groq faces lawsuit from jilted engineers

https://www.ft.com/content/93ee425d-9ac7-4543-8cc9-fef2e0670787
1•JumpCrisscross•1m ago•0 comments

AI-Ready Data: 4 Foundations for More Reliable Enterprise AI

https://devnavigator.com/2026/10/05/ai-ready-data-foundations/
1•thescienceguy93•1m ago•0 comments

TaxCalcBench: Can AI file your taxes?

https://taxcalcbench.ai/
1•michaelrbock•2m ago•0 comments

One of the reasons learning Arabic is so hard (صعب)

https://claude.ai/artifact/3siWSWUruHevarmRMZ5pHr
1•tzury•2m ago•0 comments

Mollie Pay by Bank

https://www.mollie.com/gb/payments/payment-methods/pay-by-bank
2•TechTechTech•4m ago•0 comments

MusicBrainz Picard 3.0

https://picard.musicbrainz.org/changelog/
1•pluc•5m ago•0 comments

How fast is Python 3.15?

https://blog.miguelgrinberg.com/post/how-fast-is-python-3-15
1•ibobev•5m ago•0 comments

Show HN: GoodStanding, free diagnostic for nonprofits auto-revoked by the IRS

https://goodstanding.thecompound.tech
1•kyisaiah47•5m ago•0 comments

Destroying all of humanity is hard work, even for a superintelligence

https://nibblestew.blogspot.com/2026/10/destroying-all-of-humanity-is-hard-work.html
1•ibobev•6m ago•0 comments

Counting Elements in CSS: Using Sibling-Count and Hacks

https://blog.master.dev/transferring-sibling-count-to-a-parent-element/
1•ibobev•6m ago•0 comments

Building a Faster (Rust-Based) Python Driver for ScyllaDB

https://www.scylladb.com/2026/09/28/building-a-faster-rust-based-python-driver/
1•tzach•7m ago•0 comments

Om Malik, Last Humanist

https://logranmar.substack.com/p/om-malik-last-humanist
1•speckx•9m ago•0 comments

I turned Facebook into a video game

https://domdefense.com
1•skellertor•10m ago•1 comments

Making a GTK application in Haskell, part 1

https://floreal.tech/blog/2026/making-a-gtk-app-in-haskell-part-1/
3•Vosporos•11m ago•0 comments

Cloudflare Web Search API

https://developers.cloudflare.com/web-search/
1•gregzeng95•11m ago•1 comments

Machine Identity Attestation and Auth for Bare Metal and Virtualised Systems

https://infisical.com/blog/machine-identity-attestation-bare-metal
1•FinnLobsien•11m ago•0 comments

Ask HN: Which Distro with Nividia GPU?

2•tietjens•12m ago•1 comments

Claude playing old strategy games in 2026

http://replicated.live/blog/games
1•gritzko•14m ago•0 comments

AeroBounce Game: One-finger neon chaos and endless runs to beat your high score

https://aerobounce.saposs.com/
1•jimmy_lee•14m ago•0 comments

Show HN: Pumpkins.sh – claim, carve, and display a pumpkin to the world

https://pumpkins.sh/
2•linesofcode•16m ago•0 comments

Windows 11 Adds Tilde Shortcut to File Explorer

https://thewincentral.com/windows-11-file-explorer-tilde-unicode-17/
2•layer8•17m ago•0 comments

Hacker News: An Apology

https://lee-phillips.org/byeHN/
4•leephillips•17m ago•0 comments

Turn tool usage and thinking context

https://github.com/zackemannen81/A008
1•mrwhite81•17m ago•1 comments

How Fuzzy Search Works in a Search Engine

https://serenedb.com/blog/fuzzy-search-deep-dive
1•mkornaukhov•18m ago•0 comments

Show HN: Vibivibi – End-to-end encrypted sharing of Coding Agent sessions

https://vibivibi.com/
1•hyluo•18m ago•0 comments

Windows Developers Get a Built-In Linux Container Runtime: WSL Containers

https://opensourcewatch.beehiiv.com/p/windows-developers-get-a-built-in-linux-container-runtime-w...
1•CrankyBear•19m ago•0 comments

Are any of your neighbouring countries a current enemy?

https://brilliantmaps.com/polls/are-any-of-your-neighbouring-countries-a-current-enemy/
1•cgh•19m ago•1 comments

Agentic Machine Learning Modeling at Instacart

https://tech.instacart.com/agentic-machine-learning-modeling-at-instacart-fb3ecd295ee7
1•skadamat•21m ago•0 comments

Evergreen Common Lisp (EGCL)

https://github.com/atgreen/evergreen
1•Thom2503•22m ago•0 comments