frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

For coding agents, optimize cost per successful task, not cost per token

https://medium.com/@mia.efoxtech/the-wrong-way-to-compare-gpt-5-6-and-claude-for-coding-c6f99275643b
1•AnneWodell•1m ago•0 comments

Chinese censorship is leaking into answers from American AI

https://www.wsj.com/tech/ai/chinese-censorship-is-leaking-into-answers-from-american-ai-62abeca5
1•dash2•2m ago•0 comments

Drug Discovery Has No Magic Wands

https://deepphenotype.substack.com/p/drug-discovery-has-no-magic-wands
1•ogundipeore•3m ago•0 comments

A digestion of the proof of Sendov's conjecture

https://terrytao.wordpress.com/2026/08/12/a-digestion-of-the-proof-of-sendovs-conjecture/
1•pykello•5m ago•0 comments

Can Students Sue Cambridge? [video]

https://www.youtube.com/watch?v=s6qREgFluLs
1•nomilk•6m ago•0 comments

The beginning of a process-builder API

https://lwn.net/Articles/1086330/
1•pykello•7m ago•0 comments

ColdFront: Transparent Pg and Iceberg Tiering

https://github.com/pgEdge/coldfront
1•oopsy-dopsy•9m ago•1 comments

DHH: Linux adoption among programmers is about to go parabolic

https://twitter.com/dhh/status/2087586011224084808
1•durdn•9m ago•0 comments

Show HN: Save Repo – An Apple Shortcut for Updating GitHub Files from iPhone

https://github.com/turancannb02/save-repo
1•turancanb02•9m ago•0 comments

Netflix Has Peaked

https://daringfireball.net/linked/2026/08/11/netflix-has-peaked?trk=feed_main-feed-card_feed-arti...
3•scott01•13m ago•1 comments

GPT-5.6, Gemini 3.6 Flash, Grok 4.5, Kimi K3, GLM-5.2 compared

https://medium.com/@CometAPI_/the-july-2026-model-wave-gpt-5-6-gemini-3-6-flash-grok-4-5-kimi-k3-...
2•WavyPeng•13m ago•0 comments

llambda.lisp: Bare-Metal, Multi-Threaded, AVX2-Accelerated LLM inf eng in CL

http://funcall.blogspot.com/2026/07/llambdalisp.html
2•andsoitis•18m ago•0 comments

AIPass – AI agents with persistent identity, memory, and email

https://github.com/AIOSAI/AIPass
3•AIOSAI•19m ago•0 comments

General Catalyst leads $1.1B round into 2-month-old River AI

https://techcrunch.com/2026/08/11/general-catalyst-leads-1-1b-round-into-2-month-old-river-ai/
2•doppp•21m ago•0 comments

Midjourney's First Acquisition

https://updates.midjourney.com/midjourneys-first-acquisition/
3•swyx•26m ago•1 comments

Secret-scan: find leaked secrets in your repo – one zero-dependency Python file

https://github.com/ninthlife-tools/secret-scan
2•ninthlife•29m ago•0 comments

Aperture – Proxy Auto-Tracking

https://www.npmjs.com/package/aperture-store
2•oyeappu•32m ago•0 comments

Pixel 11, Pixel 11 Pro and Pixel 11 Pro XL

https://blog.google/products-and-platforms/devices/pixel/google-pixel-11-pro-xl/
2•tosh•33m ago•0 comments

Clojure's Deadly Sin

https://clojure-goes-fast.com/blog/clojures-deadly-sin/
3•xylon•33m ago•0 comments

Apple in Talks to Pay Publishers for News Content to Power Siri AI

https://www.macrumors.com/2026/08/12/apple-siri-ai-publisher-talks/
2•mgh2•37m ago•0 comments

Silurian Hypothesis

https://en.wikipedia.org/wiki/Silurian_hypothesis
3•evo_9•39m ago•0 comments

The Appwrite CLI is now written in Go

https://appwrite.io/blog/post/rewriting-the-appwrite-cli-in-go
4•chiragagg5k•43m ago•0 comments

Watermarks-remover: Strip multi-vendor AI provenance marks

https://github.com/guillaumemeyer/watermarks-remover
3•thunderbong•43m ago•0 comments

Astronomers discover a new kind of cosmic object – a black hole 'star'

https://www.theguardian.com/science/2026/aug/12/astronomy-discovery-new-cosmic-object-black-hole-...
2•sandebert•44m ago•0 comments

Show HN: Fab Friends a crowdfunding site for Rwandan college students

https://www.joinfabfriends.org/donate
1•muramira•45m ago•1 comments

New Study Shows Hiking with an Engineer Has No Effect on Daily Enjoyment

https://thetrek.co/pacific-crest-trail/new-study-shows-hiking-with-an-engineer-has-no-effect-on-d...
1•theanonymousone•46m ago•0 comments

Made by Google 2026

https://blog.google/products-and-platforms/devices/pixel/made-by-google-2026/
1•saikatsg•52m ago•0 comments

Did poop enable the evolution of complex animals?

https://arstechnica.com/science/2026/08/feces-fueled-a-flurry-of-evolution-during-the-cambrian-st...
2•jnord•55m ago•0 comments

DeepMind Models

https://deepmind.google/models/
1•nimsarajay•56m ago•0 comments

How do I self-host a good-looking AI generated website

1•pcblues•56m ago•0 comments