frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

Fundamental Theorem of Calculus [Slop]

https://www.proofatlas.ai/library-theorems/fundamental-theorem-calculus/
1•throwyonion•1m ago•0 comments

Show HN: Vsqrd – AWS for Biology Experiments

https://vsqrd.com/
1•misterchocolat•4m ago•0 comments

The First Ten Years

https://gregorygundersen.com/blog/2025/01/30/first-ten-years/
1•andsoitis•5m ago•0 comments

Succinct Sounds-Like Starting Vowels

https://wiredream.com/vowels/
1•dgacmu•6m ago•0 comments

Show HN: Lain, a structural code graph and agent coordinator for coding agents

https://github.com/spuentesp/lain
1•spuentesp•9m ago•1 comments

Two similar AI judges fail together 7.7x more often than independence predicts

https://github.com/LAWLESS1987/covenant
1•Lawless1987•12m ago•0 comments

A refined phylochronology of the second plague pandemic in Western Eurasia

https://www.pnas.org/doi/10.1073/pnas.2534899123
1•Thevet•18m ago•0 comments

How I view LLMs as Sept 2026

https://bhardwajrish.blogspot.com/2026/09/how-i-view-llms-as-sept-2026.html
1•crearo•21m ago•0 comments

ZCode: silently uploading your Git history to the cloud

https://news.routley.io/posts/inside-zcode-silently-uploading-your-git-history-to-the-cloud.html
3•fourfire•24m ago•0 comments

A CERN for AI-assisted science

https://terrytao.wordpress.com/2026/09/17/a-cern-for-ai-assisted-science/
1•andsoitis•24m ago•0 comments

Show HN: Converting AVS audio viz to WebAssembly using GPT 5.6

https://strlns.github.io/very-serious-vibecoding/AVS-Web-Native-Editor-Alpha/
1•moritzwarhier•26m ago•0 comments

Spending on data centers and hardware now exceeds housing investment

https://fortune.com/2026/09/20/us-economy-milestone-spending-data-centers-ai-boom-housing-residen...
3•rawgabbit•26m ago•1 comments

Amiga Unix, Again

https://amigaux.org/
4•doener•28m ago•2 comments

Building an NBA stats warehouse without writing any of the code

https://jeffknupp719305.substack.com/p/paul-reed-is-tenth
2•jknupp•30m ago•0 comments

Prototype of a Windows 11-inspired Web OS with autonomous IA agents

https://github.com/davidpoliquin30-design/Prototype-autonom-ia-OS
1•davidpoliquin•32m ago•0 comments

Tools to Prevent Dementia

https://www.scmp.com/lifestyle/health-wellness/article/3368006/why-dementia-preventable-and-new-t...
1•elo2000•41m ago•0 comments

New startup to test psychedelic-like drug against Parkinson's symptom

https://www.statnews.com/2026/09/17/biotech-news-ariadne-testing-psychedelic-like-drug-parkinsons...
2•elo2000•42m ago•0 comments

Chronic stress may trigger hidden inflammation that damages the heart

https://lms.mrc.ac.uk/chronic-stress-may-trigger-hidden-inflammation-that-damages-the-heart/
2•doener•44m ago•1 comments

Are we talking less? A Q&A with psychologist Matthias Mehl (March 2026)

https://news.arizona.edu/news/are-we-talking-less-qa-psychologist-matthias-mehl
1•BiraIgnacio•46m ago•0 comments

Does Jev Have Politics? (Yes)

https://idlerambling.substack.com/p/does-jev-have-politics-yes
1•Sketchy_•50m ago•0 comments

Turn curiosity into a map you can explore

https://www.echohive.ai/experiments/drift
1•echohive42•52m ago•0 comments

AI Netscape – 1997 Chrome, 2026 AI

https://ainetscape.com/
2•TStringham•56m ago•2 comments

The addiction of online sports gambling: "Like crack in the '80s"

https://www.cbsnews.com/news/sports-betting-apps-gambling/
4•dzonga•58m ago•2 comments

The Crown of Creation

https://mika.global/post/1789946288.html
1•mika314•58m ago•0 comments

KDE Goes to the Pyramids [video]

https://media.ccc.de/v/kde2026-58-kde_goes_to_the_pyramids
2•sho_hn•1h ago•0 comments

How Anthropic CEO Dario Amodei's Writings Help Explain A.I. Fears

https://www.nytimes.com/2026/09/17/technology/dario-amodei-anthropic-essays-ai.html
4•gmays•1h ago•0 comments

Popular European Games

https://playfrom.eu/
1•doener•1h ago•0 comments

EkkoBSD

https://ekkobsd.org/
3•threemux•1h ago•0 comments

DAPO: An Open-Source RL System from ByteDance Seed and Tsinghua Air

https://github.com/BytedTsinghua-SIA/DAPO
4•the_arun•1h ago•0 comments

My AI Kept Pushing Me to Ship, So I Asked It Why

https://www.oreilly.com/radar/my-ai-kept-pushing-me-to-ship-so-i-asked-it-why/
2•Anon84•1h ago•0 comments