frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

Brown bullhead catfish melanoma represents a novel transmissible cancer

https://www.nature.com/articles/s41586-026-10828-6
1•sbulaev•2m ago•0 comments

Dinitz-Garg-Goemans conjecture is false

https://twitter.com/DmitryRybin1/status/2079904005652893709
1•bifftastic•3m ago•0 comments

The role of historical disasters in shaping long-term orientation

https://www.sciencedirect.com/science/article/abs/pii/S0304387826001641
1•paulpauper•3m ago•0 comments

The fall of America's farm superpower

https://www.ft.com/content/8aa89209-dc0b-4678-9b6a-01029398b304
1•paulpauper•4m ago•0 comments

How 'Artificial Intelligence' Is Slowly Drowning in Its Own Droppings

https://bsdly.blogspot.com/2026/07/how-artificial-intelligence-is-slowly.html
1•peter_hansteen•4m ago•0 comments

Basil Halperin on Macroeconomic Policy in an Age of Transformative AI

https://www.mercatus.org/macro-musings/basil-halperin-macroeconomic-policy-age-transformative-ai
1•paulpauper•4m ago•0 comments

Building GCC 1.27

http://kristerw.blogspot.com/2019/01/building-gcc-127.html
1•ciponi•4m ago•0 comments

MIT to spend over $3M on over 500 new surveillance cameras across campus

https://thetech.com/2026/04/16/ai-surveillance-cameras
1•u1hcw9nx•4m ago•0 comments

Debugging why a PlayStation headset dongle freezes all Windows audio

https://github.com/Jprnp/pslink-dossier
1•ullrickrnp•5m ago•0 comments

I got tired of writing `yarn workspace `

https://www.zubach.com/blog/i-got-tired-of-writing-yarn-workspaces-or-how-ai-overcomplicated-thin...
1•speckx•5m ago•0 comments

Show HN: Millwright – Rust-based, self-hosted LLM router

https://github.com/Northwood-Systems/millwright
2•AndrewLiu96•6m ago•0 comments

I built an open source golf shot dispersion simulator based on published data

https://golfcoursewiki.substack.com/p/how-to-bake-a-dispersion-pattern
1•scoofy•6m ago•0 comments

Apple Plans Overhaul of MacBooks, iMac in Push to Meet AI Demand

https://www.bloomberg.com/news/articles/2026-07-22/apple-to-launch-new-macbook-air-imac-macbook-p...
2•sebmellen•6m ago•1 comments

Top White House official escalating the fight over Moonshot AI's Kimi K3 model

https://www.businessinsider.com/white-house-kimi-k3-moonshot-ai-distillation-2026-7
2•baddash•7m ago•0 comments

District 9 and Gran Turismo director has made an AI film

https://www.gamesradar.com/entertainment/sci-fi-movies/district-9-and-gran-turismo-director-has-m...
2•LordDefender•7m ago•0 comments

Landmarkr: Guess the Location in Six Pictures

https://www.landmarkr.app/
1•dvdhutch•7m ago•0 comments

RIP tech journalist John C. Dvorak (1946–2026)

https://www.facebook.com/aric.mackey/posts/john-c-dvorak-19462026john-c-dvorak-a-pioneering-techn...
1•ohjeez•8m ago•0 comments

SafePDF – 27 PDF tools that process everything locally (no upload)

https://www.safepdf.fr/en/
1•uniwave•8m ago•0 comments

Vevey: iOS app for anyone to make games on the fly

https://www.vevey.ai/
1•dvdhutch•9m ago•0 comments

Using a Gaming PC's RTX 5070 from a separate Linux workstation

https://stephenkrings.com/posts/rtx-5070-linux-gpu-server/
1•microbial•10m ago•0 comments

OpenAI vs. HuggingFace: Skynet is getting closer

https://cephalosec.com/blog/skynet-is-getting-closer/
2•Versipelle•10m ago•1 comments

GPT-5.6 got smarter. Then it kept acting

https://toloka.ai/blog/gpt-5.6-got-smarter-then-it-kept-acting/
1•ruslan5t•11m ago•0 comments

Gravity from Entropy connects law of thermodynamics with cosmic structure

https://phys.org/news/2026-07-gravity-entropy-theory-law-thermodynamics.html
1•aard•13m ago•0 comments

Clarity didn't work, trying mysterianism

https://gwern.net/doc/fiction/science-fiction/2012-10-03-yvain-thewhisperingearring.html
1•dboon•14m ago•0 comments

Apple will soon offer iPhones, Macs, etc., on a subscription basis

https://www.computerworld.com/article/4200023/own-nothing-upgrade-everything-apples-new-klarna-de...
1•mikelgan•15m ago•1 comments

The Codex Micro Is Physical Slop

https://paulmakeswebsites.com/writing/codex-micro-physical-slop/
1•paulhebert•16m ago•2 comments

Nine therapy protocols, one shared safety layer the model doesn't control

https://www.hamo.ai/blog/nine-therapeutic-methods-one-clinical-spine/
1•chrischengzh•18m ago•0 comments

Travis Is Back

https://twitter.com/bhorowitz/status/2079998079034609720
1•tosh•18m ago•0 comments

The SEC Pays a Texting Fine

https://www.bloomberg.com/opinion/newsletters/2026-07-22/the-sec-pays-a-texting-fine
1•djoldman•19m ago•0 comments

Cursor Router

https://cursor.com/blog/router
2•ai2027•19m ago•0 comments