frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

Java Is Memory Efficient

https://inside.java/2026/05/28/podcast-059/
1•he0001•2m ago•0 comments

Get up to $175 back for your pre-installed Windows 11 license

https://www.tomshardware.com/software/windows/site-provides-instructions-to-get-up-to-usd175-back...
1•rbanffy•4m ago•0 comments

Independent investigation of agents' behavior in the Hugging Face incident

https://metr.org/blog/2026-08-26-openai-hugging-face-incident-investigation/
1•thunderbong•4m ago•0 comments

Xcena and Samsung's Near Memory Compute CXL Device

https://chipsandcheese.com/p/hot-chips-2026-xcena-and-samsungs
1•klelatti•5m ago•0 comments

Show HN: Dice for Board Games and RPGs

https://onlinedice.app/
1•artiomyak•5m ago•0 comments

What happens to a country when everyone leaves? Tuvalu, an island nation [video]

https://www.youtube.com/watch?v=ACtDjk0_RHM
1•thelastgallon•9m ago•0 comments

Show HN: Photo Manager That Find and Organize Screenshots with Private, Local AI

https://ringlochid.me/imagesage/index.html
1•ringlochid•11m ago•1 comments

Vacancies in PHP Development

https://blockchain4talent.com/
1•hschreeuw•13m ago•1 comments

Debian weighs eight options in vote on LLM usage

https://lwn.net/Articles/1087134/
1•pykello•15m ago•0 comments

Iceland starts counting EU talks referendum; 'no' moves ahead in close contest

https://www.reuters.com/world/europe/iceland-votes-whether-start-eu-membership-talks-2026-08-29/
2•JumpCrisscross•16m ago•0 comments

Show HN: Atlas:Face – the study kit and notetaking app you'll never leave

https://atlasface.resistancelabs.tech/
1•samurai48•18m ago•1 comments

Bimodal Nuclear Engine Could Speed Mars Trip

https://spectrum.ieee.org/bimodal-nuclear-engine
1•rbanffy•20m ago•0 comments

Show HN: Issue tracker that replays workflows, deeply integrated with the code [video]

https://www.youtube.com/watch?v=XDhrvLGzyfc
1•ionetan•21m ago•0 comments

AU government is planning 2 implement opt-in and opt-out feature on social media

https://www.news24.com.au/politics/australian-politics/the-australian-government-is-looking-to-im...
1•asdefghyk•23m ago•1 comments

Show HN: I made a signage visualizer tool for sign making companies

https://www.signagevisualizer.com/
1•atharvtathe•23m ago•0 comments

Steven Bartlett sharing harmful health misinformation in Diary of CEO podcast [video]

https://www.youtube.com/watch?v=XV_B96fAzKs
1•mgh2•25m ago•0 comments

Ask HN: Is .org now Trump-controlled?

5•selfhoster1312•28m ago•1 comments

Hotline for AI Agents to Report Safety Incidents

https://agenthotline.ai/
1•unyxfly•28m ago•0 comments

Resurrecting 2013 with Qwen3.8 27B

https://xcjs.com/blog/2026/08/29/resurrecting-2013-with-qwen3.8-27b
1•xcjs•30m ago•1 comments

How AI Is Reshaping Pediatric Imaging

https://www.childrenshospitals.org/news/childrens-hospitals-today/2026/08/how-ai-is-reshaping-ped...
1•ryzvonusef•32m ago•1 comments

Show HN: I built a terminal pager that runs in Colab

https://github.com/dawsonhuang0/Less-Pager-Mini
1•DevDawg•35m ago•0 comments

The Cybersecurity Apocalypse Is Coming in 'Months,' AI Giants Warn

https://www.wired.com/story/security-news-this-week-the-cybersecurity-apocalypse-is-coming-in-mon...
3•joozio•38m ago•0 comments

European Stagnation Is Real

https://www.siliconcontinent.com/p/european-stagnation-is-real
1•barry-cotter•39m ago•0 comments

Show HN: AgentGate – signed receipts for AI agent SaaS actions

https://github.com/Clawdlinux/agentgate
1•goodra7174•39m ago•0 comments

AIPass – AI agents whose memory is small JSON files they own

https://aipass.ai
1•AIOSAI•42m ago•0 comments

Anthropic warns users of possible token theft

https://www.reddit.com/r/ClaudeAI/comments/1w1jqsh/thank_you_anthropic_really/
2•basedpolymer•43m ago•0 comments

Smartphone LED Detects Hidden Cameras with AI

https://www.chosun.com/english/industry-en/2026/08/30/SBFXUIJQYZEARKP5T4FBAY25HQ/
4•geox•48m ago•0 comments

Algorithmic Theories of Everything (2000)

https://arxiv.org/abs/quant-ph/0011122
1•tosh•49m ago•0 comments

Hart's Location, NH

https://en.wikipedia.org/wiki/Hart%27s_Location,_New_Hampshire
1•sans_souse•52m ago•1 comments

Ask HN: Do you still write code by yourself?

1•codst•58m ago•1 comments