frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

Zig: Pointer Stability for ArrayLists

https://ziglang.org/devlog/2026/#2026-08-27
1•tosh•1m ago•0 comments

Scaling Domain Data Repetition in LLM Pretraining

https://arxiv.org/abs/2608.14071
1•gmays•4m ago•0 comments

Emacs vs. Vim 2026: 24.3% vs. 8% Usage, 20x Memory Gap

https://tech-insider.org/emacs-vs-vim-2026/
1•pir8life4me•6m ago•0 comments

Free and Open-Source Web Search for Your AI Agents

https://pypi.org/project/simurg/
1•lebagetdefrance•7m ago•0 comments

Remembering the King who Drowned

https://komando1.substack.com/p/remembering-the-king-who-drowned
1•inglor_cz•7m ago•1 comments

Show HN: FinBridge – Korean stock market data for AI agents (MCP)

https://www.gronox.kr/
1•gronox-kr•9m ago•0 comments

Keeping an open source project going when GitHub stars slow down

https://twitter.com/paul_h0rn/status/2094070350120091915
2•BlueBerry2001•11m ago•0 comments

Ethical Issues in Advanced Artificial Intelligence

https://nickbostrom.com/ethics/ai
2•Anon84•11m ago•0 comments

Show HN: Cushion – PouchDB on Deno KV

https://jsr.io/@mikehall314/cushion
3•mikehall314•12m ago•0 comments

DNS Abuse and Criminal Infrastructure

https://labs.ripe.net/author/andrew_campling/dns-abuse-and-criminal-infrastructure-beyond-definit...
2•jruohonen•12m ago•0 comments

Doom Played on Series of 555 Timers

https://hackaday.com/2026/08/28/doom-played-on-series-of-555-timers/
2•nickbild•12m ago•0 comments

Europe's summer drought is so extreme that desertification is a growing threat

https://fortune.com/2026/08/29/europe-summer-drought-desertification-threat-rivers-fish/
4•Brajeshwar•13m ago•0 comments

Getting into Flow with AI Coding

https://www.natashatherobot.com/p/flow-ai-coding
2•ingve•14m ago•0 comments

Founder Mode vs. Founder Drift

https://keithbrown.com/founder-mode-vs-drift/
1•tech-pulse•14m ago•0 comments

Open Pinball Database

https://opdb.org/
1•Tomte•16m ago•0 comments

Search BioNumbers – The Database of Useful Biological Numbers

https://bionumbers.hms.harvard.edu/Search.aspx?task=searchbypop
1•Tomte•17m ago•0 comments

Ask HN: Fun Productive Phone Hobbies?

2•bix6•19m ago•2 comments

Iran's Unit 4000 and the new architecture of outsourced terrorism

https://www.israelnationalnews.com/news/432418
1•pir8life4me•25m ago•0 comments

Show HN: An anonymous, no-signup alert app – try it live with GitHub incidents

https://thenotifier.app/developers/see-it-work/
1•bubbamack•27m ago•0 comments

7th Branch Prediction Championship: Part I

https://www.sigarch.org/7th-branch-prediction-championship-part-i/
1•jruohonen•27m ago•0 comments

Robots: We Can't Look Away – and Can't Stop Worrying

https://www.bajura.online/2026/08/robots-why-we-cant-look-away-and-cant.html
1•AllForAll•27m ago•0 comments

Rapidly spreading genetic variants make parasites resistant to malaria treatment

https://www.brown.edu/news/2026-08-17/malaria-drug-resistance
1•gmays•28m ago•0 comments

My BC-250 Journey (With 40 CUs Unlocked)

https://worldofmatthew.com/blog/bc250/
1•worldofmatthew•28m ago•0 comments

China Tobacco products have likely caused at least 57M deaths since 1990

https://www.theexamination.org/articles/deadliest-company-in-the-world
3•cwwc•29m ago•1 comments

Icelanders vote 'no' to talks aimed at joining the EU in close referendum

https://www.npr.org/2026/08/29/nx-s1-5948759/iceland-european-union-referendum-vote
3•elsewhen•31m ago•0 comments

The Law of Leaky Abstractions

https://www.joelonsoftware.com/2002/11/11/the-law-of-leaky-abstractions/
2•maayank•31m ago•2 comments

Dario Amodei says Anthropic is 'not interested in destroying anyone'

https://www.businessinsider.com/dario-amodei-anthropic-not-destroying-saas-2026-8
1•potatobox•32m ago•1 comments

Show HN: Open SAR-COP – AI-generated common operating picture for disasters

https://open-sar-cop.github.io
1•Euraxluo•33m ago•0 comments

Open science mid-year review: ten news stories from early 2026

https://blogs.helsinki.fi/thinkopen/open-science-review-2026-mid-year/
1•jruohonen•33m ago•0 comments

METR and Redwood Offer Holy %^ Postmortem of the HuggingFace Hack

https://thezvi.wordpress.com/2026/08/29/metr-and-redwood-offer-holy-postmortem-of-the-huggingface...
1•catbird•36m ago•0 comments