frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.

Dozens hurt and hundreds arrested in wave of French school protests

https://www.bbc.co.uk/news/articles/c862epne7glyo
1•FridayoLeary•6m ago•0 comments

Best Bitcoin Commercial

https://twitter.com/bramk/status/2105780416141738076
1•cft•7m ago•0 comments

Stop trying to make xenophobia sound smart

https://www.theargumentmag.com/p/stop-trying-to-make-xenophobia-sound
1•NomNew•11m ago•0 comments

Show HN: Giving Opus 5.5 a simulated paint canvas

https://stillwet.art/
2•alstonite•13m ago•0 comments

Superhuman AI for Stratego

https://arxiv.org/abs/2511.07312
2•droidjj•13m ago•0 comments

NASA has a Dragon dilemma, and there appear to be no good answers

https://arstechnica.com/space/2026/09/nasa-has-a-dragon-dilemma-and-there-appear-to-be-no-good-an...
1•Terretta•16m ago•0 comments

Quest: Open-source, safety-focused harness

https://github.com/electric-capital/quest
2•puntium•18m ago•1 comments

Where Are the Prepared Minds? The Impact of AI on Scientific Thinking

https://ai.nejm.org/doi/full/10.1056/AIe2601110?query=featured_home
1•Alien1Being•20m ago•0 comments

Google's Project Suncatcher prototype satellite is in orbit

https://blog.google/innovation-and-ai/models-and-research/google-research/project-suncatcher-prot...
4•xnx•33m ago•0 comments

Fortress(Tilion YC F26): A stealth Chromium so your agents stop getting blocked

1•armanluthra26•34m ago•2 comments

Kuhnian Shift to "Agent Science and Engineering"

https://arunis100.medium.com/agent-science-and-engineering-the-future-of-cs-and-data-science-c70c...
1•davi•35m ago•0 comments

Thought-Terminating Cliché

https://en.wikipedia.org/wiki/Thought-terminating_cliché
6•tiagod•36m ago•0 comments

Show HN: Soothsay - Your AI types curl | sh. soothsay checks the script first

https://github.com/rijuld/soothsay
1•rijuldahiya•36m ago•0 comments

Show HN: Zero Slop – Open-source skill that finds and fixes AI slop in writing

https://github.com/manavmishra/ZeroSlop
2•mmishra7•37m ago•0 comments

A lightweight Python CLI to audit and detect orphaned AWS resources

https://github.com/Ace7-coder/aws-finops-auditor
1•Ace7-coder•39m ago•0 comments

Bashy: A pure-Go Bash 5.3 for Linux, macOS and Windows that also speaks Bash

https://github.com/qiangli/bashy
2•thunderbong•40m ago•0 comments

Hiring Breakpoints

https://staysaasy.com/hiring-breakpoints/
1•zdw•42m ago•0 comments

Show HN: Bun-install – Global CLI installer for Bun with local monorepo support

https://github.com/joeycumines/bun-install
2•joeycumines•43m ago•0 comments

What do you want from AI

https://www.anthropic.com/research/your-thoughts-on-ai
1•benwen•46m ago•0 comments

Mass exploitation of Citrix NetScaler: What we currently know

https://www.cybersecuritydive.com/news/exploitation-citrix-netscaler-what-we-know/831780/
4•wicket•47m ago•0 comments

C++ Insights – See your source code with the eyes of a Compiler

https://github.com/andreasfertig/cppinsights
1•rramadass•47m ago•1 comments

Show HN: Grist, open source coding harness

https://grist.lol/
1•pranav6226•1h ago•0 comments

What is zero-knowledge, end-to-end encryption?

https://blog.mega.io/what-is-zero-knowledge-end-to-end-encryption
4•devonnull•1h ago•0 comments

Noodle AI Bots Mac App

https://usenoodle.app/
1•handfuloflight•1h ago•0 comments

A fifth of Swiss Alpine ice has vanished in five years

https://www.swissinfo.ch/eng/glaciers-permafrost/a-fifth-of-swiss-alpine-ice-has-vanished-in-five...
7•Teever•1h ago•0 comments

Show HN: Pyxel – A Python retro game engine with built-in art and sound editors

https://github.com/kitao/pyxel
3•kitao•1h ago•1 comments

The Softmax function and its derivative

https://eli.thegreenplace.net/2016/the-softmax-function-and-its-derivative/
2•costco•1h ago•0 comments

Tiny Brutalism

https://placeholders.itch.io/tiny-brutalism
3•abetusk•1h ago•0 comments

Show HN: Shootsolo – a free camera app for filming yourself without a crew

https://shootsolo.com/
1•KarimMuhtar•1h ago•0 comments

Butterflies use optical illusions to dodge predators

https://www.essex.ac.uk/news/2026/09/30/butterflies-use-optical-illusions-to-dodge-predators
8•gmays•1h ago•2 comments