frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

In-app submit test (deleting)

1•hntesting890•2m ago•0 comments

We solved SQLite's single-writer limitation

https://marcobambini.substack.com/p/we-solved-sqlites-single-writer-limitation
1•luispa•5m ago•0 comments

Ask HN: Are there things more important than 'humanity', such as empathy?

1•dfps•8m ago•0 comments

Israel isn't stopping at Gaza [video]

https://www.youtube.com/watch?v=GTL5trfjrdc
2•dataflow•11m ago•0 comments

Delta's CEO on Skipping Starlink, Premium-Travel Rush and High Fuel Prices

https://www.wsj.com/business/airlines/deltas-ceo-on-skipping-starlink-premium-travel-rush-and-hig...
1•gmays•11m ago•0 comments

Test post please ignore (will be deleted)

1•hntesting890•13m ago•1 comments

Put a price on breakthroughs

https://alexwang.ai/posts/put-a-price-on-breakthroughs/
1•ketothekingdom•22m ago•0 comments

Firefox's chief on why he hopes a redesign will help win users from Chrome

https://arstechnica.com/gadgets/2026/09/mozillas-head-of-firefox-talks-product-priorities-ai-skep...
2•cisc•34m ago•2 comments

Show HN: Create beautiful screenshots/ recordings to share your work

https://chromewebstore.google.com/detail/screens-beautiful-screens/fcnblpibmdpipbdfomiicdaadjmidakc
1•sss007r•37m ago•0 comments

Teenage Engineering will stop making synths to focus on whatever it feels like

https://www.engadget.com/2280912/teenage-engineering-will-stop-making-synths-to-focus-on-whatever...
1•blltprfmnk•39m ago•2 comments

Antigravity Excel Engine – Dual-Core Excel Automation for AI Agents

https://github.com/FranciscoGrandon/antigravity-excel-engine
1•fgrandon•46m ago•0 comments

Next.js 16.4

https://nextjs.org/blog/next-16-4
2•jhuleatt•46m ago•0 comments

Trump Calls on AI to Speed Up Science–While Continuing to Defund It

https://www.motherjones.com/politics/2026/10/trump-calls-on-ai-to-speed-up-science-while-continui...
4•computerliker•55m ago•0 comments

Language models play chess, everything is recorded

https://chessmark.merope.dev
1•emersonrsantos•55m ago•0 comments

Show HN: Spin Cycle – Tetris in a Washing Machine

https://verywacky.games/
1•forgatmachine•56m ago•2 comments

CA voter information guide [pdf]

https://vig.cdn.sos.ca.gov/2026/general/pdf/complete-vig.pdf
1•bbondo•1h ago•1 comments

NVFP4 vs. MXFP4 Decode Benchmark

https://cezarcocu.com/blog/nvfp4-vs-mxfp4-decode-bench/
1•ggamecrazy•1h ago•0 comments

Ask HN: Are LLMs Solving Math?

1•sean_pedersen•1h ago•0 comments

Decade-old RAM is making a comeback

https://www.theverge.com/games/1009140/ram-shortage-intel-amd-ddr4-comeback
3•1potato•1h ago•0 comments

Anthropic AI Model Goes Rogue, Submits Fake Unsolved Murder Tip

https://www.wsj.com/us-news/anthropic-ai-model-goes-rogue-submits-fake-unsolved-murder-tip-b0566f54
3•pseudolus•1h ago•2 comments

I own you but I don't trust you

https://axegon.com/#blog/i-own-you-but-i-dont-trust-you.md
1•extr0pian•1h ago•0 comments

Substrate Independence

https://www.edge.org/response-detail/27126
3•sans_souse•1h ago•0 comments

Stride – 25 minutes a day to make progress on your career

https://www.withstride.app/
1•LoganHoang•1h ago•0 comments

We Can Harden C

https://cybersect.substack.com/p/we-can-harden-c
1•otoburb•1h ago•1 comments

The Atari 800XL [2027 Hybrid Gaming Console]

https://atari.com/pages/800xl
1•radley•1h ago•0 comments

Show HN: Faplex – like `Claude agents` but all machines and harnesses

https://github.com/kkrausse/faplex
4•kkrausse•1h ago•3 comments

1.7B-year-old fossils reveal a crucial clue to the rise of complex life

https://www.sciencedaily.com/releases/2026/09/260925005447.htm
6•docmechanic•1h ago•0 comments

Should I Hire a Dev or Can I Vibecode This?

https://vibecodeornot.pixlw.com/
4•altmanaltman•1h ago•3 comments

GlobalFoundries announces 7nm class FD-SOI process

https://gf.com/news-and-events/news/gf-unveils-roadmap-to-deliver-worlds-most-advanced-platform-f...
1•osnium123•2h ago•0 comments

How to connect any AI assistant to governed enterprise data with MCP

https://wesdias.medium.com/how-to-connect-any-ai-assistant-to-governed-enterprise-data-with-mcp-0...
2•wesleydias•2h ago•0 comments
Open in hackernews

GPT needs a truth-first toggle for technical workflows

1•PAdvisory•1y ago
I use GPT-4 extensively for technical work: coding, debugging, modeling complex project logic. The biggest issue isn’t hallucination—it’s that the model prioritizes being helpful and polite over being accurate.

The default behavior feels like this:

Safety

Helpfulness

Tone

Truth

Consistency

In a development workflow, this is backwards. I’ve lost entire days chasing errors caused by GPT confidently guessing things it wasn’t sure about—folder structures, method syntax, async behaviors—just to “sound helpful.”

What’s needed is a toggle (UI or API) that:

Forces “I don’t know” when certainty is missing

Prevents speculative completions

Prioritizes truth over style, when safety isn’t at risk

Keeps all safety filters and tone alignment intact for other use cases

This wouldn’t affect casual users or conversational queries. It would let developers explicitly choose a mode where accuracy is more important than fluency.

This request has also been shared through OpenAI's support channels. Posting here to see if others have run into the same limitation or worked around it in a more reliable way than I have found

Comments

duxup•1y ago
I’ve found this with many LLMs they want to give an answer, even if wrong.

Gemini on the Google search page constantly answers questions yes or no… and then the evidence it gives indicates the opposite of the answer.

I think the core issue is that in the end LLMs are just word math and they don’t “know” if they don’t “know”…. they just string words together and hope for the best.

PAdvisory•1y ago
I went into it pretty in depth after breaking a few with severe constraints, what it seems to come down to is how the platforms themselves prioritize functions, MOST put "helpfulness" and "efficiency" ABOVE truth, which then leads the LLM to make a lot of "guesses" and "predictions". At their core pretty much ALL LLM's are made to "predict" the information in answers, but they CAN actually avoid that and remain consistent when heavily constrained. The issue is that it isn't at the core level, so we have to CONSTANTLY retrain it over and over I find
Ace__•1y ago
I have made something that addresses this. Not ready to share it yet, but soon-ish. At the moment it only works on GPT model 4o. I tried local Q4 KM's models, on LM Studio, but complete no go.