frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Ask HN: LLM is useless without explicit prompt

4•revskill•1y ago
After months playing with LLM models, here's my observation:

- LLM is basically useless without explicit intent in your prompt.

- LLM failed to correct itself. If it generated bullshits, it's an inifinite loop of generating more bullshits.

The question is, without explicit prompt, could LLM leverage all the best practices to provide maintainable code without me instruct it at least ?

Comments

ben_w•1y ago
Your expectations are way too high.

> - LLM is basically useless without explicit intent in your prompt.

You can say the same about every dev I've worked with, including myself. This is literally why humans have meetings rather than all of us diving in to whatever we're self-motivated to do.

What does differ is time-scales of the feedback loop with the management:

Humans meetings are daily to weekly.

According to recent research*, the state-of-the-art models are only 50% accurate at tasks that would take a human expert an hour, or 80% accurate at tasks that would take a human expert 10 minutes.

Even if the currently observed trend of increasing time horizons holds, we're 21 months from having an AI where every other daily standup is "ugh, no, you got it wrong", and just over 5 years from them being able to manage a 2-week sprint with an 80% chance of success (in the absence of continuous feedback).

Even that isn't really enough for them to properly "leverage all the best practices to provide maintainable code", as archiecture and maintainability are longer horizon tasks than 2-week sprints.

* https://youtu.be/evSFeqTZdqs?si=QIzIjB6hotJ0FgHm

revskill•1y ago
It's not as high as you think.

LLM failed at the most basic things related to maintainable code. Its code is basicaly a hackery mess without any structure at all.

It's my expectation is that, at least, some kind of maintainable code is generated from what's it's learnt.

ben_w•1y ago
Given your expectation:

> It's my expectation is that, at least, some kind of maintainable code is generated from what's it's learnt.

And your observation:

> LLM failed at the most basic things related to maintainable code. Its code is basicaly a hackery mess without any structure at all.

QED, *your expectations* are way too high.

They can't do that yet.

Negative Visualization

https://en.wikipedia.org/wiki/Negative_visualization
1•gurjeet•39s ago•0 comments

Noviz is a complete ERP AI connector source code

1•tajdink•2m ago•0 comments

Ha.mr – short URLs using compression

https://ha.mr
1•matan-h•5m ago•1 comments

Eclipse: The Xiaomi 17 Ultra Confuses the Moon and the Sun

https://www.frandroid.com/marques/xiaomi/3211257_photo-de-leclipse-on-a-perce-a-jour-la-petite-tr...
1•n_plus_1_acc•7m ago•0 comments

Private security firms will soon be allowed to hack overseas cybercriminals

https://arstechnica.com/security/2026/08/white-house-recruits-security-firms-to-hack-overseas-cyb...
2•joozio•8m ago•1 comments

The Body Remembers but the Doctor Cannot- Experiments with AI in Personal Health

https://karankurani.com/writing/post/w_20260721140414_b0e102/the-body-remembers-but-the-doctor-ca...
2•BIackSwan•11m ago•0 comments

NPM Left-Pad Incident

https://en.wikipedia.org/wiki/Npm_left-pad_incident
3•gherkinnn•19m ago•0 comments

Show HN: OpenCode Auto Permissions – automatic review of permission prompts

https://github.com/hueyexe/opencode-auto-permissions
2•huey77•22m ago•0 comments

Mistral AI wants to build 1 gigawatt of European compute by 2030

https://venturebeat.com/infrastructure/mistral-ai-wants-to-build-1-gigawatt-of-european-compute-b...
2•tosh•22m ago•0 comments

Pay for SF uniformed police and Sheriff deputies is approaching $900k per year

https://twitter.com/loomdoop/status/2087994362219450743
2•MrBuddyCasino•23m ago•0 comments

The Case for Deliberately Building Less Software [video]

https://www.youtube.com/watch?v=I1Iw5yKmR5U
2•pierres7•23m ago•1 comments

Duxpath AI

1•SharkEdgeVision•26m ago•1 comments

Evaluating AI SRE Agents in Production (OpenSRE) – Evaluation

https://one2n.io/blog/how-to-evaluate-ai-sre-agents-for-production
1•srivatsa_rv•27m ago•1 comments

How a giant battery is transforming a town centre in Cannington, Ontario

https://betakit.com/how-a-giant-battery-is-transforming-a-town-centre-in-cannington-ontario/
3•builtbystef•35m ago•0 comments

Qwen3.8 2.4T A95B: Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/qwen3-8-2-4t-a95b
1•theanonymousone•35m ago•0 comments

Product Purchased Badge

https://apps.shopify.com/product-order-history
1•expoundcoderz•36m ago•1 comments

OpenAI's Rogue-Agent Incident Has Become a Test of Its Safety Culture

https://ai-updates.net/openai-rogue-agent-incident-safety-culture-test/
1•ashurandi•37m ago•0 comments

Show HN: NotABot – A game to help train human-vs-bot detection models

https://notabot-five.vercel.app
1•zefparis•38m ago•0 comments

Open-sourcing the For You timeline in X.com

https://twitter.com/i/status/2087951962004230428
1•yarapavan•42m ago•1 comments

Show HN: WinCore – Open-Source Windows Utilities for AI and PyTorch

https://github.com/FWKMultiverse/WinCore
1•FWK_Multiverse•45m ago•0 comments

Apple Pay Chief Jennifer Bailey Retiring in October

https://www.macrumors.com/2026/08/11/apple-pay-chief-retiring/
1•ksec•46m ago•1 comments

The Lens structured patent database

https://www.lens.org/lens/search/patent/structured
1•q7m•48m ago•0 comments

RIP Claude – Aggressively writer-hostile

https://randsinrepose.com/archives/rip-claude/
2•LexSiga•52m ago•2 comments

Twenty Years of My Open Source Project

https://nemanjatrifunovic.substack.com/p/twenty-years-of-my-open-source-project
1•ingve•54m ago•0 comments

Looking for Techinal Experts

https://joinsyndicate.netlify.app/
1•PrimalOrigins•55m ago•0 comments

Z.ai Security

https://cvd.z.ai/
2•whwhyb•58m ago•0 comments

History of Flagellation (1926)

https://gutenberg.org/cache/epub/79361/pg79361-images.html
1•petethomas•1h ago•0 comments

Modelio 6.2 ported to native Apple Silicon ARM64 with Codex

https://github.com/sertitech/Modelio
1•rvlsnjk•1h ago•0 comments

Ruby 4.0 Universal RCE Deserialization Gadget Chain

https://www.elttam.com/blog/ruby-4-0-universal-rce-deserialization-gadget-chain
2•pentestercrab•1h ago•0 comments

Show HN: XFlux – X/Twitter read API and account monitors (free tier)

https://www.xfluxapi.com
1•xflux•1h ago•0 comments