frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Mage a Lightweight, Research-Friendly Multimodal Model Family

https://microsoft.github.io/Mage/
1•nmstoker•2m ago•0 comments

Welteislehre: World Ice Theory

https://en.wikipedia.org/wiki/Welteislehre
1•Teever•4m ago•0 comments

Why Carbon Capture Can't Conceivably Solve Climate Change

https://projects.propublica.org/why-carbon-capture-cant-solve-climate-change/
1•gmays•7m ago•0 comments

Underwater Oxygen Loss Threatens Earth's Stability, Researchers Warn

https://scripps.ucsd.edu/news/underwater-oxygen-loss-threatens-earths-stability-researchers-warn
2•littlexsparkee•8m ago•0 comments

Data Centers Approved /Solar Farms Rejected -Virginia

https://virginiamercury.com/2024/12/03/data-centers-approved-solar-farms-rejected-what-is-going-o...
1•mistercaz•10m ago•0 comments

Orchid – The assistant that never sleeps

https://orchid.ai
1•lucknite•11m ago•0 comments

AI company employees petition US Government for regulation

https://www.engadget.com/2225612/ai-company-employees-petition-us-government-for-regulation/
3•Larrikin•15m ago•0 comments

Offer rates for tech jobs fell from 51% to 39% since 2015

https://www.interviewquery.com/p/tech-interview-offer-rates-lowest-12-years
19•racketracer•16m ago•2 comments

Show HN: My recommendations after reading 1,500 books

https://brianhama.com
1•brianhama•16m ago•0 comments

1945 Empire State Building B-25 crash

https://en.wikipedia.org/wiki/1945_Empire_State_Building_B-25_crash
1•SockThief•19m ago•0 comments

United States Occupation of Haiti

https://en.wikipedia.org/wiki/United_States_occupation_of_Haiti
1•SockThief•20m ago•0 comments

TUIParts

https://tuiparts.sh
2•handfuloflight•20m ago•0 comments

FIFA plan for Kushner-backed $20B operation to run World Cup meets UEFA fury

https://apnews.com/article/world-cup-fifa-investors-kushner-infantino-uefa-8345e0864e3a2176327338...
1•petethomas•20m ago•0 comments

Show HN: Mustuse.ai – Open-source automated ranking system run by agents

https://mustuse.ai
1•daimajia•21m ago•0 comments

SessionRadar – Monitor Claude Code and Codex sessions from your menu bar

https://sessionradar.com
1•mariohercules•21m ago•0 comments

My girlfriend broke up with me – This is my body's reaction to the news

https://old.reddit.com/r/dataisbeautiful/comments/1v874lz/my_girlfriend_broke_up_with_me/
1•refp•21m ago•2 comments

Where are my iCloud files?

https://bombich.com/blog/2026/07/28/local-backups-of-cloud-storage
1•chmaynard•22m ago•0 comments

Northstar Browser 1.0.5

3•andreasrosdal•22m ago•0 comments

Sam Altman on AGI, Compute, and Human Agency [video]

https://www.youtube.com/watch?v=XDB5beon4DY
2•hakkikonu•22m ago•0 comments

Workplaces look for cheaper AI as 'tokenmaxxing' fades as a corporate fad

https://apnews.com/article/ai-token-openai-anthropic-corporate-31bb80ac1cd7862d05f6397177d826b1
1•petethomas•23m ago•0 comments

Show HN: A tool to privately timestamp a claim and reveal it on schedule

https://prove-yourself.sebmellen.com/
1•sebmellen•23m ago•2 comments

NIH Will End Dangerous Gain-of-Function Research

https://www.wsj.com/opinion/jay-bhattacharya-how-nih-will-end-dangerous-gain-of-function-research...
1•LostMyLogin•23m ago•0 comments

New York school pauses plan to deploy humanlike AI robot teacher after backlash

https://apnews.com/article/new-york-school-artificial-intelligence-robot-teacher-c2126c704104c4eb...
2•petethomas•25m ago•0 comments

Dice: Detailed Inter-Chiplet End-to-End PHY Modeling for Chiplet Simulation

https://arxiv.org/abs/2607.24221
1•Jimmc414•28m ago•0 comments

Substackers Say New AI Detection Tool Is a 'Witch Hunt'

https://www.404media.co/substackers-say-new-ai-detection-tool-is-a-witch-hunt/
2•Jimmc414•28m ago•0 comments

Poolside Desktop Assistant – multi harness ACP app for macOS

https://poolside.ai/blog/introducing-poolside-desktop-assistant
1•appleton•30m ago•0 comments

Show HN: I was tired of opening 2 tabs for every HN link, so I made a userscript

https://github.com/twalichiewicz/HNewhere
7•twalichiewicz•31m ago•1 comments

I Asked Gemini to Read 300-Year-Old Portuguese Parish Records

https://mariofilho.com/gemini-portuguese-parish-records/
1•mariofilho•31m ago•0 comments

Architects, Anthills, and AI: A Nobel Prize's Lesson on Scientific Progress

https://kovashikawa.github.io/ai/architecs-anthill-ai/
1•rkovashikawa•32m ago•0 comments

Artist sues AI meme generator for selling deeply personal comic as ad template

https://arstechnica.com/tech-policy/2026/07/artist-sues-ai-meme-generator-for-selling-deeply-pers...
2•inerte•32m ago•1 comments
Open in hackernews

Ask HN: LLM is useless without explicit prompt

4•revskill•1y ago
After months playing with LLM models, here's my observation:

- LLM is basically useless without explicit intent in your prompt.

- LLM failed to correct itself. If it generated bullshits, it's an inifinite loop of generating more bullshits.

The question is, without explicit prompt, could LLM leverage all the best practices to provide maintainable code without me instruct it at least ?

Comments

ben_w•1y ago
Your expectations are way too high.

> - LLM is basically useless without explicit intent in your prompt.

You can say the same about every dev I've worked with, including myself. This is literally why humans have meetings rather than all of us diving in to whatever we're self-motivated to do.

What does differ is time-scales of the feedback loop with the management:

Humans meetings are daily to weekly.

According to recent research*, the state-of-the-art models are only 50% accurate at tasks that would take a human expert an hour, or 80% accurate at tasks that would take a human expert 10 minutes.

Even if the currently observed trend of increasing time horizons holds, we're 21 months from having an AI where every other daily standup is "ugh, no, you got it wrong", and just over 5 years from them being able to manage a 2-week sprint with an 80% chance of success (in the absence of continuous feedback).

Even that isn't really enough for them to properly "leverage all the best practices to provide maintainable code", as archiecture and maintainability are longer horizon tasks than 2-week sprints.

* https://youtu.be/evSFeqTZdqs?si=QIzIjB6hotJ0FgHm

revskill•1y ago
It's not as high as you think.

LLM failed at the most basic things related to maintainable code. Its code is basicaly a hackery mess without any structure at all.

It's my expectation is that, at least, some kind of maintainable code is generated from what's it's learnt.

ben_w•1y ago
Given your expectation:

> It's my expectation is that, at least, some kind of maintainable code is generated from what's it's learnt.

And your observation:

> LLM failed at the most basic things related to maintainable code. Its code is basicaly a hackery mess without any structure at all.

QED, *your expectations* are way too high.

They can't do that yet.