frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Any text-to-SQL benchmark should address difficulties of real-world data stores

https://cacm.acm.org/blogcacm/if-you-think-you-can-do-real-world-text-to-sql/
17•shenli3514•1h ago

Comments

m12k•29m ago
Friend of mine built a startup around this, allowing non-techies to "query their database using natural language": https://www.blazesql.com/ Not sure how he achieved it (if TFA is to be believed), but it's my impression that his query generation and results are fairly robust.
programmertote•27m ago
Interesting... at my current job, my team and I are solving this problem of how to make sure LLM understand our SQL data warehouse to answer analytical questions for clients. We have a 20-year old database with rotting schema as the blog post describes. So we had to rebuild the database in a way that is well structured and governed. Once that hard work is done, we slap dbt models on top of business metrics with model YAML files (and some common MD files) carrying a lot of semantic and metadata info for them.

Then our software engineering team ingest the dbt models (we have to tactically create dbt models; that is, always think "what would make LLM hallucinate less" as we are implementing them) and info from the semantic layer to build context for the LLM, and use that to answer analytical questions. So far, it's been promising. The accuracy isn't zero like the blog's author suggested though. We have built like 30 metrics in dbt and semantic layer in the last quarter, and asked the research and analytics teams to do internal testing on the LLM app. I will find out how accurate this approach is from the feedback soon.

semiquaver•23m ago

  > It is widely known that an LLM can only find data it has seen before.
What does this mean? It’s obviously not literally true, so what is the author trying to convey?

Medici family mystery may be solved after more than 400 years

https://www.cnn.com/2026/07/15/science/medici-family-mystery-dna-malaria
35•effects•1h ago•4 comments

GigaToken: ~1000x faster Language model tokenization

https://github.com/marcelroed/gigatoken/
307•syrusakbary•5h ago•57 comments

Quality non-fiction books are the antithesis of AI slop

https://resobscura.substack.com/p/quality-non-fiction-books-are-the
40•benbreen•8h ago•23 comments

Terrence Tao's ChatGPT Conversation about the Jacobian Conjecture Counterexample

https://chatgpt.com/share/6a5fdc7a-d6f8-83e8-bbea-8deb42cfed56
496•gmays•5h ago•299 comments

Malleable Computing, Emacs, and You

http://yummymelon.com/devnull/malleable-computing-emacs-and-you.html
42•kickingvegas•1h ago•5 comments

Safari Technology Preview 248 Released

https://webkit.org/blog/18162/release-notes-for-safari-technology-preview-248/
51•Erenay09•2h ago•11 comments

Show HN: Bento - An entire PowerPoint in one HTML file (edit+view+data+collab)

https://bento.page/slides/
578•starfallg•7h ago•140 comments

Are AI Labs Pelicanmaxxing?

https://dylancastillo.co/posts/pelicanmaxxing.html
328•dcastm•5h ago•131 comments

John C. Dvorak has died

https://twitter.com/na_announce/status/2079952538040672302
406•coleca•3h ago•115 comments

Everyone Should Know SIMD

https://mitchellh.com/writing/everyone-should-know-simd
183•WadeGrimridge•5h ago•55 comments

Any text-to-SQL benchmark should address difficulties of real-world data stores

https://cacm.acm.org/blogcacm/if-you-think-you-can-do-real-world-text-to-sql/
18•shenli3514•1h ago•3 comments

Show HN: Cactus Hybrid: We taught Gemma 4 to know when it's wrong

https://github.com/cactus-compute/cactus-hybrid
18•HenryNdubuaku•5h ago•3 comments

Nobody knows what a used GPU cluster is worth

https://ciphertalk.substack.com/p/nobody-knows-what-a-used-gpu-cluster
146•rbanffy•1w ago•122 comments

Making

https://beej.us/blog/data/ai-making/
247•erikschoster•7h ago•103 comments

Launch HN: Unlayer (YC W22) – Add email and document builders to your app

https://unlayer.com
43•adeelraza•7h ago•22 comments

Fairphone 6 wide camera experimental Linux support

https://nondescriptpointer.com/articles/fairphone-6-wide-camera-linux/
31•helonaut•2h ago•0 comments

The startup's Postgres survival guide

https://hatchet.run/blog/postgres-survival-guide
286•abelanger•10h ago•162 comments

I Inspected My Take-Home Interview Project. It Was a Whole Operation

https://citizendot.github.io/articles/fake-job-interview-git-hook-malware/
199•CITIZENDOT•2h ago•43 comments

How we made our LeRobot video reader up to 15× faster

https://www.eventual.ai/blog/how-we-made-our-lerobot-video-reader-up-to-15x-faster
6•ykev•5d ago•0 comments

Taking OCaml and Eio for a Spin

https://mattjhall.co.uk/posts/taking-ocaml-eio-for-a-spin.html
25•mattjhall•2d ago•2 comments

So Reddit has decided that plain HTML is unsafe

https://www.cole-k.com/2026/07/21/reddit/
195•montroser•10h ago•194 comments

Businesses with ugly AI menu redesigns

https://blog.fiddery.com/businesses-with-ugly-ai-menu-redesigns/
158•speckx•10h ago•123 comments

Nvidia DGX Spark as a daily driver

https://daniel.lawrence.lu/blog/2026-07-15-dgx-spark-as-daily-driver/
68•plun9•3d ago•54 comments

Fretboard Memorisation with Modular Arithmetic

https://ohaodha.ie/blog/fretboard-memorisation-with-modular-arithmetic/
4•ohaodha•52m ago•0 comments

Ghost Cut – or why Cut and Paste is broken everywhere

https://ishmael.textualize.io/blog/ghost-cut/
112•willm•8h ago•80 comments

Back to Kagi

https://blog.melashri.net/micro/back-to-kagi/
171•speckx•9h ago•146 comments

Why do we love music? (2018)

https://pmc.ncbi.nlm.nih.gov/articles/PMC6353111/
13•jstrieb•2d ago•1 comments

Can a MUD evaluate LLMs? A $99 proof of concept

https://cruciblebench.ai/
91•Davisb135•7h ago•57 comments

Show HN: ValuePair – a friendship app that cares about values first

https://valuepair.app
3•zloy88•42m ago•2 comments

Which streaming service was that on again?

https://www.timwehrle.de/blog/which-streaming-service-was-that-on-again/
39•weetii•8h ago•59 comments