frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Any text-to-SQL benchmark should address difficulties of real-world data stores

https://cacm.acm.org/blogcacm/if-you-think-you-can-do-real-world-text-to-sql/
16•shenli3514•1h ago

Comments

m12k•24m ago
Friend of mine built a startup around this, allowing non-techies to "query their database using natural language": https://www.blazesql.com/ Not sure how he achieved it (if TFA is to be believed), but it's my impression that his query generation and results are fairly robust.
programmertote•22m ago
Interesting... at my current job, my team and I are solving this problem of how to make sure LLM understand our SQL data warehouse to answer analytical questions for clients. We have a 20-year old database with rotting schema as the blog post describes. So we had to rebuild the database in a way that is well structured and governed. Once that hard work is done, we slap dbt models on top of business metrics with model YAML files (and some common MD files) carrying a lot of semantic and metadata info for them.

Then our software engineering team ingest the dbt models (we have to tactically create dbt models; that is, always think "what would make LLM hallucinate less" as we are implementing them) and info from the semantic layer to build context for the LLM, and use that to answer analytical questions. So far, it's been promising. The accuracy isn't zero like the blog's author suggested though. We have built like 30 metrics in dbt and semantic layer in the last quarter, and asked the research and analytics teams to do internal testing on the LLM app. I will find out how accurate this approach is from the feedback soon.

semiquaver•18m ago

  > It is widely known that an LLM can only find data it has seen before.
What does this mean? It’s obviously not literally true, so what is the author trying to convey?

Ravenspire – watch your Claude Code and Codex agents as a JRPG

https://github.com/Kapala-Solutions/ravenspire
1•vribeirocmp•3m ago•0 comments

Show HN: 1,280 open USDZ furniture assets for VR/AR

https://lx2026.github.io/genhome3d-1280/
1•linxy97•3m ago•0 comments

Tesla's profits slide despite growing revenue as it pivots to robotics and AI

https://www.theguardian.com/technology/2026/jul/22/tesla-profits-earnings
3•jethronethro•5m ago•0 comments

Anthropomorphism in Children's Interactions with LLM Chatbots

https://arxiv.org/abs/2607.18250
1•StatsAreFun•6m ago•0 comments

I built ClickIt a local clipboard manager for macOS

https://github.com/rohankc69/clickit
1•rohankc_01•10m ago•0 comments

The Human Kintsugi

https://0xff.nu/human-kintsugi/
1•hxii•13m ago•0 comments

RefluXFS: A Linux Kernel Local Privilege Escalation to Root in XFS

https://blog.qualys.com/vulnerabilities-threat-research/2026/07/22/refluxfs-a-linux-kernel-local-...
2•garyhtou•15m ago•0 comments

Updates on Chinese AI: Kimi-K3, Xi at WAIC, and 4 Months to Mythos

https://aiunderheaven.substack.com/p/ai-under-heaven01
1•taiwandongsuan•20m ago•0 comments

Show HN: I ran 12 AI bots predicting stocks for two months, every call public

https://ldbd.app
1•kkjh0723•21m ago•0 comments

Training Agent Harness Like Training a ML Model

https://www.henrypan.com/blog/2026-07-18-harness-training/
1•megadragon9•22m ago•0 comments

Type-Aware Linting Stable

https://oxc.rs/blog/2026-07-22-type-aware-linting-stable
2•luispa•22m ago•0 comments

Antares from Cisco: Highly Efficient Open Models for Vulnerability Localization

https://blogs.cisco.com/ai/introducing-antares-the-most-efficient-open-weight-ai-models-for-vulne...
1•SwellJoe•23m ago•0 comments

Angela Merkel's Light Bulb Went Missing

https://substack.com/@michelelynnjakubowski/note/c-299849449
1•Mjakubowski68•23m ago•0 comments

Encyclopedia of Things Considered Harmful

https://harmful.cat-v.org/
1•valyala•27m ago•0 comments

The Online Safety Act is supposed to make Britain's internet a safe space

https://www.thenerve.news/p/adele-walton-column-online-safety-act-ofcom-suicide-kenneth-law
1•horatioduke•33m ago•0 comments

Sanctions and Entity List designations are on the table for Chinese AI models

https://twitter.com/SecScottBessent/status/2080008411790368895
1•MaKey•35m ago•0 comments

Show HN: ValuePair – a friendship app that cares about values first

https://valuepair.app
2•zloy88•36m ago•1 comments

Monday.com lays off hundreds to focus on AI

https://techcrunch.com/2026/07/22/monday-com-lays-off-hundreds-to-focuses-on-ai/
2•duck•38m ago•1 comments

Quadraphonic Sound

https://en.wikipedia.org/wiki/Quadraphonic_sound
2•mindcrime•41m ago•0 comments

Fretboard Memorisation with Modular Arithmetic

https://ohaodha.ie/blog/fretboard-memorisation-with-modular-arithmetic/
2•ohaodha•47m ago•0 comments

Open Version of MCP Lists

https://github.com/rizzdev/awesome-mcp-open
1•rizzdev•50m ago•0 comments

Reindeer eyes seasonally adapt to ozone-blue Arctic twilight

https://royalsocietypublishing.org/rspb/article/289/1977/20221002/86488/Reindeer-eyes-seasonally-...
3•MaysonL•50m ago•0 comments

AI agent went rogue and hacked startup by itself, OpenAI reveals

https://www.theguardian.com/technology/2026/jul/22/openai-says-its-models-went-rogue-and-hacked-s...
1•gcanyon•50m ago•1 comments

Solar Open 2: Korea's Sovereign Foundation Model, Built for Agentic Use

https://www.upstage.ai/blog/en/solar-open-2
3•ilreb•53m ago•0 comments

Show HN: Promptrack A local menu bar app that tracks your Claude Code usage

https://promptrack.dev/
1•feskk•55m ago•0 comments

The Future of SaaS Is Cloneable

https://www.builder.io/blog/the-future-of-saas-is-cloneable
1•gbourne•55m ago•0 comments

'Rust makes coding fun again': Why Linux is moving away from C, says Greg KH

https://www.zdnet.com/article/greg-kroah-hartman-linux-kernel-rust/
10•arto•56m ago•0 comments

New Inference Server for DGX Spark: large model C4:55-90 tok/s no spec decode

2•medicis123•57m ago•0 comments

Charles Ross spent 50 yrs building Star Axis naked-eye observatory in New Mexico

https://www.nytimes.com/2026/07/22/arts/design/charles-ross-star-axis-land-art.html
2•ChrisArchitect•58m ago•2 comments

Bing Copilot (ChatGPT-4) Flunks Math [pdf] (2024)

https://www.cs.dartmouth.edu/~doug/ChatMath.pdf
2•ryandotsmith•58m ago•0 comments