frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Lathoa, a math app for kids where the AI is wrong on purpose

https://lathoa.ai/en
11•thanouil1411•8h ago
I made this for kids around 10 to 14. A robot called Errol solves a math problem step by step and one of the steps is wrong. The kid has to find it and say what's wrong with it. Sometimes nothing is wrong, so just saying "there's a mistake" every time doesn't work. The user needs to enter an explanation if she finds an error to gain more XP; speed matters also for more points. There is no direct interaction or chatting with an LLM. Lathoa's harness is stable and has many evaluation steps to catch inconsistencies and prompt injections.

You can play one on the homepage without signing up.

The part that surprised me: it's hard to get an LLM to be wrong on purpose. Half the time it gives you the right answer and calls it wrong, or a "mistake" that's actually correct. So every case gets checked before a kid sees it. Where it can, a plain arithmetic check redoes the math exactly. A second model also solves the problem without seeing Errol's work. If anything disagrees, the case is thrown away.

The weak spot is that the second model can make the same mistake as the first. The arithmetic check is there for that, but it only works on English cases so far. German and Greek write decimals with a comma and I haven't got the parsing right yet.

What I'd really like to know: does finding someone else's mistake teach anything that solving the problem yourself doesn't? I'm not sure, and I'd like to hear from people who teach.

Comments

satisfice•25m ago
It helps us practice critical thinking. What it doesn’t do is help us practice knowing WHEN to use critical thinking.

Gemini 4 Argon

https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/
793•bradleyg223•3h ago•526 comments

The top secret URSALA, RAQUEL, and FARRAH satellites

https://www.thespacereview.com/article/4951/1
65•Bluestein•1h ago•8 comments

Surprisingly complex waves reveal the brain's inner workings

https://www.quantamagazine.org/surprisingly-complex-waves-reveal-the-brains-inner-workings-20260930/
98•ibobev•4h ago•35 comments

EDG C++ front-end goes public

https://edgcpp.org/#transition
116•iandinwoodie•3h ago•41 comments

Launch HN: Magnitude (YC S25) – Self-optimizing inference engine for agents

https://github.com/magnitudedev/magnitude
115•anerli•5h ago•49 comments

Why the Bronze Age Collapsed

https://www.worksinprogress.news/p/why-really-caused-the-bronze-age
42•AnodicElegy•1d ago•29 comments

Singapore govt dating app uses Gale-Shapley stable marriage algorithm

https://twitter.com/tuakdotsol/status/2105105417760391258
111•rzk•13h ago•37 comments

Halfspace experimental IDE for solid modeling with distance fields

https://www.mattkeeter.com/projects/halfspace/
50•luu•3h ago•6 comments

5x faster Edge Functions: V8 isolates to Firecracker MicroVMs

https://www.netlify.com/blog/edge-functions-firecracker-microvms/
92•jbott•4h ago•35 comments

Doing a Machine Learning PhD While Working in Japan

https://www.tokyodev.com/articles/doing-a-machine-learning-phd-while-working-in-japan
29•pwim•15h ago•6 comments

Dear Software Makers

https://blog.jim-nielsen.com/2026/dear-software-makers/
48•speckx•3h ago•23 comments

A brief history of the Bloomberg terminal

https://spectrum.ieee.org/bloomberg-terminal
203•rbanffy•8h ago•81 comments

Coltrane's Tone Circle

https://jtomschroeder.com/blog/tone-circle/
15•jtomschroeder•8h ago•1 comments

You said no MCP

https://earendil.com/posts/you-said-no-mcp/
581•yarapavan•13h ago•329 comments

CS240 AI Cheating Retrospective

https://turkeyland.net/thoughts/ai.php
68•ArchAndStarch•3h ago•44 comments

Functional Ultrasound Imaging (fUSI) from scratch

https://www.neuroai.science/p/functional-ultrasound-imaging-from
17•pminimax•3h ago•3 comments

CHOMPI portable sampler instrument is now open-source (hardware and software)

https://www.chompiclub.com/opensource
22•lashkari•5h ago•3 comments

Bild AI (YC W25) Is Hiring a Founding Product Engineer

https://www.ycombinator.com/companies/bild-ai/jobs/dAbC3Gd-founding-product-engineer
1•rooppal•6h ago

Before pixels: Modular industrial dashboards

https://unsung.aresluna.org/before-pixels-modular-industrial-dashboards/
27•leephillips•4h ago•6 comments

I could've accessed 17T Microsoft records

https://blog.faav.net/how-i-couldve-accessed-17-trillion-microsoft-records
235•luispa•2d ago•103 comments

Show HN: Lathoa, a math app for kids where the AI is wrong on purpose

https://lathoa.ai/en
14•thanouil1411•8h ago•1 comments

Great Dirhombicosidodecahedron ("Miller's Monster")

https://www.software3d.com/MillersMonster.php
21•cobbzilla•6h ago•1 comments

The last time my family was replaced by technology

https://manuel.darcemont.fr/posts/the-last-time-my-family-was-replaced-by-technology/
159•megalomanu•10h ago•391 comments

What TLA+ can and can't check

https://buttondown.com/hillelwayne/archive/what-tla-can-and-cant-check/
123•b-man•9h ago•29 comments

Burning Man death rates – A short lesson in statistics

https://ihavenapkinthoughts.substack.com/p/burning-man-death-rates-a-short-lesson
90•viraj_shah•2d ago•112 comments

SDF vs. MSDF vs. Slug: GPU Text Rendering

https://alphapixeldev.com/sdf-vs-msdf-vs-slug-vs-rive-gpu-text-rendering/
124•ibobev•9h ago•49 comments

Responsible Release of AI-Generated Mathematics

https://agmai.org/general-sep29/
68•aureianimus•20h ago•78 comments

Gitea 28.0

https://blog.gitea.com/release-of-28.0.0/
58•porridgeraisin•2h ago•23 comments

Gemini 4 Argon (High): Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/gemini-4-argon
64•theanonymousone•2h ago•28 comments

Commit description as a thinking tool

https://yedhu.me/posts/commit-description-as-a-thinking-tool/
104•yedhukrishnan•5h ago•60 comments