frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Ask HN: How do you correct spatial reasoning of LLMs?

5•mstaoru•6h ago
I'm working on a complex multi-part 3D-printed product, and I needed some engineering input from LLMs. I use highest Gemini Pro reasoning levels, Fable, and K3 Max with everything tuned to highest.

The problem I'm facing is mounting a sensor on a barbell sleeve. The sensor is a semi-hollow 100mm cylinder, 60mm diameter, with a 50mm diameter 50mm deep tube cutout for the sleeve. (It's a bit more complex than that.)

I'm trying to brainstorm different clamping mechanisms, and evaluate durability and manufacturability of several designs.

No matter what I try, the output of all LLMs is absolutely wild, especially if you ask to make a technical drawing. Anthropic reliably been the worst, and Gemini, surprisingly, the best.

But none understand what goes where in any detail, besides spewing paragraph and paragraph of wild ideas including rotational cam clamps over a recessed ring TPU collet (apparently, a "brilliant pivot") and stuff like this.

Is this the "unconquered frontier"? Do I need to blow dust off my AutoCAD? :)

Comments

spottedmarley•6h ago
Everybody is working on (or waiting for) world models. Language models are not sufficient for doing work in the meatspace.
krapp•5h ago
It's wild to me that this is even a question, because of how AI brained the entire world has become. I mean, you have a tool that you know works, and a tool that you know doesn't, but you still feel the need to find a way to make the square peg fit into the round hole.

I'm sure there are ways to "correct the spatial reasoning of LLMs" (whatever that even means) but I'd question whether you even need an LLM to begin with. I know that isn't the answer you're looking for, I'm just a slack-jawed knuckle-dragging Luddite. But this seems like the kind of simple, common problem for which established solutions already existed, like "talk to an actual engineer and not a chatbot that can't even do basic math."

And if it's a matter of not wanting to pay someone, consider that if you're already paying for an inadequate simulation of knowledge, you should be willing to pay for actual knowledge.

mstaoru•3h ago
I'm with you here, the same slack-jawed knuckle-dragging Luddite. I'm not happy about the state of the world, but it is what it is.

I am working on it alone and don't have access or extra money to pay an actual engineer. If I can squeeze 10% of what an actual engineer would've told me, for $20/month I'm in.

And it's not an either-or game, I can have some rough, potentially totally stupid ideas and then validate them with an engineer, instead of paying $10k to make it from scratch.

xyzzy123•5h ago
I would consider asking it to model its proposed designs in openscad or build123d (ideally something query-able). Then have it render and examine plausibility / suitability from different angles. Get it to render the part in use also and give instructions to think about forces and motion.

Recommend doing this in a coding harness not a chat box.

The reason I think you might have more success with this is that the model is mostly thinking about the part in words, which it can convert to a part design in CAD in code. LLMs are really good at coding. Also means it can use relative positioning and relationships.

You will be able to iterate more easily, compare things, compute properties, commit to git etc. The process is more reproducible and steerable than generative production of images.

When the LLM can look at renders of the geometry it generated, it’s easier for it to discriminate when it’s producing nonsense like misaligned parts, things that don’t fit, etc. It’s still going to kind of suck, but it will be better. The whole process of code -> render -> inspect forces the model to put up or shut up and provides grounding. Meshes > bloviating.

As far as I know today's LLMs don't have a "visual imagination" but a process like this could be a slow approximation of one. They clearly do have SOME spatial understanding (pelican tests show us that!) but it feels really non-human.

One thing missing from this is kinesthetics. Personally I am mostly not thinking in accurate visuals in mechanical design. I am imagining how the parts feel and kind of how they move and what slips first and what bends and what feels heavy. Imagining what my hands would feel. But I don't think I trust LLMs to evaluate that stuff by writing simulation code yet.

Read HN twice a day for the last decade. Here's my list of S-Tier HN links

34•vivzkestrel•3h ago•5 comments

Empower the people not the AI – self containing OS

6•OnemanBSD•4h ago•0 comments

Ask HN: Who wants to be hired? (August 2026)

147•whoishiring•2d ago•430 comments

Ask HN: Did GitHub remove the stargazers list?

19•reconnecting•10h ago•1 comments

Ask HN: Anyone interested in building a harness-only benchmark?

4•GodelNumbering•7h ago•5 comments

Ask HN: What Happened to Spec-Driven Development?

3•vivekyyy•4h ago•2 comments

Ask HN: Who is hiring? (August 2026)

221•whoishiring•2d ago•254 comments

Ask HN: Show your micro-SaaS / MRR updates (August 2026)

5•genekrapivin•5h ago•0 comments

Blitz Agent Your specialized agent -> https://blitzagent.studio

2•rvey•5h ago•0 comments

Ask HN: How do you correct spatial reasoning of LLMs?

5•mstaoru•6h ago•4 comments

Ask HN: I built bribes.fyi, now I am stuck what to with it

2•neverenderr•6h ago•3 comments

AI Agents for Logistics, Pitfall?

2•srguarapo•7h ago•0 comments

RNet lets users use one AI credit balance across multiple apps [demo]

2•rNetAi•9h ago•0 comments

Ask HN: How can I improve my products? Which one to keep working on?

2•bchhabra2490•10h ago•1 comments

Ask HN: Dear Anthropic, can we please have thought traces back?

7•exabrial•20h ago•6 comments

Robotic Evals

3•andrewlyu•11h ago•0 comments

Tell HN: "Update to iOS 18.7.8" is updating users to iOS 26; Apple PSA

3•lynndotpy•13h ago•1 comments

Do You Think OpenAI Is Apple Circa the 1980s?

2•Taikhoom2010•14h ago•1 comments

Tell HN: Wife used Chat GPT to set up her new biz domain, email, and website

3•jvanderbot•14h ago•1 comments

ArXiv

5•fred123123•1d ago•1 comments

Ask HN: What was your big failure? How did you get around it?

7•jspann•22h ago•5 comments

Ask HN: When do you choose RAG over Fine-Tuning?

4•Harish_0089•6h ago•0 comments

Why remote roles are region specific and not 100% remote?

5•moizrocky1•1d ago•9 comments

Ask HN: What is a good format for a tool to report data to a LLM?

5•michaelmure•19h ago•6 comments

You've reached the end!