Ask HN: Are You Using Finetuning?

3•nate•1h ago

How? For what?

fintuneing seems to be out of fashion (if it were really ever in fashion), but I still see folks like Karpathy mention reaching for it as a tool.

But is anyone in any business capacity on here doing that? Are you finetuning any remote LLM or something self-hosted? What for?

I’m just curious where the line is of “oh this is better encoded in the models weights rather than in RAG/thinking over context stuff it needs to figure out.

Comments

BoredPositron•1h ago

We mainly do full finetunes on diffusion models and their text encoders like z-image, flux2 klein to adapt them to our clients visual style and train LoRas for people and products. The quality goes up immensely if the model has a better grasp of professional visual terms. Training the right kind of leather or plastic (mainly for the pattern) helps when you are scaling to 12-16k and want 99.9% reproduction, everything becomes a texture at that size and if you don't have them trained it's a mess.

nate•1h ago

Ah. That makes sense. Is this something where you do it once and you are done? Or is it something you re-finetune based on performance or reviews you get back from the client. i.e. Client doesn't like something so you go back for another cycle of

Also, is this something that's a pain in the ass to manage multiple versions of the model? One (maybe more in draft mode) for each client?

BoredPositron•1h ago

We do one finetune on the base model to iron out a few of its problems, like plastic skin and its poor understanding of visual terms and reproduction. It also really helps it understand the normal maps we use for perspective templating.

What we are mostly producing are LoRAs, and we put them through a staged training process. The first stage is all about the textures, the second stage focuses on the product itself, and the last stage dials in the exact perspectives we need.

Despite what the research out there says, we actually get better results sticking with LoRAs instead of LoKRs. The pain is generating the dataset because you have to adapt it for every product. The actual training is basically just fire and forget.

S3 Files

What AI is Doing to the Workforce [video]

Where are we converging with all AI companies running the same business model

Show HN: Hive – full dev workspace using (Kanban/chat mode,multi-repo,agent-SDK)

Eyes on the Far Side of the Moon

The V32 numbers station: a mysterious Cold War spying system revived in Iran

SCOTUS overturns 5th Circuit ruling that told ISP to kick pirates off Internet

Ask HN: What is the most useful website you've discovered?

Give Your AI Eyes: Introducing Chrome DevTools MCP

Will you run one reproducibility check? (Voynich structural analysis)

Show HN: I built an AI coding agent 50% cheaper than Claude Code (same prompts)

Duolingo-style GitHub streak widget for READMEs

Before We Start on Quantum

Visual DLQ monitoring and replay for RabbitMQ

Why IPv6 is the only way forward

Roger Fenton's Valley of the Shadow of Death (1855)

Apple Studio Display XDR Now FDA-Cleared for Diagnostic Radiology Use

Spacebot: Agentic AI system where LLM process has a dedicated role, OpenClaw alt

Hotcopy Coding CLI with no context ceiling and agents that learn across sessions

China is winning one AI race, the US another – but either might pull ahead

Building a framework-agnostic Ruby gem (and making sure it doesn't break)

Another Memory Corruption Case

NRR doesn't have to compress as you scale (data from 37 devtools)

Show HN: A VS Code extension that points tickets based on tech debt

We fix your broken Rhino models

Tailslayer: Library for reducing tail latency in RAM reads

Wireless festival cancelled after Kanye West banned from entering UK

RAM Has a Design Flaw from 1966. I Bypassed It [video]

Show HN: A Little Excursion