frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

LLM Classification Is Feature Engineering

https://minimallysufficient.com/posts/llm-classification-is-feature-extraction/
26•minsufficient•55m ago

Comments

ltbarcly3•37m ago
I don't understand the point they are trying to make.

It's very often (always?) the case that something general also solves particular problems.

    A sorting algorithm is an implementation of min()

    A parser also is a syntax checker.

    A route planner is a reachability checker. 

    A computer algebra system is a basic arithmetic calculator.

    A general constraint solver is a Soduku hint maker.

It's true that LLM output can be used as an input to another classifier, this is also true of any classifier. The improvement on top of the straight LLM classification is relatively small, and I would argue that working on the prompt or just including in the prompt for the LLM what features might be useful to consider would likely work even better.

Fundamentally I read this article as: We want to build a simpler, dumbed down clone of Mathematica, so we cobbled together the following pieces... We also needed a way to do arithmetic, so we also include a copy of Mathematica to do basic arithmetic.

michi883•34m ago
I took the point as: don't make the LLM the classifier. Use it to turn messy input into useful features, then let a normal model make the actual decision. That gives you thresholds/calibration you can inspect.

What I'm not sure about is how stable those features are when you switch the underlying LLM or model version.

softwaredoug•24m ago
In my work on LLM as a judge, I prefer to use LLM decisions as features in a downstream classic ML model for the final decision. It works really well

https://softwaredoug.com/blog/2025/01/21/llm-judge-decision-...

xerlait•20m ago
Why does he first ask to label "ironic" or "not", and then answer the feature questions? Wouldn't it be better to reverse the order?
twelfthnight•19m ago
Why not use a text embedder for the unstructured data and concatenate with the structured data?

For example you could freeze most of the layers of the embedder but let the final ones learn. Then you wouldn’t need to do either feature or prompt engineering?

dist-epoch•12m ago
> Calibration / Threshold Control

The amount of thinking is relatively calibrated. Ask an obvious classification, you get an instant answer. Ask a tricky one, much more thinking.

drabbiticus•11m ago
I really wish people would define terms when using math. What is y? What is LLM(x)? Presumably it evaluates to some real number so that it can be fed to the logistic sigmoid function. If it is the logistic function, then why does beta going to infinity matter? It seems to just collapse the output of the sigmoid function to 1 and make the value of LLM(x) meaningless instead of their claim that it recovers the LLM classifier. What is the function I()?

Maybe these are well understood terms in some field? Maybe I'm just lost?

Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA

https://global.fujitsu/en-global/pr/news/2026/09/14-02
270•my123•1d ago•103 comments

Rate limits on GitLab.com are changing

https://about.gitlab.com/blog/rate-limit-change-2026/
43•darkwater•1h ago•32 comments

Artificial intelligence now beats some of the best human forecasters

https://www.economist.com/science-and-technology/2026/09/16/artificial-intelligence-now-beats-som...
31•ddp26•1h ago•16 comments

OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance

https://asiaai.fyi/openai-misalignment-framework-global-governance/
26•ghernando•1h ago•33 comments

CrowdSec Source Code Leak

https://www.crowdsec.net/blog/crowdsec-statement-source-code-exposure
16•eccgecko•1h ago•1 comments

LLM Classification Is Feature Engineering

https://minimallysufficient.com/posts/llm-classification-is-feature-extraction/
27•minsufficient•55m ago•7 comments

Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents

https://skillsync.com
3•cat-whisperer•13m ago•0 comments

One Year of Sponsored Servo Development

https://servo.org/blog/2026/09/15/one-year-of-sponsorship/
277•AshleysBrain•8h ago•122 comments

Vinix – A modern operating system written in V

https://vinix-os.org/
14•hggh•55m ago•5 comments

Whoisinspace.com/

https://whoisinspace.com
19•Egg-Man•36m ago•4 comments

CCC invites all model citizens to 40C3

https://events.ccc.de/en/2026/09/12/40c3-model-citizens/
213•antonly•8h ago•62 comments

Nvidia announces native GPU programming in Rust

https://developer.nvidia.com/blog/introducing-cuda-rust-two-tracks-for-writing-gpu-kernels/
881•nonmaskable•1d ago•349 comments

Show HN: Share your AI Setup, Learn from others

https://mysetup.ai/
58•steveybrown•3h ago•29 comments

hister

https://github.com/asciimoo/hister
4•bookofjoe•10m ago•0 comments

Show HN: Die With Me – Claude and Codex rate limits as AIM away messages

https://diewithme.co/join
3•monijz•10m ago•0 comments

My temporary PHP fix from 2014 has nearly 20M installs. Today I'm deprecating it

https://jakeasmith.com/blog/http-build-url/
268•jakeasmith•1d ago•70 comments

The Relation Between Mathematics and Physics by Paul Dirac (1939)

https://www.damtp.cam.ac.uk/events/strings02/dirac/speach.html
129•rramadass•3d ago•36 comments

Keys Not Included: recovering the signing keys for US driver's license barcodes

https://ryan.science/blog/keys-not-included
252•Ryan5453•13h ago•125 comments

GLM Built Its Own Inference Infrastructure

https://z.ai/blog/glm-built-its-inference-infrastructure
237•whiteros_e•8h ago•198 comments

Show HN: I built a new version of my fun spatial 3D online meeting app

https://flat.social
75•pawelwentpawel•3h ago•48 comments

Grand MS-DOS Gaming General MIDI Showdown

https://blog.johnnovak.net/2023/03/05/grand-ms-dos-gaming-general-midi-showdown/
3•ibobev•2d ago•0 comments

Ask HN: How to recover Google auth after phone stolen?

11•keymasta•18m ago•9 comments

Better Vector Search for Long Documents: Chunking Inside Manticore Search

https://manticoresearch.com/blog/auto-chunking/
66•GloriaVinogrado•6h ago•10 comments

Xiaomi Mimo 2.6 live post-training dashboard

https://mimo.xiaomi.com/rl/
517•krackers•20h ago•147 comments

Lucasart's Afterlife

https://togameforlife.wordpress.com/2023/12/09/on-lucasarts-afterlife/
82•Bondi_Blue•1d ago•40 comments

Online Z3 Guide

https://microsoft.github.io/z3guide/
55•Bluestein•2d ago•14 comments

Cloudflare/Security-Audit-Skill

https://github.com/cloudflare/security-audit-skill
164•donk8r•11h ago•33 comments

Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations

https://github.com/arnegiacomo/fugleramme
2248•arnemunthekaas•2d ago•250 comments

Comparison of Malloc() Algorithms

https://egbert.net/blog/articles/comparison-of-arena-architecture-in-malloc.html
121•egberts1•1d ago•32 comments

Mastering Layout Engines in Graphviz: Dot vs. Neato vs. Twopi vs. Circo

https://guides.visual-paradigm.com/mastering-graphviz-layout-engines-dot-neato-twopi-circo/
6•vismit2000•2d ago•1 comments