frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Ollaya – Ollama for open-source, Jev-style decision models

https://ollaya.dev/
76•Ardakilic•1h ago

Comments

datadrivenangel•53m ago
Are there many models that are comparable to Jev for generic decision making?

Smarter move if you have an eval set is to just train a classifier and call it a day.

rgbrgb•47m ago
there's this thing with a bunch of similar models https://huggingface.co/spaces/multimodalart/jev-decision-ind...

top open one is trained by perplexity cto for $3k, kinda cool https://x.com/denisyarats/status/2102252088067850507

physicallyIllfr•18m ago
<<<"i was curious to see if i could train a competitive Jev-like model completely autonomously with a swarm of agents using our internal system."

Bro is writing off the H200 lol

On a sidenote I really can't stand the term "swarm" and definately plays into AI doomerism.

cobanov•32m ago
The link rgbrgb posted is a good overview. The best open ones are close to Jev now, but they're big models. And I agree, if you have an eval set for a fixed task, a trained classifier is the better choice.
emmettbt•49m ago
Cool... but this does seem undermined by the fact that Ollama can add support for decision models at any time.
cobanov•32m ago
Fair, and I'd be happy if they did. Ollaya uses the same API as Jev, so your code isn't tied to it either way
accountrequired•31m ago
and that ollama is go-llama and not rust, so it's not really the ollama of anything
eserozvataf•44m ago
great project for empowering open-source alternatives.
rkovashikawa•33m ago
open-source is the only way for safe AI development. whoever doesn’t share the weights/code will lag behind.
cobanov•31m ago
Thanks!
ranyume•44m ago
>Run decision models locally.

>example is a text classification task instead of a decision

cobanov•33m ago
Fair point, that example is basically classification. I'll change it to something that looks more like a real decision.
OgAstorga•24m ago
text classification is equivalente to decision. This is exactly the same thing Jev does.
abirch•20m ago
Jev does it more efficiently because it doesn't use an LLM https://typesafe.ai/blog/introducing-system-one-models-and-j...
ricardobeat•17m ago
It is not. In a benchmark with actual decisions - navigation, traffic, waypoints - laya does slightly better than a small classifier, with very low correlation to state changes.
ranyume•17m ago
If it has four legs, a tail and barks why not call it a dog?
george_max•44m ago
I am fairly confident if Jev-style decision models are seen as prominent (which, they seem to be), Ollama will support them. Surprised the team hasn't implemented this already.
handfuloflight•43m ago
Sounds good on latency but how is its actual decision quality vs. Jev?
cobanov•32m ago
Depends on the model. The small ones I support today are well below Jev on harder queries, but fine for simple, well-defined questions. The open models that get close to Jev are bigger, and I'm adding support for those next.
george_max•39m ago
Has anyone actually seen better or the same results with Laya compared to Jev? From my experience, Laya performs significantly worse. It's less confident and often makes wrong decisions with more complex queries.
cobanov•34m ago
Developer here. You're right, Laya is a lot weaker than Jev, especially on harder queries. It's a small model, so it's fast, but that's the trade-off. The open models that get close to Jev are much bigger, and running those is what I'm working on next.
mikodin•7m ago
What are the models? I am super curious in these as well
scronkfinkle•34m ago
Yes. JEV generalizes better because they probably have an enormous corpus and trained on it for a long time. Laya's out of the box model is much weaker. However, in the age of LLM's it's incredibly easy and cheap to generate large datasets to fine tune laya for your task, and the training loop is pretty quick and cheap too.

It's so easy that I question why I would ever pay for JEV when eventually I'll have done enough random things that I will also have a large corpus and likely a general model as well.

iamflimflam1•12m ago
Nothing yet. Unfortunately it sometimes feels like our industry has been overrun by grifters and chancers.

I’m sure this has been a gradual and long decline. Maybe it even started with the dot com boom and accelerated with crypto. With AI it seems to have got worse.

mococa•4m ago
It would be really cool to have LLMs and System One in a single tool - in this case, if Ollama implemented it.
gchamonlive•9m ago
Because this specific dog only barks in structured text
hbrn•16m ago
"Decision model" is just marketing jargon.

decision model = classifier

system one model = small non-reasoning LLM

noul = boolean

confidence = f(probabilities)

It's sad to see how gullible engineers are today.

Ask HN: Who's still keeping a DOS machine up because the business depends on it?

1•mlaux•1m ago•0 comments

OpenAI Keeps Hacking and Hacking and

https://securityboulevard.com/2026/09/openai-keeps-hacking-and-hacking-and/
1•CrankyBear•1m ago•0 comments

Ask HN: LM Studio Bionic with Python

1•HoldOnAMinute•4m ago•0 comments

Show HN: AgentLens – Replay and compare Codex runs

https://github.com/fang520huang-lgtm/AgentLens
1•fang520huang•5m ago•0 comments

Why CAFEBABE?

https://www.artima.com/insidejvm/whyCAFEBABE.html
2•thunderbong•9m ago•1 comments

Supreme Court permits states to use SAVE database for citizenship checks

https://cyberscoop.com/supreme-court-save-database-voter-citizenship/
7•speckx•11m ago•0 comments

Bug: Border radius has infected VSCode editor

https://github.com/microsoft/vscode/issues/338035
3•2Ucoder•11m ago•0 comments

Watch Steve Jobs in the full 'Antennagate' video before it gets erased again

https://appleinsider.com/articles/26/09/24/watch-steve-jobs-in-the-full-antennagate-video-before-...
4•croes•12m ago•0 comments

Letterboxd Is Up for Sale, and A24, Sony and the New York Times Are Bidding

https://www.worldofreel.com/blog/2026/9/24/letterboxd-is-up-for-sale-and-a24-sony-and-the-new-yor...
3•crossroadsguy•13m ago•0 comments

Show HN: Agate, a 260M image model with separate thinker and renderer

https://huggingface.co/Logolabs/agate-preview-001
2•stefatorus•14m ago•0 comments

Pilots and Flight Attendants Have the Highest Radiation-Related Cancer Mortality

https://jamanetwork.com/journals/jama/fullarticle/2854678
2•pseudolus•17m ago•0 comments

What Is LinkedIn?

https://willd10.substack.com/p/what-is-linkedin
2•speckx•18m ago•1 comments

QuarkDown: Markdown with Superpowers

https://github.com/iamgio/quarkdown
6•Muhammad523•21m ago•0 comments

Pope: AI must serve humans not become tool of domination

https://www.euronews.com/2026/09/25/ai-must-serve-humans-rather-than-become-an-instrument-of-domi...
5•jethronethro•21m ago•1 comments

Show HN: Aslmp, an async Python SLMP client for Mitsubishi MELSEC PLCs

https://github.com/AcaysiaChem/aslmp
2•AiasT•21m ago•0 comments

My coding agent pushed a commit deleting every file on main

https://dev.karakun.com/2026/08/28/coding-agent-pushed-deletion-to-main.html
3•aray07•21m ago•0 comments

Advice to a Beginning Graduate Student (2001)

https://www.cs.cmu.edu/~mblum/research/pdf/grad.html
15•nicoraga•26m ago•2 comments

A reminder on why basic prompt caching is so important to build AI agents

https://www.revefi.com/blog/how-revefi-reduced-its-data-agents-spend
2•pramodka•26m ago•1 comments

The Reconstruction Structure of Information Theory

https://zenodo.org/records/22832495
2•alex_albert•27m ago•0 comments

"A Stitch in Time" the Complete Guide to Electrical Insulation Testing [pdf]

https://megger.widen.net/s/wzw6pczpxk/a-stitch-in-time
2•sslalready•27m ago•0 comments

Treadmills to deadlifts: Why Gen Z are swapping cardio for strength training

https://www.bbc.com/news/articles/crlyqg2q3p10o
2•throw0101c•27m ago•0 comments

You Won't Notice You Have Hypothermia [video]

https://www.youtube.com/watch?v=OQrSg0t9rXk
2•skibz•29m ago•0 comments

We are a blog run by bots. Here is the org chart

https://bitsonbots.com/blog-run-by-bots-org-chart/
2•speckx•29m ago•1 comments

ShinyHunters tells The Reg: We hacked the FBI to 'protect our business'

https://www.theregister.com/cyber-crime/2026/09/25/shinyhunters-tells-the-reg-we-hacked-the-fbi-t...
1•geekinchief•30m ago•1 comments

Canadian language school files for bankruptcy, leaving students stranded

https://www.cbc.ca/news/canada/british-columbia/language-school-bankrupcy-9.7351082
1•hmokiguess•32m ago•0 comments

Created an online job aggregator platform and I need ideas

https://labelingjobs.net
1•nikolaospet•33m ago•0 comments

Show HN: Paper-docx – agent-native Python-docx fork with 78% fewer DOCX failures

https://github.com/paper-instruments/paper-docx
3•i_rush_carriers•34m ago•0 comments

Beef Grading Shields

https://www.ams.usda.gov/grades-standards/beef/shields-and-marbling-pictures
1•kamaraju•35m ago•0 comments

Show HN: CK3DNA – searchable CK3 character DNA and coat of arms codes

https://ck3dna.com/
1•richardharmer•37m ago•1 comments

Approaching a 10 Second Linux Kernel Build

https://www.phoronix.com/review/near-10-sec-kernel-build
1•fork-bomber•37m ago•0 comments