frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

https://www.bloomberg.com/news/articles/2026-08-26/china-s-z-ai-made-ox-alpha-stealth-model-that-rivals-deepseek
66•garo-pro•1h ago

Comments

garo-pro•1h ago
Unfortunately I can't find sources other than this for now but this seems to be legit.
mohsen1•6m ago
> The company on Wednesday confirmed speculation that the Ox Alpha model is a new iteration of its GLM series and said it will release the weights for it tonight, in response to queries by Bloomberg News.

Seems legit.

It's really hard to know how good it is. So much hype around it.

dgellow•37m ago
Do we know the size of the model?
j_maffe•33m ago
Anyone has a link to a report of its capabilities? I can't find a reliable source.
vblanco•24m ago
completely vibes based, but ive been using it to port Mindustry game from Java to C# with agents, and its been working for 50 hours (its 15-20 tks so super slow inference). Its done a fantastic work and its almost finished now. Better results than deepseek flash and gpt luna by a mile on this kind of long term work. Less good than gpt sol or opus. We dont know the param count but my guess is 200-300 range.
daveyoung•17m ago
likely a distilled glm 5.3 that will punch within 20% of that at 2-3x less size. you'll find that capability is typically very jagged on models that are distilled
esskay•24m ago
I'd be interested to know what was going on with it during the public test as there were numerous reports of it improving considerably at tasks it was asked to do early on in the test compared to later in it.
rfoo•20m ago
lol don't shout out the obvious
daveyoung•19m ago
Two potentials from my pov:

1. Just variance in pass@K. If you prompt any model multiple times you'll see a large variance. N=1, but I find chinese open source models have a higher variance than higher-RL'd models like fable/opus.

2. They legitimately shipped a new RL checkpoint over the 7 days, which I find hard to believe.

I am leaning towards 1.

WithinReason•21m ago
Mixed signals, here it's performing below even GPT-5.4 Nano:

https://livebench.ai/

while here it outperforms Fable by a significant margin:

https://oxalpha.com/

but if the latter is true, will people still say it was "distilled" from Fable?

sunbum•17m ago
the 2nd website is not official, just something someone slopped together for some reason.
epolanski•11m ago
GLM 5.3 was a great model, so this would be strange to release a regressed model
ImprobableTruth•7m ago
It's probably GLM 5.3 flash, so weaker but cheaper.
tosh•16m ago
my guess is this is a small model punching way above its weight

on toy benches it made quite a few mistakes but was able to fix all of them on its own

(meaning more tokens, more turns, more tool calls — but same outcome as gpt 5.6 sol)

giamma•6m ago
https://unwall.app/www.bloomberg.com/news/articles/2026-08-2...

Oldinsurancemaps.net is now a Charter Project

https://openstreetmap.us/news/2026/08/oim-charter-project/
55•altilunium•2h ago•7 comments

RAG Is Simpler Than You Think

https://www.lighthousenewsletter.com/p/rag-is-simpler-than-you-think
71•j0selit0•2h ago•36 comments

Show HN: Buslens – where can I get to by bus? (UK)

https://rupertlinacre.com/buslens/
39•RobinL•3h ago•25 comments

Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

https://www.bloomberg.com/news/articles/2026-08-26/china-s-z-ai-made-ox-alpha-stealth-model-that-...
72•garo-pro•1h ago•18 comments

Value Classes Still Need Compiler Sympathy

https://johan-sjolen.github.io/post/compiler-sympathy/compiler-sympathy/
22•lichtenberger•2h ago•11 comments

Apple introduces M6 and M5 Ultra

https://www.apple.com/newsroom/2026/08/apple-introduces-m6-and-m5-ultra-for-a-big-leap-in-perform...
1167•interpol_p•22h ago•1139 comments

Stalking the Wily Hacker: 40 years later – Cliff Stoll [video]

https://www.youtube.com/watch?v=656058JxTM0
121•zoenolan•4d ago•34 comments

FDA authorizes first wearable device that monitors ketone and blood sugar levels

https://www.fda.gov/news-events/press-announcements/fda-authorizes-first-wearable-device-continuo...
419•sunnynagra•16h ago•201 comments

Queryable Executables

https://fzakaria.com/2026/08/24/actually-queryable-executables
206•rguiscard•10h ago•57 comments

OpenAI Jalapeño: Better than Nvidia Blackwell

https://newsletter.semianalysis.com/p/openai-jalapeno-better-than-nvidia
501•bmulholland•21h ago•320 comments

Harvest (IBM 7950): Supercomputer for cryptanalysis at the NSA in the Cold War

https://spectrum.ieee.org/cold-war-codebreaker-nsa-ibm
46•jnord•6h ago•12 comments

New Mac Studio with M5 Max and M5 Ultra

https://www.apple.com/newsroom/2026/08/apple-introduces-new-mac-studio-with-m5-max-and-m5-ultra/
773•interpol_p•22h ago•518 comments

Beyond Recall and the Illusion of Competence

https://var0.xyz/posts/beyond-recall-and-the-illusion-of-competence.html
4•tuxie_•1h ago•0 comments

Black hole singularity is a surface not a point

https://arxiv.org/abs/2608.21590
260•raattgift•18h ago•187 comments

New Mac mini, featuring M6 and M5 Pro

https://www.apple.com/newsroom/2026/08/apple-unveils-a-more-powerful-mac-mini-featuring-the-all-n...
503•runako•22h ago•317 comments

Maiao: Gerrit-style code review workflow for GitHub, GitLab, Gitea, others

https://github.com/runetes/maiao
87•zdw•12h ago•51 comments

Social media use on the rise among Australian under-16s after ban: data

https://www.france24.com/en/live-news/20260826-social-media-use-on-the-rise-among-australian-unde...
22•giuliomagnifico•1h ago•20 comments

When str.lower() is a security vulnerability in Python

https://sethmlarson.dev/when-str-lower-is-a-security-vulnerability
133•rbanffy•14h ago•56 comments

U.S. gov't moves to suppress pushback on data centers

https://www.tomshardware.com/tech-industry/data-centers/u-s-govt-moves-to-suppress-pushback-on-da...
25•rbanffy•29m ago•5 comments

Nitter and XCancel receive cease and desist notices

https://github.com/zedeus/nitter/issues/1442
981•Banditoz•18h ago•816 comments

Building a backyard office, the build and cost breakdown

https://www.imkylelambert.com/articles/building-a-backyard-office-the-build-and-cost-breakdown
356•surprisetalk•20h ago•220 comments

Bomb fishing is wreaking havoc on Indonesia's coral reefs

https://e360.yale.edu/digest/bomb-fishing-coral-reefs
325•speckx•20h ago•171 comments

The Feeling of Power (Asimov, 1958)

https://archive.org/details/1958-02_IF
8•irdc•48m ago•1 comments

Tooltips need a delay, and then they need to skip it

https://blog.master.dev/tooltips-need-a-delay-and-then-they-need-to-skip-it/
186•ibobev•18h ago•56 comments

Run OpenBSD on DigitalOcean for $4/month

https://nil.wallyjones.com/run-openbsd-on-digitalocean-for-4month/
171•speckx•17h ago•78 comments

Agentic Context Management: Memory and Cost as Architecture Problems

https://arxiv.org/abs/2607.21503
59•gdad•8h ago•20 comments

Don't Wordle

https://dontwordle.com/
355•Hbruz0•23h ago•126 comments

C2PA Cameras Do Not Survive Contact with Reality

https://www.da.vidbuchanan.co.uk/blog/android-c2pa.html
162•Retr0id•15h ago•106 comments

How credit card rewards became a $9.2B wealth transfer

https://www.library.hbs.edu/working-knowledge/how-credit-card-rewards-became-multibillion-dollar-...
191•conbrian•23h ago•338 comments

More than half of adults in U.S. say they lack basic statistical understanding

https://www.psu.edu/news/research/story/more-half-adults-us-say-they-lack-basic-statistical-under...
110•giuliomagnifico•5h ago•156 comments