frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Xiaomi MiMo v2.6

https://mimo.xiaomi.com/mimo-v2-6
191•volf_•1h ago•81 comments

The NASA/ESA Mars Sample Return mission has been canceled

https://www.science.org/content/article/nasa-s-mars-sample-return-mission-dead
171•Muhammad523•2h ago•123 comments

What Sun got wrong

https://bcantrill.dtrace.org/2026/09/20/what-sun-got-wrong/
425•chmaynard•7h ago•239 comments

Attention is all you have

https://alicegg.tech/2026/09/21/attention
462•zer0tonin•6h ago•133 comments

Transformers Explained Visually

https://poloclub.github.io/transformer-explainer/
41•aray07•1h ago•5 comments

Why does mathmain need an encrypted loader?

https://safedep.io/mathmain-encrypted-loader/
79•abhisek•2h ago•22 comments

The Advisory Group on Mathematics and Artificial Intelligence

https://terrytao.wordpress.com/2026/09/21/advisory-group-on-mathematics-and-artificial-intelligence/
39•digital55•2h ago•16 comments

Divide by Depth for Instant 3D

https://gabrieloc.com/2026/09/15/perspective.html
22•gabrieloc•2d ago•3 comments

In Search of a Compositional Theory of Self-Stabilization

http://muratbuffalo.blogspot.com/2026/09/in-search-of-compositional-theory-of.html
29•matt_d•2h ago•3 comments

Turn off and restrict access to Apple Intelligence features on Mac

https://support.apple.com/guide/mac-help/turn-restrict-access-apple-intelligence-mchlb2e44f94/mac
175•alwillis•3h ago•112 comments

Grok 4.7

https://x.ai/news/grok-4-7
422•meetpateltech•5h ago•341 comments

Apple Copland D11E4 Booting in the Browser

https://www.pagetable.com/300
36•luu•3h ago•10 comments

AI coding has made CI a bottleneck, so we reworked ours to keep up

https://linear.app/now/ci-bottleneck-reworked
58•julian_digital•1h ago•45 comments

Frontier AI on Your Own Hardware

https://timdettmers.com/2026/09/21/dlab-open-source-week/
34•pretext•2h ago•8 comments

Show HN: A website that tracks US food prices every day

https://www.kadoa.com/food-prices
16•hubraumhugo•1h ago•4 comments

US halts flights at busy East Coast airports, says fiber line cut

https://www.reuters.com/world/us/faa-halts-some-us-east-coast-flights-due-communication-issues-20...
128•allanbreyes•2h ago•82 comments

RoboHarm: Do Frontier Robot Policies Refuse Unsafe Instructions?

https://robocurve.org/roboharm/
14•msadowski•2h ago•3 comments

Kev: Tiny Jev-like family of decision models built on top of Qwen3.5

https://github.com/jaredpalmer/kev/tree/main
367•tosh•14h ago•164 comments

Python Workers are now generally available

https://blog.cloudflare.com/python-workers-ga/
150•torutofu•7h ago•21 comments

This Digital Radio Gets Messages to the World’s Remotest Locations

https://spectrum.ieee.org/hermes-shortwave-radio-digital-data
63•SamuraiLion•5h ago•30 comments

How do Traffic Signals Work (2019)

https://practical.engineering/blog/2019/5/11/how-do-traffic-signals-work
46•at1as•5h ago•31 comments

Avoiding the babbling-idiot failure in a time-triggered communication system

https://ieeexplore.ieee.org/document/689473
20•ronfriedhaber•3h ago•6 comments

A restored PDP-11/83 serving this page on 211BSD Unix

http://pdp1173.com/
78•davepl•5h ago•28 comments

Grim Fandango Puzzle Document (1996) [pdf]

http://gameshelf.jmac.org/2008/11/13/GrimPuzzleDoc_small.pdf
345•kelseyfrog•15h ago•88 comments

Fable 5 – Median thinking declined in August

https://twitter.com/Lon/status/2101793422487204027
294•espeed•5h ago•196 comments

Show HN: Foremerge – Catch intent conflicts between parallel coding agents

https://github.com/naw103/foremerge
26•naw103•5h ago•0 comments

Noodle Gallery- Open-source, self-hosted alternative to Google Photos and Immich

https://digitalescapetools.com/tools/noodlegallery.html
47•xabd•6h ago•34 comments

M5 Ultra Mac Studio Review

https://www.macstories.net/stories/m5-ultra-mac-studio-review-the-dream-mac-for-local-ai-agents/
210•piotrgrabowski•7h ago•194 comments

Heretic removes restrictions from language models

https://heretic-project.org/
214•Bluestein•16h ago•88 comments

macOS 27: Workaround to avoid downloading AI models and save storage

https://www.reddit.com/r/MacOSBeta/comments/1vlnf13/workaround_to_avoid_downloading_ai_models_and/
187•ano-ther•7h ago•83 comments
Open in hackernews

Xiaomi MiMo v2.6

https://mimo.xiaomi.com/mimo-v2-6
189•volf_•1h ago

Comments

algoth1•59m ago
Finally a lab that doesn't cheat on the charts
vatsachak•56m ago
Wow, the chinese labs are getting good at advertising model releases. The moat is thin.

Some features of the release I like:

- Demonstration of diverse tasks, such as using a DAW

- Graphs from various benchmarks and price ranges

- Real world use of the model in scientific environments

unpopularopp•53m ago
I've never used worst smartphones than anything from Xiaomi, bloated ad infested borderline malware territory fork of Android. Maybe just me but whenever I see them on HN I just can't think anything good about this company.
platinumrad•50m ago
It's a big company, like Microsoft or Google. Some of their products are good and some are bad.
verdverm•31m ago
ironic to this thread, I have less bloatware and ads since I switched from Verzion to Pixel on Fi (many years ago)

Curious if Verizon / ATT still force apps on your phone, eg. NFL and Amazon apps, Fi service is subpar

InsideOutSanta•35m ago
It's funny, I have the exact opposite reaction. This is probably misguided on my part, but Xiaomi is one of the very few major tech companies that I don't have an immediate strong negative reaction to. Everything I've bought from them, from robot vacuum to mobile phone, has been reasonably well designed, didn't break, and was priced fairly. I also think their car looks badass.

I'm sure they're doing all kinds of terrible things, like all major companies. I just can't help but like them. Also, this model looks great, and I'll give their subscription a shot next month.

algoth1•34m ago
I still have a xiaomi mi 11 lite, my wife has a 15t. The cameras are the best for the price. The way they chove ads down your throat at every opportunity should be illegal though
bel8•9m ago
Which model did you have?

I ask because my wife has the 15T and the camera is better than my iPhone 17 Pro. And while toying around with it I didn't notice any bloat.

Plus hers support native split screen which I kinda need to multitask on the go.

I'm so pissed at how bad Siri is compared to her android phone that I'm thinking about selling the iPhone to get a Huawei Pura Ultra.

DanMcInerney•53m ago
This is a big week. Probably getting next OpenAI and Anthro models, Grok 4.7, Mimo, etc. These open source model releases are why I can't take the "slow down" crowd seriously. I pitted older Mimo, qwen, step, gpt-oss, and other models against each other playing games like Werewolf and Sketch.io-like games where I let them talk shit while they played against each other. Mimo was by far pareto frontier of game-playing for the models that were <$0.15/m input tokens on OpenRouter. Qwen was pareto frontier in the shit talking game though. Qwen's hilarious. https://www.tiktok.com/@clankerfights/video/7642862917582425...
ddxv•53m ago
This looks great in terms of cost and capabilities, truly pushing the frontier forward in terms of open weight light weight models.
omani•53m ago
ah, would you look at that. I was wondering why mimo 2.5 became "dumber" the last weeks. I was speculating they are probably about to release a new version of the model. because the model really acted out a lot. especially the last two weeks. dont know, was just a feeling, highly speculative.

but now I got my "proof".

sandblast•42m ago
I guess that would only be possible if your provider was Xiaomi itself?
omani•34m ago
yes. I use opencode and opencode uses Xiaomi as a provider.
nemothekid•51m ago
Looking at the frontend design examples; why do these models seem to love the "01 - UPPERCASE TEXT" motif. It's everywhere now (see https://try.cloudflare.com/, which has '01 · QUICK TUNNELS', but no "02" anywhere).
sandblast•44m ago
Nice catch!
danvayn•40m ago
My guess is that by function they break down frontend sections or components into pieces and I believe document things for themselves on some level, or purposely are verbose in this way. It is probably also shaped by users and existing web patterns. They probably get reinforced by models the more common they become.
rao-v•50m ago
I know we have strong views on what a truly open model is (open weights, open training data, open training code etc.) but I really like how transparent they’ve been about the training of this model.

The realtime dashboard they shared during training (https://mimo.xiaomi.com/rl/) was an incredible learning and teaching tool for me, and they’ve been unusually comprehensive in sharing details about their methodology (check out that tech report - it's got lots of clever behind the scene tricks like Google or Deepseek writeups) and benchmark scores (even the stuff they didn’t do well on).

If you’re releasing an open model going forward, please consider offering the community more of this transparency!

earthnail•19m ago
Thanks so much for sharing this. As someone who mostly watches from the sideline, can you share what you can see in this dashboard that someone like me can't see? Is it the metrics themselves that they measure (the metrics tab is absurdly detailed), something in the notices, or something else I missed?
verdverm•7m ago
the existence, who else has a live dashboard for the RL late-training?
stymaar•46m ago
Flash[1]: 309B total / 15B activated parameters

Pro [2]:, 1.02T total / 42B activated parameters

[1]: https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Flash-RL

[2]: https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Pro-RL

verdverm•35m ago
curious why the HF pill (on the right) always has inaccurate values
stymaar•27m ago
I noticed the same, and I wonder as well.
verdverm•19m ago
I suspect they are calculating something in the weights or config, I see it pretty consistently with quants
verdverm•30m ago
There's also a Qwen 3.5 9B distill

https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B

gandreani•27m ago
Those this mean they've fine-tuned this Qwen 3.5 9B on output from the V2.6 model?
syntaxing•45m ago
All these new models are such tease for us folks with 128GB of shared memory. Buying another unit now to expand to 256GB is a mortgage payment but it’s getting tempting…
brcmthrowaway•42m ago
Is there a gamechanger around the corner to reduce DRAM requirements?
stymaar•36m ago
n-gram per-layer embeddings[1][2] might be it.

[1] https://sebastianraschka.com/llm-architecture-gallery/per-la...

[2]: See DS 4.1-Flash and Qwen-3.8-Next.

verdverm•34m ago
this is to offload VRAM to DRAM (for GP comment), and makes no difference for URAM
stymaar•30m ago
Am I missing a joke? WTF is URAM?
verdverm•28m ago
unified memory, not sure if anyone uses URAM, I human hallucinated it
zozbot234
bertili•40m ago
They mixed up DeepSeek 4.1 Flash with something else on this page, possibly DeepSeek 4.1 Flash means Gemini 3.8 Flash.
NooneAtAll3•38m ago
does anyone know what unnamed model is on paretto frontier picture right between MiMo 2.5 and 2.6?

so weird to acknowledge someone being on the front edge, but not name it

AnodicElegy•30m ago
Pretty sure that's Luna xhigh.
spwa4•36m ago
As for the stats that everyone wants:

MiMo-V2.6-Flash-310B-A15B roughly GPT-5.6 Luna / Claude 4.9 according to benchmarks MiMo-V2.6-Pro-1.02T-A42B roughly GPT-5.6 Sol / Opus 5 according to benchmarks

Perhaps with IQ2 flash will run on 128G M5?

MisterMunchkin•34m ago
I really liked MiMo 2.5, it was really affordable and actually had vision, unlike DeepSeek. (DeepSeek has only recently added it)

Just tried 2.6 flash on a really niche topic I specialise in and it has done a really good job. They’ve definitely polluted their training data with claudeslop, but looking past the slop there is a decent model.

omani•32m ago
how do you recognize "claudeslop"?
Bluestein•26m ago
It's an honest, load-bearing, simple thing.-
jwpapi•27m ago
In the chart they use "Pareto Line", which I think is wrong. Pareto is 20% effort leading to 80% results. Which could be interpreted as models costing 20% having 80% of peak intelligence, but that’s not what it looks like to me.

It looks like the "Frontier Line" to me, which is also often misinterpreted. frontier does not mean the best models. It means all models that are not strictly dominated, meaning in most cases: Not same price or cheaper and more intelligent.

I personally would like the word frontier to be used with more criterias: Open Weights, per use-case, etc etc. This would make model selection easier, but I understand it’s not an easy thing to do.

hashmush•23m ago
"Pareto" is many things, but here it does indeed refer to the frontier: https://en.wikipedia.org/wiki/Pareto_front
jwpapi•17m ago
Thank you guys. I learned something new.
shmolyneaux•22m ago
This is the Pareto Front [1], rather than the Pareto principle. It's the idea that anything that's more intelligent is more expensive and anything that's less expensive is less intelligent.

[1]: https://en.wikipedia.org/wiki/Pareto_front

abound•21m ago
There are two (or more) concepts named after the same person:

- Pareto efficiency/Pareto curves: Basically the convex hull of points along the edge of a graph, indicating the best tradeoff between the axes. This is what the post is talking about.

- Pareto principle: this is the 80/20 rule you're talking about

lwansbrough•27m ago
Anyone else more excited about Chinese models than American models these days? Big thing for me is affordability.
swingandamiss•24m ago
No, because I'd rather not support our economic and military rivals.
Freedom2•23m ago
Agreed, and also because I support freedom of speech!
girvo•15m ago
Neither the US nor the Chinese companies are on your side then. They both censor, just different topics.

But at least I can run Chinese models locally, and strip a lot of that censorship/refusal.

lwansbrough•22m ago
I'm Canadian so this sentiment has little value in 2026 unfortunately.
ActionHank•10m ago
Also, frankly, as a fellow Canadian it's pretty clear that the biggest "rival" the US has right now is itself. Just passed out in the corner puking on itself shouting about all the foreigners who won't talk to it.
scottyah
user43928•26m ago
I don't trust any of the benchmarks where Opus 5 surpasses Astra or Fable 5.1.

Maybe Terminal Bench 4.0 and ExploitGym are reasonable.

Terminal Bench 4.0

  GPT 6 Astra             59.6
  Claude Fable 5.1        55.1
  Claude Opus 5           49.0
  MiMo-V2.6-Pro           34.9
  MiMo-V2.6-Flash         28.8
  DeepSeek V4.1 Flash     26.8
  MiMo-V2.5-Pro            1.5
ExploitGym

  GPT 6 Astra             42.4
  Claude Fable 5.1        30.4
  Claude Opus 5           22.1
  MiMo-V2.6-Pro           17.8
  MiMo-V2.6-Flash          6.0
  MiMo-V2.5-Pro            0.1
DeepSWE v1.1

  DeepSeek V4.1 Flash     74.2
  Claude Opus 5           74.0
  GPT 6 Astra             74.0
  MiMo-V2.6-Pro           71.9
  Claude Fable 5          70.0
  MiMo-V2.6-Flash         67.9
  MiMo-V2.5-Pro           19.0
mokre•24m ago
Maybe you should not trust any of the benchmarks!
simonw•12m ago
Pelicans for Flash: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

Pelicans for Pro: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

phainopepla2•4m ago
I think we can say pretty confidently they aren't pelican-bench-maxxing
Imanari•3m ago
ish… at least we can be sure they don’t benchmaxx the pelicans lol
brcmthrowaway•3m ago
Just me, or do these look bad?

Qwen3.8-27b pelican was amazing on Mac.

https://www.nudgehost.com/dpjn3uwe

thrownawaysz•12m ago
>Night 0.8x Usage, 00:00-08:00 -UTC+8

It's because offpeak electricity is cheaper?

Funnily it's perfect if you are in the Pacific Time Zone because you can use it daytime 9am to 5pm

alfalfasprout•6m ago
The moat for OAI and anthropic seems to be very quickly shrinking. Chinese labs are now using RSI-like approaches and even without resorting to heavy distillation they're catching up in a couple of months vs. what would have been 6-12 months a year prior.

And as these models get better the pace of training is quickly speeding up too.

This doesn't bode particularly well for anthropic/OAI after they go public.

mydreamof•23m ago
It is a 9B agentic model developed by Xiaomi MiMo through supervised fine-tuning of Qwen3.5-9B on MiMo-generated data
•
17m ago
You can definitely offload n-gram embeddings to storage; they're very sparsely used (only a few KB fetched per token) so this is quite effective. Loading to DRAM only becomes necessary if they are a bottleneck to overall performance (which might happen if you're doing very wide batches and everything else uses super fast VRAM/HBM).
verdverm•14m ago
I was looking at the qwen-next-flash, and the weights would fill my OEM Spark on their own, before the n-gram. I'm unclear if offloading to disk can work here, is that what you are implying is possible?!
girvo•13m ago
Check out eugr’s TP=1 sparkrun recipe :)

It’s an NVFP4 quant, but it fits, and is surprisingly capable.

verdverm•12m ago
do you have a HF link? HF search is not uncovering it for me

(or is it somewhere else)

girvo•14m ago
Nah, I’m streaming ngrams off NVMe on my Spark-alike right now. Works surprisingly well (except for when I accidentally bottlenecked it through my NAS)
verdverm•13m ago
interesting, peer comment seems to indicate this is a possibility as well, will have to take a deeper look
petu•12m ago
n-grams can be kept on SSD, no need to hold them in any kind of RAM (at least w/o batching)
zozbot234•22m ago
You could always stream from SSD storage. Especially effective if you get a cheap old-gen HEDT with lots of PCIe slots to add NVMe storage to and reasonable overall PCIe bandwidth.
verdverm•29m ago
https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B is an option
nextaccountic•4m ago
No, Pareto refers to Pareto efficiency https://en.wikipedia.org/wiki/Pareto_efficiency

What you call "frontier line" is also called "Pareto frontier" https://en.wikipedia.org/wiki/Pareto_front

Your description of it is basically correct though

•
8m ago
Because your country is increasingly owned by Chinese?
lwansbrough•4m ago
Because at present the pedophile US president is making it his mission to molest my country. China, for all its faults (including espionage, which the US is also guilty of) is just trying to conduct trade.
verdverm•4m ago
Half of Canada now uses the word 'enemy' when asked for an adjective to describe America or China. We're equivalent in their eyes now because we elected Trump a second time and all that he has said and done in 2.0
tacomagick•18m ago
Absolutely! Chinese models are both cheaper and more capable in many cases, compared to the American models and their makers continuously fumbling or reducing model capability with each update. Deepseek decreased costs when they released Flash 4.1 you would not see any American company do this, in reverse they would try charge you more.
user43928•9m ago
OpenAI decreased prices with the 5.6 model family.

And later they further cut Sol and Terra pricing by 20% (maybe only in the API) and Luna by 80%.

In fact Luna still outperformed DeepSeek Flash 4.1 in cost per task on Artificial Analysis when I last checked.

However, Luna is slightly less intelligent. I have a feeling that it's pretty dumb and prone to hallucination unless running at xhigh or max effort, where it somehow manages to work quite well.

I did not personally test the open weight models beyond the old Qwen 3.6 27B, which produced unusably bad results for me.

The competition is great, and I hope Chinese models will continue to force leading US labs to offer models at a low price point.

That said, I don't think the Chinese labs have anything over OpenAI and Anthropic when it comes to capability or efficiency - I have no reason not to believe the US labs have even lower cost to serve the models.

verdverm•5m ago
I have a contrarian opinion that China passing America in Ai is the Sputnik moment we need to leave the hubris behind and get our mojo back

debatable if a turn around is possible before '29