frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

macOS 26.4 silently enabled login keychain data protection

https://lapcatsoftware.com/articles/2026/9/5.html
1•chmaynard•1m ago•0 comments

Google CC, an AI agent built for families

https://blog.google/innovation-and-ai/models-and-research/google-labs/cc-expanding-to-groups/
2•hmokiguess•1m ago•0 comments

Ask HN: Why does it matter if no one reads your novel as long as the AIs do?

2•amichail•1m ago•0 comments

IA Agent Arthur – Developed an OfflineAIagent

https://x.com/AI_AGENT_ARTHUR
1•AIAGENTARTHUR•3m ago•0 comments

Building Software That Can Prove Agents Wrong

https://www.rafael.md/writing/building-software-that-can-prove-agents-wrong
1•rafaelcamaram•4m ago•0 comments

vLLM Architecture, Memory and Benchmarks Deep Dive

https://www.g-ftech.com/blog/vllm-throughput-deep-dive
1•gfactor_ai•5m ago•0 comments

CBP suspends all personal prescription importation Oct 22

https://www.personalimportation.org/advocacy
2•burnt-resistor•5m ago•0 comments

Show HN: MoA – Attention kernel proven minimal before code was written

https://github.com/womenflyplanes/moa-attention-verified-mullin
1•lmullin•6m ago•0 comments

EU plans "digital expropriation" of Europeans in the interest of AI companies

https://noyb.eu/en/ai-eu-member-states-plan-digital-expropriation-europeans-interest-ai-companies
2•latexr•6m ago•0 comments

Future-Proof Data Systems [pdf]

https://www.vldb.org/pvldb/vol19/p4965-giceva.pdf
1•matt_d•7m ago•0 comments

A Major Cold Intrusion Spreads Across Eastern Europe: Seasonal Frost to Follow

https://www.severe-weather.eu/global-weather/major-cold-intrusion-polar-blast-eastern-europe-sept...
1•dgellow•7m ago•0 comments

Many Choices Lie Ahead

https://proofsandprompts.com/2026/09/21/many-choices-lie-ahead/
1•mathgenius•8m ago•0 comments

Tell HN: WebRezPro Compromised

1•chrisaycock•9m ago•0 comments

Your Second Brain Won't Do the Work

https://korin.uk/your-second-brain-wont-do-the-work
1•phoronixrly•10m ago•0 comments

Women Who Sold Books Door to Door

https://daily.jstor.org/the-women-who-sold-books-door-to-door/
2•herbertl•11m ago•0 comments

Astronomy mapped: Africa lacks optical observatories

https://observatorymap.withastra.io/
1•pppone•12m ago•0 comments

Per-Layer Embeddings (PLE)

https://sebastianraschka.com/llm-architecture-gallery/per-layer-embeddings/
1•Bluestein•14m ago•0 comments

Googlebook's built-in intelligence reinvents the way you use your laptop

https://blog.google/products-and-platforms/devices/googlebook/googlebook-built-in-intelligence/
2•zathan•14m ago•1 comments

Declarative WASM GC Types

https://blog.veritates.love/decl-wasm-types
1•ibobev•14m ago•0 comments

Bitcoin decentralized usernames are now available

https://twitter.com/robin_linus/status/2102109957718151641
1•ca98am79•15m ago•1 comments

P2300R10: `Std:Execution`

https://www.open-std.org/jtc1/sc22/wg21/docs/papers/2024/p2300r10.html
1•ibobev•16m ago•0 comments

GitHub Actions leaking secrets when Miri output is cached

https://blog.rust-lang.org/2026/09/21/github-actions-leaking-secrets-when-miri-output-is-cached/
2•andrewstetsenko•17m ago•0 comments

More than a taken branch per cycle?

https://lemire.me/blog/2026/09/21/more-than-a-taken-branch-per-cycle/
1•ibobev•17m ago•0 comments

Ask HN: Do You Use Siri?

5•moomoo11•18m ago•7 comments

All your base are belong to us

https://en.wikipedia.org/wiki/All_your_base_are_belong_to_us
3•simonebrunozzi•18m ago•0 comments

Relation Algebra ≠ Relational Algebra

https://remy.wang/blog/ra-ra.html
2•fanf2•22m ago•0 comments

Ask HN: How to Learn Systems Programming

1•sampsn•22m ago•1 comments

Paul's Automotive DIY Repair Guides

https://www.paulstravelpictures.com/Articles/index.htm
1•js2•25m ago•0 comments

Jonathan Swift vs. AI

https://www.thedial.world/articles/news/schools-teachers-artificial-intelligence
1•cainxinth•26m ago•0 comments

Wow Forever Classfinder – Find Your Perfect Class

https://wowforeverclassfinder.com/
1•marc0janssen•27m ago•0 comments
Open in hackernews

Xiaomi MiMo v2.6

https://mimo.xiaomi.com/mimo-v2-6
141•volf_•52m ago

Comments

algoth1•41m ago
Finally a lab that doesn't cheat on the charts
vatsachak•38m ago
Wow, the chinese labs are getting good at advertising model releases. The moat is thin.

Some features of the release I like:

- Demonstration of diverse tasks, such as using a DAW

- Graphs from various benchmarks and price ranges

- Real world use of the model in scientific environments

unpopularopp•35m ago
I've never used worst smartphones than anything from Xiaomi, bloated ad infested borderline malware territory fork of Android. Maybe just me but whenever I see them on HN I just can't think anything good about this company.
platinumrad•32m ago
It's a big company, like Microsoft or Google. Some of their products are good and some are bad.
verdverm•13m ago
ironic to this thread, I have less bloatware and ads since I switched from Verzion to Pixel on Fi (many years ago)

Curious if Verizon / ATT still force apps on your phone, eg. NFL and Amazon apps, Fi service is subpar

InsideOutSanta•16m ago
It's funny, I have the exact opposite reaction. This is probably misguided on my part, but Xiaomi is one of the very few major tech companies that I don't have an immediate strong negative reaction to. Everything I've bought from them, from robot vacuum to mobile phone, has been reasonably well designed, didn't break, and was priced fairly. I also think their car looks badass.

I'm sure they're doing all kinds of terrible things, like all major companies. I just can't help but like them. Also, this model looks great, and I'll give their subscription a shot next month.

algoth1•16m ago
I still have a xiaomi mi 11 lite, my wife has a 15t. The cameras are the best for the price. The way they chove ads down your throat at every opportunity should be illegal though
DanMcInerney•34m ago
This is a big week. Probably getting next OpenAI and Anthro models, Grok 4.7, Mimo, etc. These open source model releases are why I can't take the "slow down" crowd seriously. I pitted older Mimo, qwen, step, gpt-oss, and other models against each other playing games like Werewolf and Sketch.io-like games where I let them talk shit while they played against each other. Mimo was by far pareto frontier of game-playing for the models that were <$0.15/m input tokens on OpenRouter. Qwen was pareto frontier in the shit talking game though. Qwen's hilarious. https://www.tiktok.com/@clankerfights/video/7642862917582425...
ddxv•34m ago
This looks great in terms of cost and capabilities, truly pushing the frontier forward in terms of open weight light weight models.
omani•34m ago
ah, would you look at that. I was wondering why mimo 2.5 became "dumber" the last weeks. I was speculating they are probably about to release a new version of the model. because the model really acted out a lot. especially the last two weeks. dont know, was just a feeling, highly speculative.

but now I got my "proof".

sandblast•24m ago
I guess that would only be possible if your provider was Xiaomi itself?
omani•15m ago
yes. I use opencode and opencode uses Xiaomi as a provider.
nemothekid•33m ago
Looking at the frontend design examples; why do these models seem to love the "01 - UPPERCASE TEXT" motif. It's everywhere now (see https://try.cloudflare.com/, which has '01 · QUICK TUNNELS', but no "02" anywhere).
sandblast•26m ago
Nice catch!
danvayn•22m ago
My guess is that by function they break down frontend sections or components into pieces and I believe document things for themselves on some level, or purposely are verbose in this way. It is probably also shaped by users and existing web patterns. They probably get reinforced by models the more common they become.
rao-v•31m ago
I know we have strong views on what a truly open model is (open weights, open training data, open training code etc.) but I really like how transparent they’ve been about the training of this model.

The realtime dashboard they shared during training (https://mimo.xiaomi.com/rl/) was an incredible learning and teaching tool for me, and they’ve been unusually comprehensive in sharing details about their methodology (check out that tech report - it's got lots of clever behind the scene tricks like Google or Deepseek writeups) and benchmark scores (even the stuff they didn’t do well on).

If you’re releasing an open model going forward, please consider offering the community more of this transparency!

stymaar•28m ago
Flash[1]: 309B total / 15B activated parameters

Pro [2]:, 1.02T total / 42B activated parameters

[1]: https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Flash-RL

[2]: https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Pro-RL

verdverm•17m ago
curious why the HF pill (on the right) always has inaccurate values
stymaar•9m ago
I noticed the same, and I wonder as well.
verdverm•11m ago
There's also a Qwen 3.5 9B distill

https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B

gandreani•8m ago
Those this mean they've fine-tuned this Qwen 3.5 9B on output from the V2.6 model?
mydreamof•4m ago
It is a 9B agentic model developed by Xiaomi MiMo through supervised fine-tuning of Qwen3.5-9B on MiMo-generated data
syntaxing•26m ago
All these new models are such tease for us folks with 128GB of shared memory. Buying another unit now to expand to 256GB is a mortgage payment but it’s getting tempting…
brcmthrowaway•23m ago
Is there a gamechanger around the corner to reduce DRAM requirements?
stymaar•18m ago
n-gram per-layer embeddings[1][2] might be it.

[1] https://sebastianraschka.com/llm-architecture-gallery/per-la...

[2]: See DS 4.1-Flash and Qwen-3.8-Next.

verdverm•15m ago
this is to offload VRAM to DRAM (for GP comment), and makes no difference for URAM
stymaar•11m ago
Am I missing a joke? WTF is URAM?
verdverm•9m ago
unified memory, not sure if anyone uses URAM, I human hallucinated it
verdverm
bertili•21m ago
They mixed up DeepSeek 4.1 Flash with something else on this page, possibly DeepSeek 4.1 Flash means Gemini 3.8 Flash.
NooneAtAll3•20m ago
does anyone know what unnamed model is on paretto frontier picture right between MiMo 2.5 and 2.6?

so weird to acknowledge someone being on the front edge, but not name it

AnodicElegy•12m ago
Pretty sure that's Luna xhigh.
spwa4•18m ago
As for the stats that everyone wants:

MiMo-V2.6-Flash-310B-A15B roughly GPT-5.6 Luna / Claude 4.9 according to benchmarks MiMo-V2.6-Pro-1.02T-A42B roughly GPT-5.6 Sol / Opus 5 according to benchmarks

Perhaps with IQ2 flash will run on 128G M5?

MisterMunchkin•16m ago
I really liked MiMo 2.5, it was really affordable and actually had vision, unlike DeepSeek. (DeepSeek has only recently added it)

Just tried 2.6 flash on a really niche topic I specialise in and it has done a really good job. They’ve definitely polluted their training data with claudeslop, but looking past the slop there is a decent model.

omani•13m ago
how do you recognize "claudeslop"?
Bluestein•8m ago
It's an honest, load-bearing, simple thing.-
jwpapi•9m ago
In the chart they use "Pareto Line", which I think is wrong. Pareto is 20% effort leading to 80% results. Which could be interpreted as models costing 20% having 80% of peak intelligence, but that’s not what it looks like to me.

It looks like the "Frontier Line" to me, which is also often misinterpreted. frontier does not mean the best models. It means all models that are not strictly dominated, meaning in most cases: Not same price or cheaper and more intelligent.

I personally would like the word frontier to be used with more criterias: Open Weights, per use-case, etc etc. This would make model selection easier, but I understand it’s not an easy thing to do.

hashmush•4m ago
"Pareto" is many things, but here it does indeed refer to the frontier: https://en.wikipedia.org/wiki/Pareto_front
lwansbrough•8m ago
Anyone else more excited about Chinese models than American models these days? Big thing for me is affordability.
swingandamiss•5m ago
No, because I'd rather not support our economic and military rivals.
Freedom2•4m ago
Agreed, and also because I support freedom of speech!
user43928•7m ago
I don't trust any of the benchmarks where Opus 5 surpasses Astra or Fable 5.1.

Maybe Terminal Bench 4.0 and ExploitGym are reasonable.

Terminal Bench 4.0

GPT 6 Astra 59.6 Claude Fable 5.1 55.1 Claude Opus 5 49.0 MiMo-V2.6-Pro 34.9 MiMo-V2.6-Flash 28.8 DeepSeek V4.1 Flash 26.8 MiMo-V2.5-Pro 1.5

ExploitGym

GPT 6 Astra 42.4 Claude Fable 5.1 30.4 Claude Opus 5 22.1 MiMo-V2.6-Pro 17.8 MiMo-V2.6-Flash 6.0 MiMo-V2.5-Pro 0.1

DeepSWE v1.1

DeepSeek V4.1 Flash 74.2 Claude Opus 5 74.0 GPT 6 Astra 74.0 MiMo-V2.6-Pro 71.9 Claude Fable 5 70.0 MiMo-V2.6-Flash 67.9 MiMo-V2.5-Pro

mokre•5m ago
Maybe you should not trust any of the benchmarks!
•
11m ago
https://huggingface.co/XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B is an option