frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Xiami Mimo 2.6 Live Post-Training Dashboard

https://mimo.xiaomi.com/rl/
48•krackers•56m ago

Comments

wolttam•22m ago
Hah, it would be great to see more labs pick this up.
krm01•21m ago
This is pretty neat. What would be a good reason for the other Model providers to not do this?
kibae•10m ago
Speculating here, but I assume researchers can make a reasonable estimate of the size of closed models based on factors like training time, training speed, and the number of tokens processed.

Also, Anthropic and OpenAI probably want to keep each other on their toes so they don’t end up on the wrong side of another Opus 4.6 / GPT-5.3-Codex situation, where one lab releases a model only for the other to drop a better one hours later.

speedgoose•18m ago
I didn't know 2 thirds of the training data would be source code.
jerrygenser•12m ago
that is the the "data used to improve the model" when signing up for the subscription plans
leothetechguy•12m ago
this is the rl run, not the pretraining run
thehamkercat•14m ago
This is crazy, but sadly anthropic/openai will never do this, what has happened to this world, where chinese companies are more open than US or even EU companies
rozab•13m ago
Why are they doing this? To try head off accusations about distillation?
bayindirh•6m ago
Sometimes you're confident about what you're doing and show how you work to the world.

Keeping the garage door open, or at least making the door translucent. It's always cool.

liuliu•12m ago
When you run benchmarks while training, isn't that the definition of contamination? Asking because I am not sure if this is normal in big labs now.
lucrbvi•10m ago
They are using it to evaluate checkpoints during the training, they are probably not using the benchmarks for training the models. It's a common practice for big reinforcement learning runs.
jampekka•5m ago
Kinda yet. The benchmarks become part of the validation set, which means they get slightly overfit to them if they are used as criteria for stopping the training. But a lot less compared to using them in the training data.

I'd guess everybody uses at least some benchmarks as stopping criteria, which is kinda sensible, but also does mean some benchmaxxing, and explains why the newest models always tend to eke out in benchmarks.

https://en.wikipedia.org/wiki/Training,_validation,_and_test...

ProfessorLayton•12m ago
2.6 Pro: >started 2026-09-15 10:32 UTC

For some reason I thought training took much, much longer than what the progress bar suggests.

This is really neat, I'm currently using mimo 2.5 pro, and it's decent (or great given the price). Hopefully their next one is multimodal.

GaggiX•6m ago
These are post-training reinforcement learning steps.
joelwallis•11m ago
I been using MiMo-V2.5 to do most of my work as software engineer, on a variety of projects I'm working on, and I been VERY happy with ROI. The model is very powerful! Not perfect – I've run in hallucination loops once or twice, but nothing a stop-then-continue wouldn't solve.

The cost is unbelievably low, and the quality of intelligence I get is equivalent to when I was working mostly with Anthropic models (late last year/early this year). I'm fully invested in MiMo and I'm very happy with it.

-- PS: I also check almost daily to see if other models are capable of doing such great work. And they do – DS4F is powerful and DS41 is impressive, GLM 5.3 Flash gets a job done well, etc. – but when I add cost of M-token in the ROI math, Jeez! MiMo is an order of magnitude better.

james2doyle•7m ago
2.5 Pro or the regular 2.5?

I always found that those Mimo models to be really good at tool calling and following instructions

levocardia•8m ago
You'd think they would make it less obvious that they are running their whole operation with Claude

Reversing AIs tech job displacement

https://gist.github.com/mxfactorial/34ec46591752cc6955c842d436b01604
1•mxfactorial•1m ago•0 comments

Interakt, an open-source self-hosted search and AI chat for your website

https://github.com/alphasolutionsrepo/interakt
1•interakt•4m ago•0 comments

Gnome 51, "A Coruña"

https://release.gnome.org/51/
1•plaguna•4m ago•0 comments

Artemis II Launch 35mm High Speed Film (01.apr.2026)

https://images.nasa.gov/details/KSC-20260401-MH-AJN01-0001-Artemis_II_Launch_35mm_High_Speed_Film...
1•HelloUsername•5m ago•0 comments

A 3D Rasterizer for Embedded Devices

https://github.com/CubeCoders/Jet
1•SamuraiLion•5m ago•0 comments

Breaking the 1.58-bit Barrier for Ternary LLMs

https://arxiv.org/abs/2609.16338
1•matt_d•6m ago•0 comments

2026 Small World in Motion Competition – Nikon Small World

https://www.nikonsmallworld.com/galleries/2026-small-world-in-motion-competition
1•xnx•8m ago•0 comments

RustFS 1.0.0 GA: Production-Ready, Open Source, S3-Compatible Object Storage

https://rustfs.com/blog/announcing-rustfs-1-0-0-ga/
1•frozeus•9m ago•0 comments

Zuck's bot increased our Vercel bill 10x

1•Marciok•10m ago•0 comments

How VoltDB Works

https://docs.voltdb.com/UsingVoltDB/IntroHowVoltDBWorks.php
1•ronfriedhaber•11m ago•0 comments

Ask HN: People unemployed for >2 years, how do you spend your time?

3•DeepLogin•12m ago•2 comments

Ukrainian USV Attack on Sochi Reveals Rarely Seen Russian Navy Trained Dolphins

https://www.hisutton.com/Sochi-Attack-Dolphins.html
1•EA-3167•12m ago•0 comments

Copyrightability of LLM-Generated Code

https://fsfe.org/news/2026/news-20260825-01.fr.html
1•erlend_sh•13m ago•0 comments

I'm getting sued by a data center

https://technically.beehiiv.com/p/why-data-centers-fight-to-keep-state-pitches-secret
2•herbertl•13m ago•0 comments

Trump's war on EVs derailed America's auto-factory revival

https://economictimes.indiatimes.com/news/international/business/how-trumps-war-on-evs-derailed-a...
8•devonnull•17m ago•0 comments

Show HN: A simple Mac extension for recent screenshots

https://github.com/KanishkVashisht/screenshot-shelf
1•kva•19m ago•1 comments

Everything we know about good agent design

https://rubriclabs.com/blog/everything-we-know-about-good-agent-design
1•handfuloflight•20m ago•0 comments

R/DoohickeyCorporation

https://www.reddit.com/r/doohickeycorporation/
1•rhet0rica•20m ago•1 comments

WASM to Go Source Code

https://github.com/goccy/wasm2go
1•jcbhmr•21m ago•0 comments

Ask HN: How would you anthropomorphize the Anthropic Claude logo?

1•standeven•22m ago•0 comments

Automattic's interim CEO and legal chief signed reciprocal severance deals

https://techcrunch.com/2026/09/16/automattics-interim-ceo-and-legal-chief-signed-reciprocal-sever...
3•miguelxpn•25m ago•1 comments

Arrow 2 and Arrow 2 Telos

https://quiver.ai/blog/introducing-arrow-2-0/
1•handfuloflight•25m ago•0 comments

History of Forgotten Chips [video]

https://www.youtube.com/watch?v=7Cd4NQu1OG0
3•rbanffy•28m ago•0 comments

An air taxi that floats: Regent's Seaglider takes flight with people on board

https://www.smartcitiesdive.com/news/regent-seaglider-first-human-crewed-flight/829828/?_hsenc=p2...
3•rmason•29m ago•0 comments

Brushstrokes: Avant-Garde and the Psychology of Perception

https://eclecticlight.co/2026/09/16/brushstrokes-avant-garde-and-the-psychology-of-perception/
1•Brajeshwar•30m ago•0 comments

The Rhythm of Your Breath Leaves a Fingerprint on Your Thoughts

https://nautil.us/the-rhythm-of-your-breath-leaves-a-fingerprint-on-your-thoughts-1285059
2•Brajeshwar•30m ago•0 comments

New satellites and artificial intelligence are transforming wildfire detection

https://www.theguardian.com/us-news/2026/aug/17/satellites-ai-wildfire-detection
2•rmason•30m ago•0 comments

How to use HDR brightness as a graphic design hack

https://hdrlogo.com/after-white
1•alex_trfmv•31m ago•1 comments

DOE vs. GitHub, INC: LLM generated-content not a DMCA violation [pdf]

https://cdn.ca9.uscourts.gov/datastore/opinions/2026/09/16/24-7700.pdf
1•telotortium•32m ago•0 comments

Racing driver is about to race 100 karts at once

https://arstechnica.com/cars/2026/09/the-worlds-best-racing-driver-is-about-to-race-100-karts-at-...
1•Brajeshwar•32m ago•0 comments