Analysis paralysis stifles not just human intelligence, but other intelligences too.
monkey_monkey•6m ago
[delayed]
erichocean•6m ago
Need this done for DeepSeek, ideally one of the Flash models.
tomrod•5m ago
Well done, and great iteration.
The pareto frontier needs clearer distinction. Benchmarks miss half the story. What, if any, capability is lost by the token reduction (for example, was it like super awesome at Golang before and now kind of sucks? that kind of distinction).
esafak•4m ago
It looks like it would be similar to GLM 5.3 Flash, had they tested it...
jamienk•4m ago
Ignoring for the moment issues of what "counts" as open, won't open models rapidly advance due to stuff like this in ways that it's less possible for the proprietary ones to do? This is exactly how Linux & Wikipedia, for example, overtook their "frontiers", right?
ls612•4m ago
On the smaller end, Quen 3.8, while being extraordinarily capable for a small local model, also suffers from extreme thinking. I wonder if the techniques described here generalize to other models too.
andsoitis•7m ago
Analysis paralysis stifles not just human intelligence, but other intelligences too.