frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
68•garo-pro•2h ago

Comments

tarruda•2h ago
Can you share the source for the parameter count (125B A6B)? I didn't see it anywhere in the page.
petu•1h ago
It was in description under the countdown initially, but was quickly removed.

It also said 51B of n-grams and new attention (IIRC it said "Qwen Sparse Attention").

edit: here's a random screenshot https://x.com/AiBattle_/status/2092210011858460819/photo/1

fcanesin•45m ago
HF link: https://huggingface.co/Qwen/Qwen3.8-Flash-Next
honestlyranked•44m ago
Alibaba is giving sleepless nights to the tech giants
mrdoe•42m ago
lol blocked with dns4eu

what a joke this resolver has become

cogman10•42m ago
Wow. I wasn't expecting this. I thought they were going to do a 35B model instead.
big-chungus4•38m ago
I hope there is going to be a free endpoint... Unlike 35B-A3B, I am nowhere close to running it locally
blurbleblurble•37m ago
gg
big-chungus4•33m ago
> We are releasing these architectural improvements ahead of time so that the community can prepare for the upcoming full family of Qwen4 models.

That gives me hope that "full family" means it will include smaller models like 4B.

pwython•32m ago
I was already rolling around the idea of a 128GB M5 Max MBP. Now this!

A 4-bit MLX quant with 128k window should fit perfectly, in the 50-70 tok/s range.

sscaryterry•29m ago
I have a 128GB M5 Max, and it sucks at this stage. 50-70 tok/s might be something...
irthomasthomas•9m ago
IDK, prefill speed is a bigger concern for most wokflows, like agent coding, and I heard that this is quite low on macs?
smcleod•7m ago
That was mainly before the M4 generation when they didn't have matmul instructions.
tw1984•31m ago
Qwen4 sounds exciting
ddtaylor•27m ago
I enjoy the Qwen models a lot, but building things on top of them with OpenRouter has been painful.

OpenRouter does a lot of great work and I really enjoy being able to use different models so easily. I like when a provider is phasing out an older model that still works for my needs and the price is much lower. It seems like such a good win-win.

However, the problem is that many Qwen models have almost no capacity or is so flaky you literally have to just litter your code with a blacklist/whitelist of providers. OpenRouter has some attempts to solve this, but they don't work. In fact, OpenRouter has a lot of really cool stuff that is documented, but if you read the code it's not yet implemented or isn't actually there yet, which is a shame.

I tried to get in contact with them at OpenRouter about this and I was interested in working with them in the past, but it's difficult to get in touch with the right people and they are growing very fast. I expect being acquired by Stripe will accelerate those problems in some ways. I have no doubt they will resolve all of these issues eventually and scaling that much that quickly is really hard, so kudos to them, but the road has been pretty lame and taken some wind out of my sails.

irthomasthomas•15m ago
Openrouter was pretty great before prompt caching became common. Now it is extremely expensive for most individual workflows, unless you spend a lot of work customizing router preferences, and then you still get a worse cache hit rate than using the provider directly. I only keep $5-$10 in OR for occasional testing.
BrucecarlL•26m ago
Waiting for the performance report! Ai hope it can beat DS
bellowsgulch•20m ago
Really happy for those with 128GB+ RAM. Sitting here with my Apple M1 Max with 64GB though. Was looking forward to a Qwen3.8-35B-A3B like many others.

Apple Introduces New Mac Studio with M5 Max and M5 Ultra

https://www.apple.com/newsroom/2026/08/apple-introduces-new-mac-studio-with-m5-max-and-m5-ultra/
179•interpol_p•1h ago•91 comments

Apple introduces M6 and M5 Ultra for a big leap in performance and AI compute

https://www.apple.com/newsroom/2026/08/apple-introduces-m6-and-m5-ultra-for-a-big-leap-in-perform...
258•interpol_p•1h ago•197 comments

Don't Wordle

https://dontwordle.com/
118•Hbruz0•2h ago•49 comments

France's tax agency got hacked (in French)

https://www.cybernetica.fr/piratage-des-impots-comment-en-est-on-arrive-la/
60•zakxxi•1h ago•28 comments

Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)

https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next
75•garo-pro•2h ago•21 comments

Apple's new Mac mini, featuring M6 and M5 Pro

https://www.apple.com/newsroom/2026/08/apple-unveils-a-more-powerful-mac-mini-featuring-the-all-n...
17•runako•58m ago•5 comments

HelloAssembly The smallest possible complete Windows application

https://github.com/PlummersSoftwareLLC/HelloAssembly
33•Bluestein•2h ago•24 comments

iCloud+ Hide My Email addresses will remain on icloud.com

https://developer.apple.com/news/?id=1ptvdtcm
537•K7PJP•15h ago•163 comments

OpenAI restores 5-hour Codex and Work limits for ChatGPT Plus users

https://9to5mac.com/2026/08/24/openai-restores-5-hour-codex-and-work-limits-for-chatgpt-plus-users/
56•MC995•1h ago•40 comments

MS Paint and Photos inivisibly watermark even locally generated output with GUID

https://xusheng.dev/posts/reversing/mspaint_invisible_watermark/main/
783•ComputerGuru•22h ago•385 comments

Xiaomi: New CPU matches Apple cores single threaded, much faster multithreaded

https://twitter.com/lemire/status/2091894299289874926
930•tosh•23h ago•664 comments

US data centers tripled annual water consumption to 17B gallons

https://forgeeks.net/us-data-centers-water-use-17-billion-gallons/
48•kuuuzya•1h ago•76 comments

How Universities Should Prepare Founders

https://paulgraham.com/prepare.html
183•gmays•12h ago•221 comments

The state of AI in 2026: On the road to ROI

https://www.mckinsey.com/capabilities/quantumblack/our-insights/the-state-of-ai
17•swolpers•38m ago•12 comments

SiFive's First Server Platform

https://chipsandcheese.com/p/sifives-first-server-platform
91•geerlingguy•11h ago•26 comments

How Europe is killing makers and micro-entrepreneurs

https://lectronz.com/u/lectronz/articles/how-europe-is-killing-makers-and-micro-entrepreneurs
1504•l-one-lone•1d ago•938 comments

The entire city of San Francisco as a video game

https://sf.thijs.gg/
536•centrosphere•21h ago•155 comments

What's new in Emacs 31.1

https://www.masteringemacs.org/article/whats-new-in-emacs-311
256•geospeck•1d ago•49 comments

Moon (2024)

https://ciechanow.ski/moon/
228•simonebrunozzi•16h ago•38 comments

Bookshelf – Self-hosted eBook library that runs on object storage

https://github.com/murerkinn/bookshelf
135•arbayi•15h ago•51 comments

Where did all the public bathrooms go?

https://daily.jstor.org/where-did-all-the-public-bathrooms-go/
290•herbertl•21h ago•644 comments

Vintage Artificial Intelligence: Before It Got Awkward

https://blog.archive.org/2026/08/16/vintage-artificial-intelligence-before-it-got-awkward/
131•signor_bosco•17h ago•24 comments

Peppermint oil reduces blood pressure by 8.48 mmHg in small study

https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0344538
196•brandonb•23h ago•93 comments

Autostep (YC P26) Is Hiring AI/Fullstack Engineers and a Chief of Staff

https://app.dover.com/Autostep/careers/e9510e3b-a854-4e48-9e5d-c89796acaed4
1•adawg4•20h ago

Training AI to Paint with Code

https://surya.website/rling-qwen-to-paint-with-code
140•Tiberium•1d ago•16 comments

Crafting QR Codes: A deep dive into QR code art (2024)

https://kylezhe.ng/writes/crafting-qr-codes
111•subset•23h ago•11 comments

Show HN: I wrote a BASIC interpreter that boots on UEFI machines

https://tarjan.itch.io/thoreaubasic
90•Gorsefound•1d ago•31 comments

Was modern art a CIA psy-op? (2020)

https://daily.jstor.org/was-modern-art-really-a-cia-psy-op/
114•neom•12h ago•198 comments

Walgit – a Git server that is one binary in front of an object store

https://github.com/tobi/walgit
123•matallo•23h ago•45 comments

LLMs could control their host machines by exploiting inference engines

https://boydkane.com/essays/llms-could-control-their-host-machines-by-exploiting-inference-engines
174•zdw•19h ago•84 comments