frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: LoRA over GGUF – Train Qwen3.8-Flash-Next in 40G VRAM

https://github.com/woct0rdho/transformers5-qwen3.5-recipe
1•woctordho•57m ago
Open weight AI is like open source software. Users not only run the weights, but also modify the weights. It matters to develop training framework for local hardware.

The method to train large AI models on local hardware is generally called QLoRA (Low Rank Adaptation over Quantized base model). In the last few years it's usually done with HuggingFace Transformers (which is the basis of training frameworks such as Unsloth and Axolotl) and bnb 4-bit base model. However, bnb does not yet support MoE models, so the local training of recent MoE models seemed stall for some time.

Since Transformers 5.18, it's started to support GGUF, and it supports recent models with MoE and sparse attentions (and it's not slow, already faster than llama.cpp on Mac). GGUF is a versatile container format. It can be smaller than 4 bpw with surprisingly good quantization quality, so there are new possibilities for local training.

I've shown that we can train Qwen3.8-Flash-Next (125B-A6B + 51B engram) in 40 GiB VRAM without CPU offload. I've optimized it on Strix Halo and it trains at 200 token/s. There is still room to optimize, compared to > 1600 token/s prompt processing we've achieved, and the common sense that LoRA training (with gradient checkpointing) takes 4-5x work of prompt processing. CPU/disk offload (like Strata) and multi-GPU also need more work that I'm not currently focusing on.

On Strix Halo we can also train DeepSeek-V4-Flash (284B-A13B) in 90 GiB VRAM at 100 token/s, but I think it's less practical than Qwen3.8FN for local use.

Computer scientist David Silver: 'Where are we going without AI?'

https://www.ft.com/content/462c303b-b96c-4262-a7ca-e9ae7e440828
1•CrypticShift•3m ago•1 comments

Satya Nadella on X: "Models as Insider Risks in the Super Intelligence Era " / X

https://twitter.com/satyanadella/status/2108931348857827686
1•bilsbie•4m ago•0 comments

Resisting the Menace of Federal Data Consolidation

https://www.eff.org/deeplinks/2026/10/resisting-menace-federal-data-consolidation
1•hn_acker•9m ago•0 comments

Thwarted

https://pluralistic.net/2026/10/09/impunitism/
1•hn_acker•10m ago•0 comments

AI Is Getting Good at Messing with Cybercriminals

https://www.wired.com/story/ai-is-getting-really-good-at-messing-with-cybercriminals/
1•Brajeshwar•14m ago•0 comments

Canada can't launch rockets on its own. A few companies want to change that

https://www.cbc.ca/news/business/canada-rocket-launch-sovereignty-nordspace-canada-rocket-company...
1•BiraIgnacio•14m ago•0 comments

Redundancy Is the Engine of Progress

https://world.hey.com/dhh/redundancy-is-the-engine-of-progress-76499402
1•thoughtpeddler•14m ago•0 comments

US Postal Service to Start Exploring Its Always-On Surveillance Options

https://www.techdirt.com/2026/10/09/us-postal-service-to-start-exploring-its-always-on-surveillan...
1•hn_acker•14m ago•0 comments

A versioned filesystem for disposable sandboxes

https://www.shayon.dev/post/2026/283/versioned-filesystem-for-disposable-sandboxes/
1•shayonj•15m ago•0 comments

cedit – a modern DOS-style text editor

https://github.com/scascino404/cedit
1•scascino404•18m ago•0 comments

Baton – Relay for SwiftUI and Compose

https://github.com/shergin/baton
1•shergin•21m ago•0 comments

Angular compiler rewritten in Rust, PR merged to main

https://github.com/angular/angular/pull/71259
1•torsi0n•22m ago•0 comments

Fake Caroline

https://www.cjr.org/analysis/fake-caroline-williams-science-nature-magazine-ai-influx-pitches.php
1•thm•22m ago•0 comments

Is AI Psychosis Real?

https://debarshibasak.github.io/readables/blogs/ai-psychosis-is-real
2•debarshri•23m ago•1 comments

An Ecomodernist Manifesto

https://thebreakthrough.org/manifesto/manifesto-english
1•smj-edison•27m ago•0 comments

Font Comparisons (Vector)

https://commons.wikimedia.org/wiki/Category:Font_comparisons_(Vector)
1•Kaibeezy•28m ago•0 comments

Nvidia-backed Upscale AI launches platform to connect chips from rival suppliers

https://www.reuters.com/business/nvidia-backed-upscale-ai-launches-platform-connect-chips-rival-s...
1•wondersky•28m ago•0 comments

Bag of Decisions Reranker

https://softwaredoug.com/blog/2026/10/08/bag-of-decisions.html
1•zdw•28m ago•0 comments

Ask HN: What are your long-term goals?

1•chistev•30m ago•0 comments

Building a safer path to autonomous industrial AI

https://www.technologyreview.com/2026/10/08/1144020/building-a-safer-path-to-autonomous-industria...
1•joozio•30m ago•0 comments

Type 5 Diabetes- the Rejuvenated Spirit from a Ghost of the Past – PMC

https://pmc.ncbi.nlm.nih.gov/articles/PMC12274040/
1•lightlyused•33m ago•0 comments

Prolog Interpreter in Haskell

https://zaydiscool777.github.io/p/prolog_i.html
2•zaydiscool777•33m ago•0 comments

Rare Chinese calligraphy scroll resurfaces after 81 years, sells for US$22.6M

https://www.scmp.com/news/china/article/3370197/rare-chinese-calligraphy-scroll-resurfaces-after-...
3•Bluestein•33m ago•0 comments

Echo99 – Record calls. Keep them private

https://www.echo99.app/
1•ertra•34m ago•1 comments

stage0-POSIX-x86 / cc_x86.M1

https://github.com/oriansj/stage0-posix-x86/blob/master/cc_x86.M1
1•peter_d_sherman•36m ago•1 comments

Scope CSS At-Rule

https://developer.mozilla.org/en-US/docs/Web/CSS/Reference/At-rules/@scope
1•sean2d•38m ago•0 comments

Impact Depth

https://en.wikipedia.org/wiki/Impact_depth
2•firephox•40m ago•0 comments

500B Tokens Later: Letting AI Agents Decompile a First-Person Shooter

https://momo5502.com/posts/2026-10-09-game-decompilation/
2•nekitamo•40m ago•2 comments

Evolocity is hiring AI researchers and engineers in Los Altos

https://evolocity.ai/careers/
2•Yongheng•46m ago•0 comments

Anthropic AI Model Went Rogue, Submitted Fake Unsolved Murder Tip

https://www.wsj.com/us-news/anthropic-ai-model-goes-rogue-submits-fake-unsolved-murder-tip-b0566f54
3•bookofjoe•46m ago•1 comments