frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Thomson: Continual Learning of Frontier Models for SovereignAI [pdf]

https://huggingface.co/spaces/tri-fair-lab/publications/blob/main/Thomson_1_0_Technical_Report.pdf
1•schwarzjn•53m ago

Comments

schwarzjn•53m ago
The development of frontier models is commonly perceived to be in the exclusive remit of a small number of heavily funded players, creating an information, economic and power asymmetry between developers and the diverse user base of modern AI. Recent public discourse acknowledges this concern, calling for SovereignAI (an organisation's capability to independently build, deploy and govern AI use), but often providing little concrete advice on how this can be achieved in the short term under a diversity of funding settings.

In this report, we argue that frontier performance can be achieved by a wide range of institutions through Continual Learning on readily available open-weight models. As opposed to existing limited approaches such as small-scale fine-tuning, prompt engineering, or tool-augmentation with a frozen model, our Continual Learning approach takes advantage of the effectiveness of a modern mid- & post-training stack while introducing safeguards preserving both plasticity and stability at each training stage and seeking to make the minimal number of high-impact interventions on the parameters.

This strategy results in model improvements comparable to the gains typically seen across multiple successive model generations. Crucially, such results are achievable with compute and personnel budgets substantially lower than commonly thought, making ownership of large parts of the SovereignAI stack (model, tool infrastructure, values & data privacy) viable for a wider range of actors.

To demonstrate this, we introduce Thomson, a new general-purpose frontier model trained with an enhanced focus on high-stakes professional work: domains commonly predicted to undergo large productivity improvements through AI. Through a unique focus on Continual Learning, data-centricity, and efficiency, we demonstrate that Thomson performs competitively with recent frontier models on a wide range of domains and capabilities, ranging from agentic tasks to safety, legal, tax & multilingualism, to comprehensive large-scale Deep Research. Thorough evaluations show a distinctive π-shaped pattern: distinct improvements across a wide range of capabilities (including those not explicitly targeted), while almost completely eliminating the forgetting problem common to narrow domain adaptation.

EPA aims to exempt datacenters from disclosing air pollution, advocates warn

https://www.theguardian.com/us-news/2026/aug/25/datacenters-air-pollution-epa
1•pseudolus•32s ago•0 comments

Confidential tokens: Routing is coming for the Frontier AI Labs

https://www.axios.com/2026/08/25/routing-is-coming-for-the-frontier-ai-labs
1•ljlolel•1m ago•0 comments

Show HN: Zoetrope – Watch a Claude Code session as a live flow graph

https://github.com/furkankly/zoetrope
1•furkankly•1m ago•0 comments

Key takeaways from investigation into Edison's role in the deadly Eaton fire

https://www.latimes.com/business/story/2026-08-20/key-takeaways-from-investigation-into-edisons-r...
1•newsomix9xl•3m ago•0 comments

Sablejs 2.0 – An AOT sandbox for AI-generated JavaScript

https://github.com/ErosZy/sablejs
1•zyeros•3m ago•0 comments

Resolving Energy Bottlenecks

https://nomagicpill.site/knowledge/energybottlenecks.html
2•surprisetalk•5m ago•0 comments

Show HN: Candlebar – crypto and stock prices in the Mac menu bar

https://candlebar.app
1•toddynho•5m ago•0 comments

Linux 7.3 Device Mapper Sees Many Fixes, Including Code Cleanups by Claude Opus

https://www.phoronix.com/news/Linux-7.3-Device-Mapper
1•DemiGuru•6m ago•0 comments

Codec Wars Get Real: Lessons from Disney's European 4K Blackout

https://www.streamingmedia.com/Articles/ReadArticle.aspx?ArticleID=176155
1•cisc•6m ago•0 comments

Ask HN: There is current insistence upon learning what LLM's do under the hood

1•joebig•6m ago•0 comments

I ran deconvolution on my own redaction output – here is what came back

https://nexoradia.com/blog/gaussian-blur-is-not-redaction/
1•farshid_dev•7m ago•0 comments

Trump considers renaming Lake Ontario to 'Lake America' as trade war escalates

https://www.theglobeandmail.com/canada/article-trump-considering-renaming-lake-ontario-to-lake-am...
6•bellgrove•7m ago•1 comments

When Code Is Abundant

https://about.gitlab.com/blog/when-code-is-abundant/
1•hoodkg•8m ago•0 comments

Show HN: Crilio – pytest for prompts

https://github.com/mukundzha/crilio
4•mukundzzha•8m ago•0 comments

The Eight Pillars of User Research

https://medium.com/researchops-community/the-eight-pillars-of-user-research-1bcd2820d75a
1•speckx•8m ago•0 comments

Nobody Can Run Your App but You

https://medium.com/ironbee/nobody-can-run-your-app-but-you-313aa9d7afe0
2•sozal•9m ago•0 comments

We had a unit test once which only failed on Sundays (2015)

https://qntm.org/unit
2•lukasgelbmann•11m ago•0 comments

Show HN: GitHub Dashboard aka GitBiased

2•skyfantom•12m ago•0 comments

Parsing IP addresses in C# at crazy speeds

https://twitter.com/lemire/status/2090154772250939471
1•PKop•12m ago•0 comments

"A monad is a monoid in the category of endofunctors" But what does that mean? [video]

https://www.youtube.com/watch?v=ZcD2gGq76FY
1•DPDmancul•12m ago•0 comments

ZenTerm, a macOS terminal with panes, drawers, and tool floats

https://github.com/praxis-labs-io/zen-term
1•jdwhite32•12m ago•0 comments

Ten ways an AI breaks under pressure. How many you can catch?

https://splabs.io/red-teamer-intuition-quotient
2•k-thimmaraju•12m ago•0 comments

Public Utility Commission seeks additional authority to regulate data centers

https://www.kxan.com/news/texas-politics/public-utility-commission-seeks-additional-authority-to-...
2•newsomix9xl•13m ago•0 comments

Show HN: Privacy-First Screen-Time Tracker and Blocker (100% Local, MV3)

https://github.com/timgioh/Privacy-First-Screen-Time-Focus-Dashboard
1•Konstantinsio•13m ago•0 comments

The state of AI in 2026: On the road to ROI

https://www.mckinsey.com/capabilities/quantumblack/our-insights/the-state-of-ai
6•swolpers•13m ago•3 comments

Show HN: Sktime-CLI A command line tool to do all time-series tasks

https://github.com/siddharth7113/sktime-cli
2•siddharth7113•15m ago•0 comments

Show HN: Pyxle – Python and React in one file, no API layer

https://pyxle.dev
1•shivamsn97•17m ago•0 comments

Scams Can't Be Stopped – We Need to Stop Pretending They Can

https://brothke.medium.com/scams-cant-be-stopped-we-need-to-stop-pretending-they-can-f89fa37f4528
4•benrothke•17m ago•0 comments

Depression may shut down the brain's ability to make new neurons

https://www.sciencedaily.com/releases/2026/08/260823094135.htm
1•samso26•17m ago•0 comments

Show HN: GitRank.lol – GitHub repo battleground funding open source

https://gitrank.lol
1•ongdevlab•18m ago•0 comments