frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Beijing is forcing a mass breakup with AI lovers

https://restofworld.org/2026/china-ai-boyfriend-ban-bytedance-doubao/
1•impish9208•41s ago•0 comments

Using math for solving real-life problems

https://app.notion.com/p/Number-Pi-Podcast-E09-3b46c4fc04cb8077aad7fb4d4b1eb6cc
1•armanj•1m ago•0 comments

Who is Anthropic's auditor – and why should we care?

https://www.ft.com/content/b93d8030-203b-4445-b4a7-49e52d9b17a5
1•petethomas•1m ago•0 comments

Harbofly, reclaim the 20–100GB of dev caches on your Mac

https://harbofly.app
1•buildcomcarlos•1m ago•0 comments

Ada 83/TLALOC: a modern Ada 83 compiler with DIANA 86 HLIR and fasmg back end

1•ViMoBr•1m ago•0 comments

Multi-tenant BYOK encryption in PostgreSQL with pgcrypto

https://xata.io/blog/multi-tenant-byok-encryption-in-postgresql-with-pgcrypto
1•tudorg•2m ago•0 comments

Data Center Foes' Newest Tack: Sue Locals over Approval Process

https://news.bloomberglaw.com/litigation/data-center-foes-newest-tack-sue-locals-over-approval-pr...
1•speckx•2m ago•0 comments

Oracle has drawn up plans for a new round of layoffs this month

https://www.businessinsider.com/oracle-is-planning-another-round-of-job-cuts-this-month-2026-8
2•DGAP•3m ago•0 comments

Optimizing SQLite for Servers

https://kerkour.com/sqlite-for-servers
1•cold_pizz4•4m ago•0 comments

Glaciers on the Climate Dashboard

https://climate.metoffice.cloud/glaciers.html
1•mooreds•4m ago•0 comments

DeepSeek V4 Pro 0813 quietly released

https://api-docs.deepseek.com/guides/responses_api/
3•HiPHInch•6m ago•1 comments

Prompt Injection Experiments with Opus-5 in Claude Code – Auto-Mode Edition

https://itmeetsot.eu/posts/2026-08-12-opus5_automode/
1•veganmosfet•7m ago•0 comments

[STORY] A Realistic Scenario

https://twitter.com/AISafetyMemes/status/2087549262427263014
1•barry-cotter•7m ago•0 comments

Multiple US sailors have tried to go overboard amid extended deployment

https://www.militarytimes.com/news/your-military/2026/08/11/multiple-uss-abraham-lincoln-sailors-...
1•Geekette•7m ago•0 comments

Buildings Calculate Forces. Software Must Calculate Meaning

https://ideal.iroha1203.dev/buildings-calculate-forces-software-must-calculate-meaning-5c9b2149744b
1•iroha1203•7m ago•0 comments

Meta, 29 states head to court; biggest test yet of youth social media litigation

https://www.reuters.com/business/meta-29-states-head-court-biggest-test-yet-youth-social-media-li...
1•1vuio0pswjnm7•8m ago•0 comments

That Test Isn't Flaky. It's Broken

https://will-keleher.com/posts/its-broken-not-flaky/
1•speckx•8m ago•0 comments

95% of the Fortune 1000 trust us with their information and assets

https://www.ironmountain.com/
1•amai•9m ago•1 comments

Shaping AI

https://www.shapingai.com/
1•sonabinu•10m ago•0 comments

MediaTek Genio upstream update: from HDMI 2.0 to booting on U-Boot and TF-A

https://www.collabora.com/news-and-blog/news-and-events/mediatek-genio-upstream-update-from-hdmi-...
1•losgehts•11m ago•0 comments

Ask HN: Responsible/trustable background check companies?

3•levinb•11m ago•0 comments

Trump's flight switch indicates severity of assassination threat by Iran

https://www.theguardian.com/us-news/2026/aug/12/trump-catering-cart-iran-threat
1•prmph•12m ago•1 comments

Less Chaos. More Life

https://spuntafacile.netlify.app/
1•ridiculouswildc•13m ago•0 comments

Show HN: Tanchi – Open-source AI prospecting agent that only automates email

https://github.com/tanchihq/tanchi
1•saysb•14m ago•0 comments

92 percent of US adults are not going to the doctor – too expensive

https://www.independent.co.uk/us/money/us-medical-costs-doctors-health-insurance-b3026479.html
4•testing22321•15m ago•1 comments

I Still Miss Apple CarPlay After 7 Years in a Tesla

https://www.josemunozmatos.com/blog/why-i-miss-apple-carplay-tesla
1•speckx•16m ago•0 comments

Qwen3.8-2.4T

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B-FP8
3•mmastrac•16m ago•0 comments

Show HN: Pane – An open-source browser that turns your work into living websites

https://github.com/abhishek-verma/Pane
1•vermaabhishek39•17m ago•1 comments

Rclone syncs your files to cloud storage

https://rclone.org/
1•yitchelle•19m ago•0 comments

Wednesday, August 12: GitHub, Incident with Pull Requests and Issues

https://www.githubstatus.com/incidents/76t89hbfb09h
7•arm32•20m ago•0 comments
Open in hackernews

Sharding a 70B model across 39 Intel laptops

https://github.com/labscommunity/cascadia
3•tatef•1h ago

Comments

tatef•1h ago
Hey everyone, we’ve been working on a project called https://cascadia.to which allows you to shard LLMs across Intel-based machines and perform inference across CPUs/GPUs/NPUs.

To get it functioning, we split models into shards and pre-compiled them to OpenVINO IR graphs.

Since then, we’ve been working on a lot of optimizations to make sharding and running models on Intel much more performant. Most of our initial improvements are thanks to the use of speculative decoding (with some interesting workarounds there) and also micro-batching and continuous batching to improve the total throughput.

We’re still in the early days, but from our work so far:

- We ran an 8B parameter model on two Intel PCs, serving two users concurrently via iGPUs at ~43 tok/s aggregate - After adding an additional PC and user, it reached 64.67 tok/s - We’ve successfully sharded larger models (e.g. with 70B parameters), and with Cascadia inference was 3.1x faster than basic sharding - And for fun, we got 39 Intel AI PCs, hooked them up via ethernet, and sharded a 70B model across them. Benchmarks aren’t super impressive yet (~1 tok/s on CPU), but we’ve been making a ton of progress there.

If you’re wondering why we’re doing this, there’s a few reasons. There is a lot of Intel hardware out there, and dedicated AI hardware right now is expensive.

Organizations that own Intel computers will be able to pool their computing power to run models locally. For hobbyists with Intel GPU rigs or Intel Xeon Server PCs, you likely want inference to be optimized for your hardware. Cascadia is aiming to be a runtime for anyone wanting to run AI on Intel hardware and squeeze the most juice out of their machines.

Cascadia is open source (Apache 2.0) and available now.

This is alpha, so we’re open to any feedback on the architecture.

Give it a go and let us know what you think :)