frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Twenty Years of RISC OS Open

https://www.riscosopen.org/news/articles/2026/06/20/twenty-years-of-risc-os-open
34•AlexeyBrin•1h ago•1 comments

Meshdiff – visually compare two STL versions in the browser, client-side

https://meshdiff.com/
66•projscope•2h ago•8 comments

Show HN: Bor – Open-source policy management for Linux desktops

https://getbor.dev/blog/2026-08-02-bor-v080-release/
74•eniac111•4h ago•14 comments

Artificial Intelligence: Ars Notoria and the Promise of Instant Knowledge

https://publicdomainreview.org/essay/ars-notoria/
52•jruohonen•3h ago•6 comments

Show HN: Fuse – statically typed functional programming language

https://fuselang.org
22•the_unproven•2h ago•2 comments

Show HN: Syncular – offline-first SQL sync with TypeScript and Rust cores

https://github.com/syncular/syncular
37•quambo•3h ago•17 comments

Go 1.27 Interactive Tour

https://victoriametrics.com/blog/go-1-27/index.html
267•Hixon10•12h ago•109 comments

Show HN: I'm a 15 Year Old Wannabe Engineer, This Is a Cycloidal Gearbox I Built

https://github.com/tom-ilan/cycloidal_gearbox
202•tomilan•11h ago•71 comments

Great Question (YC W21) Is Hiring Senior Demand Gen Manager

https://www.ycombinator.com/companies/great-question/jobs/YutDxyf-senior-demand-generation-manager
1•nedwin•1h ago

Seedance 2.5

https://seed.bytedance.com/en/blog/one-take-creation-flexible-referencing-introducing-seedance-2-5
387•njaremko•16h ago•202 comments

Cyberscript

https://cyberscript.dev
29•dtj1123•5h ago•28 comments

Has the New Cocaine Arrived?

https://playboy.substack.com/p/has-the-new-cocaine-finally-arrived
16•bookofjoe•31m ago•12 comments

Diátaxis

https://diataxis.fr/
400•ryanseys•17h ago•49 comments

Wikimedia Foundation refuses union recognition, hires union-busting law firm

https://en.wikipedia.org/wiki/Wikipedia:Wikipedia_Signpost/2026-08-02/News_and_notes
154•akolbe•2h ago•124 comments

MkLinux and the pimped-out Apple Workgroup Server 9150

http://oldvcr.blogspot.com/2026/08/mklinux-and-pimped-out-apple-workgroup.html
70•goldenskye•10h ago•4 comments

Holocloth

https://holocloth.vercel.app
39•ingve•2d ago•9 comments

An internal OpenAI Astra model solved 10 major open math and CS problems

https://twitter.com/polynoamial/status/2083467194663571701
36•wa5ina•1h ago•23 comments

Running Kimi K3 on MI355X at Better Performance per Dollar Than B300

https://www.wafer.ai/blog/kimi-k3-mi355x
158•ilreb•9h ago•74 comments

Folding Paper Globes

https://foldingglobes.com/globes
3•dango2506•4d ago•0 comments

Show HN: Katharos Functional programming and CSP-style concurrency for Python

https://github.com/kamalfarahani/katharos
6•kamalf•2h ago•1 comments

IBM i (OS/400) the Database Operating System

https://osadmins.com/en/ibm-i-os-400-the-database-operating-system/
31•naves•6h ago•16 comments

Deep-sea vehicles spot 'alien' sharks deep beneath the waves in the Pacific

https://www.science.org/content/article/deep-sea-vehicles-spot-alien-sharks-deep-beneath-waves-pa...
70•pkaeding•10h ago•35 comments

US Treasury undertakes historic intervention in yen market

https://www.ft.com/content/0f9b2fe7-bde4-4f5f-b49e-93ccb5da9ea8
36•23pointsNorth•2h ago•24 comments

Atom is better than RSS, in ways that matter

https://chrismorgan.info/atom%3Erss
105•frizlab•15h ago•59 comments

When random.bytes() runs but doesn't work

https://insider.btcpp.dev/p/when-randombytes-runs-but-doesnt
66•Funes-•11h ago•33 comments

ASRock BC-250: Building the Budget Steam Machine

https://plug-world.com/posts/2026/asrock-bc250-the-budget-steam-machine/
78•plug_world•12h ago•36 comments

A big win for Android interoperability

https://www.openhomefoundation.org/blog/a-big-win-for-android-interoperability/
160•soheilpro•1d ago•121 comments

Elena, a library for building Progressive Web Components

https://elenajs.com/
44•brianzelip•3d ago•3 comments

Unraveling the mysteries of habit formation

https://www.kyoto-u.ac.jp/en/research-news/2026-07-28
87•hhs•14h ago•33 comments

Postmortem for Kernel Soundness Bug #14576

https://leodemoura.github.io/blog/2026-8-1-postmortem-for-kernel-soundness-bug-14576/
156•juhopitk•19h ago•59 comments
Open in hackernews

The Vanishing Page: AI Firms Scan Then Destroy Rare Book Editions

https://dallasexpress.com/national/the-vanishing-page-ai-firms-scan-then-destroy-rare-book-editions/
45•throw0101a•4d ago

Comments

inigyou•54m ago
Yes, since we complained about them violating copyright so much, it does not violate copyright if they destroy the original, so they are now destroying the original. We got our demands met.
lookingdesk•51m ago
That's not how copyright works, at all.
inigyou•48m ago
Yes, actually it is
netsharc•42m ago
I'm going to buy a Taylor Swift CD, rip it as MP3s, destroy the CD, and put the MP3s online...

Will I get sued?

pfdietz•37m ago
The "putting it online" part is where you are f-ing up.
themaninthedark•22m ago
But it's transformed now!

Edit(seriously): >Transformativeness is a characteristic of such derivative works that makes them transcend, or place in a new light, the underlying works on which they are based. In computer- and Internet-related works, the transformative characteristic of the later work is often that it provides the public with a benefit not previously available to it, which would otherwise remain unavailable.(Wikipedia)

I am sure Napster or the like argued that what they were doing was transformative.

Not sure how AI companies are arguing about PUBLIC benefit if you have to pay...

pfdietz•33m ago
Destroying your copy of a copyrighted work does not violate copyright.
madaxe_again•46m ago
It absolutely is. By destroying the original they retain the one copy they were licensed. It explicitly discusses that this was judged fair use in a court of law, because of the destruction of the original, as this becomes a transformation of the original work rather than a duplication.

Don’t hate the player, hate the game.

themaninthedark•19m ago
So what are they doing about the online data they ingested?
callmeal•50m ago
Yay capitalism! Imagine how much value is being created for shareholders by this destruction! Another example of "great work" being done!

Jokes aside, I don't know what I would do in this situation either. Maybe find a way to mark "rare" or "out of print" books and sell/auction the "pages"? Or make the scanned book available via some mechanism?

madaxe_again•41m ago
You could “ingest” them all into a “model” and sell access to the information within them via “token usage”
jstummbillig•50m ago
Forever preserving the content of a book by consuming one (1) copy of a rare book is the kind of thing we should be doing more of.

Roughly nobody would have had access to that copy. Roughly everyone can now benefit from its content.

maccard•49m ago
Where can I read any of these destroyed copies from anthropic?
XorNot•4m ago
What do you imagine a rare book is? It's not the only extant copy of the writing in that book, in fact it's not even like it represents a significantly limited accessibility of that book.
Certhas•48m ago
Roughly Anthropic owners benefit from the content, unless the scans are made available somehow.
anon373839•47m ago
> Roughly everyone can now benefit from its content.

You mean, roughly everyone who uses that particular model provider’s products. For truly rare books, this has an anticompetitive flavor, since it ensures others can’t train models from the same knowledge.

polarbearballs•44m ago
I have a deep rooted doubt that "roughly everyone can now benefit from its content" If there's any value in the content it will immediately be monetized with an aggressive pay structure and had it sat in a library, that same value would be free.
polarbearballs•50m ago
Who knew Ray Bradbury could be so wrong.
pfdietz•49m ago
Nowhere are the supposed rare books listed.

Confuses "(not all that) old books" with "rare books".

Ragebait, nothing more.

madaxe_again•43m ago
They are academic titles from the last few decades for the most part, and are rare because the print runs were typically short, and the prices typically eye-watering. “Semantic analysis of cotton processing terminology in Des Moines, 1993-1996, 8th edition”, that sort of fun.
pfdietz•38m ago
I'm reminded of a quote that's been (mis)attributed to Dorothy Parker.

"This is not a novel to be tossed aside lightly. It should be thrown with great force."

eleventen•49m ago
There is good discussion on the previous thread from 6 days ago: https://news.ycombinator.com/item?id=49068738
cowanon77•43m ago
On the surface it seems bad, although many of these books were probably rotting in place rather than being read. If the companies are willing to make the digital version available this may actually prreserve the books.

More troubling to me is the focus on pre-2022 data. This indicates that the companies are seriously worried about model collapse, where the future models become worse becuase they are being fed by generated data instead of real human data.

This is actually my worse case scenario for AI: the models become good enough to disrupt entire industries in the next 10 years, but then stay frozen at that level. Model drift then kicks in because the world will keep changing but the models do not, and in a few decades we actually regress instead of progress becuase they won't be enough skilled humans to drive advances.

jmyeet•36m ago
Yeah, I really don't know where this goes from here. Like certain sources, particular Reddit and Twitter, have to largely be considered spoiled at this point. It's hard to quanitfy the amount of bots but my intuition tells me it's really high.

But changing how non-AI people write, that's an interesting angle. Because where do we go from here? In 100 years will we still be overvaluing pre-AI sources? That doesn't make sense. Of course a lot can (and will) change in 100 years.

But i'm reminded of pre-atomic steel, which is steel made before the first atomic bombs were detonated and thus have really low background radiation. This is necessary for making MRIs and such. People will go and find it from shipwrecks and such (ironically, many of which are from WW2). It's also a finite resource. What happens when we run out?

Now pre-2022 texts aren't consumed (other than destructively scanning books of course) but it is also finite. We can't make more of it.

storus•32m ago
As long as economics incentivizes next quarter thinking nothing will be done. We can't even handle human portion of global warming which is arguably more important than just some AI model collapse scenario.
DrDeese•37m ago
Why don't they just put the pdf versions on Internet Archives? Then they would at least be preserved for the future
pfdietz•35m ago
That would be illegal.
darkoob12•23m ago
Why don't they rebind them and put it back in some library? They don't have to throw it out.
XorNot•2m ago
Because legally then it's a copy and not a format conversion.

This story has a lot of details left out and is a magnet for illinformed commentators: the reason the originals are destroyed is because it is legal to do so for the purposes of format conversion, but not legal to retain the originals under copyright law.

quacked•31m ago
Humanity hasn't caught up to the reality of the times; we should adjust copyright laws so anything over say 10 years old can be hosted for free on the Internet. We should have colossal databases of music, books, movies, etc. that can be freely traded and downloaded without breaking any laws. Maybe in the past the commerical protection of art and IP was necessary but with the Internet and AI I think we need to, as a species, get over it and focus on information archival and dissemination over protecting income streams.
datakan•22m ago
10 years may be too little but generally I agree.

What's blowing my mind lately though is the people complaining about copyright today are some of the same people I saw 30 years ago screaming "Information wants to be free" and running around with the DeCSS code on their T-Shirts. But suddenly, now that they hate AI, copyright is important and the information should be locked away. I can't wait for this all to settle down.

robrain•11m ago
It's possible to be opposed to excessive copyright and also be opposed to the mass theft of media for profit and subsequent abuse of copyright that the scrapers are guilty of. The AI pillagers are taking everything in, then gatekeeping access. And as they acknowledge with their "nothing after 2022" preferences, they (or their users at least) are salting the earth for future seekers of knowledge.
datakan•9m ago
> It's possible to be opposed to excessive copyright

Define excessive. Who makes that determination and how?

riedel•
themaninthedark•23m ago
Related: https://news.ycombinator.com/item?id=49068738

https://news.ycombinator.com/item?id=49127284

skybrian•22m ago
It’s true that they’re destroying old books. The claim that they’re destroying rare books seems to be repeated without much in the way of evidence.
postalcoder•20m ago
This reminds me of the hunt for low background steel, down to the destruction of vintage shipwrecks in order to acquire the metal.

https://edconway.substack.com/p/the-eerie-story-of-low-backg...

docheinestages•16m ago
They should at least give the scanned copies to a non-profit or government entity to preserve and release them when the copyright expires.
yallpendantools•16m ago
I'm not really the most AI-friendly person around but this news cycle is just as equally annoying.

"Rare and out-of-print" is fabricating a lot of aura here. It's technically correct (the best kind of correct) but I've yet seen evidence that these are culturally significant copies being destroyed for scanning.

From https://nltimes.nl/2026/06/25/rare-book-dealers-fear-tech-fi...:

> The attachment contained 3,000 English-language titles organized by ISBN number, including books such as Distinct Element Modelling in Geomechanics by K.R. Saxena (1999); Barrett's Traditional Fairy Tales (2021), an academic study of Irish folklore; and Laser Shock Peening of Advanced Ceramics by Pratik Shukla (2018).

The oldest is from 1999! If that's the best they can actually enumerate to bolster this outrage farming cycle you could just wonder how irrelevant the rest are.

Really, please, kill this news cycle. There's a lot of issues deserving proper attention right now and this one is a straight-up nothingburger.

shrubble•14m ago
There are plenty of scanners that do not destroy the book being scanned. Here is one, I am sure there are many others: https://www.youtube.com/watch?v=b9LcTZU-HHI
64e23275635e•10m ago
It should be more than obvious, that the problem are not the scanners.
IshKebab•11m ago
Not all books are sacrosanct. There doesn't seem to be any evidence that they are destroying rare and valuable books.
NishanStepak•7m ago
When Google scanned its books, it did not destroy any of the books. This was because of the idea of the "cultural object" where the book has value as a physical artifact. The experience of reading a physical book is different than reading an electronic text. There is an argument that reading a book in its original form is better because it is closer to the original experience. This can also be said of records and tapes. The experience is designed to work in the original format. A book is a designed object with cultural significance. The container for the words can be as important as the actual words. There should be clearer distinctions and policies between mass market production where items can be destroyed and there will still be many of them left and unique and original content with limited copies available. This is not hard to do. Not to do it shows carelessness.
28304283409234•41m ago
One (1) company destroys and (illegally?) copies a rare book. This knowledge is now of that company, not of humanity. Other companies will feel the need to do the same and before you know it knowledge is walled off in the gardens of the AI companies, and the rare books are gone and destroyed.
pfdietz•36m ago
The scan doesn't go away, and can be released when the copyright expires.

Oh no, a lump of cellulose is gone. Will no one think of the fibers?

collinmcnulty•14m ago
> can be released when the copyright expires

But it ... won't be? The companies have no motivation to do so. The copyright on lots of these books are surely already expired, so if they wanted to they could be putting these up now. I'm sure internally this is viewed as a corpus of knowledge they have that their competitors do not, so they will not release them unless something forces them to.

breezybottom•37m ago
It's going to be hard for you to do more of that, considering that AI forms are destroying the books. You know, the whole point of the article...
Sha1rholder•36m ago
> Roughly everyone can now benefit from its content.

You mean every Anthropic?

datakan•18m ago
Entropy is a thing.

Ancient Rome stopped creating aqueducts because they had all the ones they needed. They failed to pass that knowledge on to the next generation and so they just forgot how to create aqueducts.

Once AI starts making the majority of content, people will simply forget how to make content. Before we know it we're all fat slobs in floating chairs like in Wall-E.

Best case scenario I think is similar to what we see in Ian M. Banks Culture books where the machines basically take care of us out of the goodness of their hearts and we just kinda fuck off into obscurity.

lapcat•12m ago
> If the companies are willing to make the digital version available this may actually prreserve the books.

They aren't. Point me to the links to the digital versions of the originals.

On the contrary, they're mangling the originals: "Character.ai, for example, offers a Books feature that permits users to rewrite public domain works such as “Pride and Prejudice” and “Frankenstein” by changing endings, settings, and inserting themselves into the narrative."

Anthropic: "We use a ‘soft codename’ for it because we don’t want it to be known that we are working on this." If destroying old books were a public service, Anthropic wouldn't need to hide it.

8m ago
People must realize that it is easy to built some sentiment against on part of legislation that seems to restrict them, however, it is much more difficult to establish legislation that serves everyone. In the end we need smart people in politics and lawmaking instead of populists and libertarian cry-babies.
myaccountonhn•8m ago
I imagine you're in favor of the artists being compensated in some other way then? Or should they just suck it up?