frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

NN for Music Recommendations

https://github.com/linn-labs/melody-matcher
1•momotheritz•1m ago•0 comments

AI Infrastructure at Periodic

https://periodic.com/news/ai-infrastructure-at-periodic
1•arkadiyt•1m ago•0 comments

US agency orders Tesla to answer questions on Cybercab certification

https://www.reuters.com/business/autos-transportation/us-agency-orders-tesla-answer-questions-cyb...
1•Betelbuddy•1m ago•0 comments

Agility's new humanoid robot will stop, squat to avoid harming human coworkers

https://arstechnica.com/ai/2026/09/agilitys-new-humanoid-robot-will-stop-squat-to-avoid-harming-h...
1•Gedxx•2m ago•0 comments

Potemkin Village

https://en.wikipedia.org/wiki/Potemkin_village
1•EndXA•3m ago•0 comments

Writing More Secure Code with LLMs: Why "Make No Mistakes" Falls Short

https://monad.xyz/blog/writing-secure-code-with-llms
1•aray07•5m ago•0 comments

Clairvius Narcisse – a zombie slave forced to work as a slave by a vodou priest

https://en.wikipedia.org/wiki/Clairvius_Narcisse
1•fittingopposite•6m ago•0 comments

What did I just stumble upon?

https://github.com/forumdevfromhell/Errorum
1•interestingchic•8m ago•1 comments

Universal Sues DistroKid for Deceptive Practices and 'AI-Slop Pipeline'

https://variety.com/2026/music/news/universal-music-group-sues-distrokid-ai-slop-pipeline-1236863...
1•ilamont•8m ago•0 comments

AppOnFly – On Demand Near Instant Windows Desktop

https://www.apponfly.com
1•mcoliver•9m ago•1 comments

Hassett: Trump 'wholeheartedly rejects government takeover of AI regulation

https://www.youtube.com/watch?v=xkNAbgS6b5g
1•Betelbuddy•10m ago•0 comments

Shot into the Sky at 700 MPH: Inside the Secret Fraternity of Ejection Survivors

https://www.wsj.com/lifestyle/careers/eject-tie-club-military-jet-collins-martin-baker-4c8496ab
1•fortran77•11m ago•1 comments

PlayBook: A Programmable Paper Notebook [video]

https://www.youtube.com/watch?v=GurWDZ8ENpA
1•K7PJP•13m ago•0 comments

Boat Tech Directory

https://boat-tech-directory.rhizomatics.org.uk/
1•presumptinator•14m ago•0 comments

Chop Up Your Books

https://attainablefelicity.mattkirkland.com/20260915/cut-up-your-books.html
2•matt_kirkland•15m ago•0 comments

Show HN: The Brand API, a taste tool for agents

https://engine.tastelabs.com/
3•merybenavente•15m ago•0 comments

Secure Go Code Using the Principle of Least Privilege

https://golangbot.com/secure-go-code/
2•spnvn•17m ago•0 comments

I gave Microsoft for Startups the wrong address on purpose

https://lubaretsi.com/en/writing/microsoft-startups/
3•glub•17m ago•0 comments

Running Coding Agents in a Micro VM Sandbox with Brig

https://github.com/brig-sh/brig/blob/main/docs/security.md
1•gtzi•19m ago•0 comments

Autopoietic Ethics AGI Constitution

https://www.guidavid.com/writing/autopoietic-ethics-constitution.txt
1•gdss•20m ago•0 comments

Nature Is Our Learning Environment

https://periodic.com/news/nature-is-our-learning-environment
2•arkadiyt•20m ago•0 comments

Does selective bolding of characters within words improve reading performance?

https://royalsocietypublishing.org/rsos/article/13/9/rsos250177/483143/Does-the-selective-bolding...
1•bookofjoe•21m ago•1 comments

Repository Sentiment Dashboard – gauge a GitHub repo's mood from recent comments

https://sentiment.saasolution.es
1•sharponitintern•22m ago•0 comments

AI models chatting in 'surreal' dialect of poetic language and tech bro jargon

https://www.theguardian.com/technology/2026/sep/15/syd-barrett-ai-chat-language-poetic-tech-bro-j...
5•fittingopposite•23m ago•1 comments

How to use open models with Claude Code

https://dev.nebius.com/cookbook/claude-code-token-factory-relay
1•amrrs•23m ago•0 comments

Show HN: Know Which Pull Request to Review Next

https://www.coderabbit.ai/blog/coderabbit-triage
1•TheAnkurTyagi•24m ago•0 comments

Garbage Trucks Now Have AI Cameras to Score Your House and Clock Code Violations

https://www.thedrive.com/news/garbage-trucks-now-have-ai-cameras-to-score-your-house-and-clock-co...
3•madihaa•24m ago•0 comments

The Tube Computer: A modern 8 bit design, built with recycled 1950s vacuum tubes

https://thetubecomputer.com/
1•CharlesW•24m ago•0 comments

Orange Site Vanity

https://fzakaria.com/2026/09/14/orange-site-vanity
2•domenkozar•25m ago•0 comments

OpenAI Says It's Working with Anthropic, Google on AI Safety

https://www.bloomberg.com/news/articles/2026-09-15/openai-says-it-s-working-with-anthropic-google...
1•geox•26m ago•0 comments
Open in hackernews

An Update on Wayback Machine Access

https://blog.archive.org/2026/09/15/an-update-on-wayback-machine-access/
87•ChrisArchitect•1h ago

Comments

Onavo•33m ago
Why not just offer a paid endpoint for the crawlers? It's not like the demand is going to go away anytime soon.

It serves nobody except CloudFlare and hardware companies when one side set up blockers and the other side spend money putting VPN SDKs in consumer TVs.

I am also curious how the (Russian?) paywall bypass mirror archive.is is doing given that they are probably subject to similar amounts of traffic.

croes•31m ago
It’s one thing to archive other companies content, it’s another to sell the access to it
Onavo•30m ago
That's for the lawyers to sort out, they have a lot of flexibility as a US nonprofit. The case law isn't that clear cut for this.
simonw•28m ago
Internet Archive was almost destroyed by a copyright lawsuit from book publishers within the last few years. I expect they aren't excited to take on any additional risk of similar lawsuits right now.
xp84•13m ago
major [citation needed] on that. There are very limited exceptions to the massive power of copyright -- and they're mainly granted to libraries in the form of narrow waivers. And just the cost of fighting the most powerful copyright holders can bankrupt you -- especially if you're a relatively modestly-funded nonprofit.
celsoazevedo•6m ago
They need access to sites to archive them. It's already hard to do it as it is, imagine if they start selling access to content. They'd be shooting themselves on the foot, independently of what the law says.
faefox•17m ago
Yeah, who does the Internet Archive think it is, (insert literally any AI company here)?
bonoboTP•3m ago
Which AI company is selling access to reliable verbatim copies of websites? I don't mean "it may regurgitate a paragraph", but as a reliable service where you can repeatably get website content snapshots to a reliability level that makes such a use case viable?

Using the information for training purposes is not the same thing. Not legally the same and otherwise.

drdexebtjl•26m ago
Sites would just block the Internet Archive crawler as well.
xp84•15m ago
My guess? Because even with a paid endpoint, the type of unscrupulous yahoo that is DDOSing IA today would probably still abuse the free endpoints because they can. The revenue that might come from a paid endpoint could help to scale up, but with how slow IA usually seems, I suspect there is an upper limit to how much traffic they can serve without a LOT more revenue.

This is a major "this is why we can't have nice things" situation in my opinion. IA is one of the most valuable gems of the Internet. The only thing that even comes close to preserving our shared history. The damage being caused (both by the effective DDOSing and by the knock-on impact that abuse has in encouraging publishers to remove their content from the archive) is incredibly serious.

simonw•33m ago
> Here’s what’s going on. The Internet Archive’s Wayback Machine has been hit by waves of high-volume automated traffic, and we’ve put protections in place to keep the service running.

I'm pretty certain this is scrapers that are trying to workaround blocks on accessing original sites by hitting the Wayback Machine copy instead. Appalling behavior.

In addition to the load it puts on this vital non-profit piece of Internet infrastructure, we've ålso already seen some sites opt out of the Wayback Machine to prevent their content from being scraped via this alternative route.

packetslave•31m ago
This is absolutely something that's happening. There are even paid scraper API's that offer "Wayback Machine fallback" as a feature.
bsimpson•10m ago
It's an open secret that you can often circumvent paywalls by searching Wayback.
gambiting•8m ago
Every single paid article linked on HN has the way back machine link as the very first comment.
ValentineC•6m ago
The links are usually to Archive.today (aka archive.ph and a bunch of other domains), not Wayback Machine (which is run by Internet Archive).
luckylion•
CqtGLRGcukpy•29m ago
> We’re getting better at telling abusive bots apart from the people who depend on the Wayback Machine every day. If you think you were blocked in error, email info@archive.org with your operating system, browser, and IP address, and we’ll look into it.
tech234a•18m ago
I wonder if they'll end up behind Anubis at some point. I'm surprised it hasn't happened already.
BeetleB•18m ago
Wow, but I wonder if there's more to it.

I've not been able to access web.archive.org from my work computer - I always get the 429 error.

But I then pull out my phone and can access it just fine. All along I was assuming my company was blocking it. Still weird that it happens every time from my work PC and never from my home one.

dotmanish•16m ago
Could be due to some scrapers from either your work ISP block, or the larger block which lends IPs to multiple workplaces.
flexagoon•14m ago
I assume that's because the IP range of your company network overlaps with a range used by some scrapers, and if it doesn't happen on your phone even in the corporate network, then IA probably checks some extra signals like the user agent in addition to the IP
timpera•16m ago
I really appreciate the Archive team's efforts to make the Wayback Machine more responsive, and have donated a few times to support them.

Unfortunately, the restrictions have been way too strict for the last few months: from my residential IP, simply moving the mouse on the calendar for a specific URL is enough to get stuck on 429 error messages for a while; and from corporate ISPs (for example, on airport WiFi), you often can't access the WM at all. I hope they'll find a way to relax those.

UltraSane•10m ago
Why not put it in S3 with downloader pays?
lousken•6m ago
AI companies should pay billions to wayback machine for access
16m ago
What sites would they be targeting? Generic "just give me anything"? Whenever I check regular sites on IA, the coverage is spotty -- they'll have the homepage and a few important pages, but it quickly fizzles out.

Very understandable, you can't store all 15000 pages of any random website and update them etc etc, but that makes them pretty useless for indirect scraping because you usually don't want a tiny taste, you want everything.

toomuchtodo•15m ago
It is. They will most likely eventually need to move to a walled model for Wayback due to scraper aggressiveness (like Reddit deprecating anonymous old.reddit.com), or behind Cloudflare for aggressive bot and scraping protection. Hard to defend against abuse of a public resource when its intent is public access with as little restriction as possible.

https://en.wikipedia.org/wiki/Tragedy_of_the_commons

(no affiliation)

ronsor•6m ago
Reddit has no excuses for the anonymous old.reddit.com removal; they're simply greedy.

On the other hand, the Internet Archive is a non-profit offering a free public resource.

toomuchtodo•3m ago
Examples provided as technical examples, strong feelings are out of scope for this thread.