frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

A domain can now say it is for sale, in DNS

https://specification.website/spec/foundations/for-sale-dns/
75•shaunpud•1h ago•41 comments

DeepMind's WeatherNext model achieves breakthrough forecasting cyclones

https://deepmind.google/blog/weathernext-ai-model-achieves-breakthrough-in-forecasting-cyclones/
196•bhavansig•6h ago•59 comments

Gateway 2000's hilariously bad ads in the 90s (Part II)

https://buttondown.com/suchbadtechads/archive/gateway-2000-part-2/
26•rfarley04•2h ago•17 comments

Voyager 1 FDS Computer Emulator

https://zaneham.github.io/voyager-fds-emulator/
21•rahen•1h ago•3 comments

Don't use your phone while you poop

https://nate.spot/no-phone-while-poop/
85•ExMachina73•1h ago•90 comments

A physicist rigged his pet hamster’s wheel to upload to Strava

https://www.runnersworld.com/news/a73355106/hamster-wheel-strava-running/
339•aanet•2d ago•78 comments

Hardware backdoors in some x86 CPUs

https://github.com/xoreaxeaxeax/rosenbridge
237•epestr•8h ago•73 comments

Gentoo bugzilla closed due AI bot scraper overload

https://social.treehouse.systems/@mgorny/117058483039362779
39•happosai•1h ago•9 comments

DeepSeek V4 Flash 0731

https://arcprize.org/results/deepseek-v4-flash-0731
706•tosh•21h ago•429 comments

BYOC Is Not Just 'Deploy into Their Cloud'

https://omnistrate.com/blog/byoc-anywhere-the-spectrum-of-bring-your-own-cloud-deployments
18•kkgupta•4d ago•1 comments

U.S. Department of Energy Launches the Genesis Open Models Initiative

https://genesisopenmodels.anl.gov/
294•moelf•17h ago•113 comments

What happens if an entire class of workers loses faith in their careers

https://www.noemamag.com/why-is-everyone-in-tech-so-sad/
838•RickJWagner•1d ago•931 comments

US Military's cyber command unit grapples with cluster of deaths by suicide

https://www.bloomberg.com/news/articles/2026-08-06/us-military-s-cyber-command-unit-grapples-with...
51•rbanffy•5h ago•62 comments

k-Coloring is Faster than Computing the Chromatic Number

https://arxiv.org/abs/2607.25973
35•matt_d•1w ago•4 comments

Assembly Hall of Shame

https://github.com/xoreaxeaxeax/asm-hall-of-shame
382•piotrgrabowski•21h ago•94 comments

Europe's free satellite service just made it easier to track wildfires

https://arstechnica.com/gadgets/2026/08/europes-free-satellite-service-just-made-it-easier-to-tra...
97•01-_-•5h ago•9 comments

ao486: x86-compatible Verilog core implementing all features of a 486 SX (2014)

https://github.com/alfikpl/ao486
50•csmantle•6d ago•7 comments

Microsoft Edge is about to lock out older ad blockers, just like Chrome did

https://www.theverge.com/tech/976880/microsoft-edge-extensions-ad-blockers-mv2-mv3
74•eternalreturn•5h ago•59 comments

Show HN: Rotating torus in terminal (but with kitty graphics protocol)

https://andreadimatteo.com/torus-v0-5.html
12•deotman•6d ago•2 comments

Ancient Library – 1,060 Greek/Latin texts, click any word to parse it

https://ancientlibrary.net/
231•aagha•20h ago•73 comments

Sensitive Info Goes into 'No Reply' Emails Constantly. This Guy Sees It All

https://www.wired.com/story/sensitive-info-goes-into-no-reply-emails-constantly-this-guy-sees-it-...
13•sbulaev•1h ago•2 comments

NASA figured out how to keep its Voyager 2 probe running for another year

https://www.space.com/space-exploration/voyager/nasa-figured-out-how-to-keep-its-48-year-old-voya...
314•wglb•13h ago•67 comments

Lost my phone at the office. Claude suggested tracking Bluetooth signal strength

https://twitter.com/un1c0rnioz/status/2084686552299634805
63•ilamont•18h ago•48 comments

2027 memory capacity is reportedly sold out

https://www.ign.com/articles/ramageddon-continues-another-year-as-2027-memory-capacity-is-reporte...
437•inigyou•1d ago•418 comments

SupererDuperer

https://www.shirtpocket.com/blog/supererduperer
139•zdw•6d ago•29 comments

Timeline of the OpenAI accidental attack against Hugging Face

https://simonwillison.net/2026/Aug/7/openai-timeline/
128•882542F3884314B•4h ago•143 comments

Managing AI Coding Costs at Scale

https://www.databricks.com/blog/managing-ai-coding-costs-scale
278•moonikakiss•20h ago•230 comments

From One Seed to a Thousand Leaves – Merkle's Authentication Tree

https://0xkrt26.github.io/math_behind_security/2026/08/03/merkle-tree.html
35•denismenace•4d ago•3 comments

The Nixpkgs core team has disbanded

https://discourse.nixos.org/t/the-nixpkgs-core-team-has-disbanded/79413
353•Meleagris•14h ago•174 comments

Carl's Required Reading

https://carlkolon.com/reading/
210•cckolon•6d ago•26 comments
Open in hackernews

Gentoo bugzilla closed due AI bot scraper overload

https://social.treehouse.systems/@mgorny/117058483039362779
39•happosai•1h ago

Comments

c-fe•47m ago
going on a bit of a tangent here - some discussion there seems to be imply that AI bots mostly use IPv4s, which makes sense to me given that they are probably bots hosted by some cloud. Whereas IPv6 may be more organic traffic from e.g. mobile users (not in this case probably)... which made me think if at some point IPv6 may at some point win against IPv4 just because its more organic traffic (i.e. not from big tech cloud), leading to pages blocking IPv4? Just speculation on my side
neuroticnews25•28m ago
This is weird, every VPS I've rented in the last 8 years had IPv6, and some didn't even had IPv4.
Retr0id•23m ago
The hardest-to-mitigate bot traffic tends to come from residential proxies, and most residential connections are still v4-only.

However, it is also easy to get large numbers of v6 addresses cheaply.

mrweasel•37m ago
While we're dealing with the same issue at work, I sometimes still wonder exactly who these scrapers are. OpenAI, Google, Anthropic and others are normally fairly well behaved (Minus Anthropic attempting to hide behind a browser-for-hire company). Mostly you can get IP range and user-agents for the large players, while it problems mostly stem from bots pretending to be Chrome.

Our largest offenders seems to be mostly limited to South-East Asia, so probably mostly Chinese AI projects, but that's speculation. I also don't recall ever seeing Grok IP ranges or a specific Grok UA, but that doesn't mean that they're hiding, perhaps they're just not interested.

marginalia_nu•9m ago
Most of the nonsense comes through botnets/residential proxies, so it's very hard to say.
shevy-java•7m ago
Well, some websites claimed China is behind it, which could make sense (I would not know either way). At the same time, though, I kind of doubt your carte blanche here for all those companies. Why would you think none of them are responsible for the AI slop spam?

> Our largest offenders seems to be mostly limited to South-East Asia, so probably mostly Chinese AI projects, but that's speculation

Ok. So you also don't know. Well, I don't know either, but I don't make a speculation by claiming x, y, and z companies to be exempt. In my book they are all responsible.

ComputerPerson•33m ago
There are definitely patterns you can use against the scrapers. This maintainer just didn't have time for it, which is understandable.

We direct scraper traffic to a bot-specific server using Cloudflare's load balancer, slowly analyzing traffic and adding conditions one at a time. No accidental scraper DDoS in a long time.

Most scrapers are relatively honest in some way shape or form.

calvinmorrison•18m ago
The year is 2026, somehow peoples basic web apps are not able to keep up with scrapers. Scrapers are not new. My side load is like .01 even with 10x the traffic of last year. It's called "serving static content", "caching" and many other things that are not new concepts.

Then you've got good old cloudflare which is free to use

shevy-java•9m ago
AI skynet is winning. It is stealing time from humans, thus forcing down their activity, as can be seen here.

A headline would be great if those AI companies would close down. I hold them all responsible for this.