frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Three sites made 215,128 "best software" pages for AI. Perplexity cites them

https://trellner.com/reports/manufactured-sources-behind-ai-recommendations/
44•jakobgreenfeld•47m ago

Comments

sph•30m ago
What protection do LLM search engines have against training off content generated by other LLMs?

Will we get to a point where AI-generated sites make up a majority of the internet, and LLMs are training upon their own regurgitations, with exponential amplification of all their lies and flaws?

Or will the pre-2022 corpus human knowledge be considered the low-background steel standard, and anything after that less and less reliable unless certified that it has been created by a human mind and untainted by hallucinations?

NegativeLatency•24m ago
They’ll train on prompts and anything else you send in. Many LLM responses are sorta finger printable: I assume this is intentional
creaturemachine•21m ago
I have a feeling we're already there.
giancarlostoro•17m ago
The weird babbling reported from Opus 5 might be a result of either a bad system prompt or bad training data.
gdulli•15m ago
> What protection do LLM search engines have against training off content generated by other LLMs?

You're talking about a scenario that won't blow itself up in the next few quarters, so it's of no interest to them.

coldpie•7m ago
I've mostly stopped using the Internet to learn new things and have gone back to books from the library. The majority is published pre-2020s and hopefully, publishing slop physically won't be profitable enough to flood that market, too.
antiloper•30m ago
Searching for products has become impossible. If you don't already know what you are looking for, you're screwed.
a2ff6eeb0•29m ago
Makes sense. Manipulating training data so that models will recommend your product is undoubtedly a big industry.
lukev•28m ago
Begun, the AI SEO wars have.
pietroppeter•19m ago
Has the industry already started to search for a better name than AI SEO for this?
properbrew•11m ago
Yes GEO (Generative Engine Optimisation) is one I've seen around.
pupppet•5m ago
Answer engine optimization (AEO)
marginalia_nu•4m ago
It does seem rather impactful.

I've seen an extremely aggressive uptick in API key requests and sales that I'm not sure where it's coming from. Like it's up 5x over the summer. Been a bit confused about this since I do basically zero traditional marketing or SEO, but I think it's AI search tools that's suggesting my services.

jpimbert•28m ago
It's difficult to read more than a few sentences, when this itself is clearly a Claude artifact.
bensyverson•27m ago
An SEO tale as old as time
Aurornis•26m ago
I used one of the 12-month free Perplexity offers when they were everywhere. It felt slightly useful at first for simple queries where I didn’t want to go through the top 10 Google results manually. If I was looking for a specific recipe I remembered or a help page or user manual it would usually find it quickly.

Then they started optimizing for speed of responses over quality of results. I can enter a query and see my results appear in a second, but they’re garbage. The links and references it gives frequently don’t match the text right next to them. It feels like someone had a KPI to make responses as fast as possible and they optimized for that above all else.

They added a “Computer” option that’s supposed to do research for you. Half the time I can’t get it to trigger through the UI. Pressing the submit button doesn’t work. When I can get it to trigger, most of those sessions will work for a while and then just stop before an answer comes back.

The only reason I keep using it is to keep observing a company that has been heavily marketed and hyped, which should have had a market leading position for something. Even non-technical people I know who listen to Joe Rogan (where Perlexity is advertising heavily, I’m told) are asking me about it.

Now there are reports of people being billed at the end of their trial period without warning, despite them saying that they will warn before this happens. There are some alarmingly bad customer support screenshots where the customer support agent (AI? Probably) acknowledges that they didn’t send the email they promised but refuse to help anyway. It takes escalating it on Twitter to get it corrected.

If I want to do actual research or AI assisted web searching I have Claude or ChatGPT do it. The results are so much higher quality and it does exactly what I ask. It may take 45 seconds instead of the instant response from Perplexity but I save time overall because the response and links are more likely to be correct

giancarlostoro•18m ago
If they had some sort of tier that was like $5 to $10 and the main thing it had on it was their custom model for research, no other model, I might re-subscribe, but they were definitely burning way too much compute trying to give it away in the hopes others kept their subscription. I was using it to trim down on my direct Claude Compute usage since they ran me unmetered for a while.

I would draft a development plan with Claude on there, then feed it to Claude Code. This isn't sustainable, but given that I had x number of months pre-paid for, I just used it.

qweqwe14•25m ago
AI;DR
alangibson•22m ago
Perplexity is about to learn that Google is an anti-spam company first, search engine second
threetonesun•14m ago
Well, was. They did a bad job of it these last few years which allowed any AI that could crawl the web to seem amazing for search because it could pick the best posts from Reddit or whatever other forum had the best context for your question, but now we're watching the AI snake eat its own tail.
marginalia_nu•12m ago
Hard to conclusively beat spam when your primary means of making money is selling the very ads that the spammers are using to make money.
jeffreyrogers•5m ago
I rarely use image search, but I went to look something up recently and I was shocked at how many obviously AI generated images showed up. I couldn't even find an image of the thing I was looking for and eventually gave up. Bing has the same problem.
dominotw•19m ago
my friend works for a company called 'profound' whose whole job is 'get found by ai' by spamming reddit and other talk sites ( among other things)
xpct•15m ago
If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I've also experienced that both Claude and Codex routinely include generated websites when I ask them to search for something. It also doesn't help that the web search tools that OAI and Anthropic have are deeply limiting: can't exclude keywords or root URLs.
Wowfunhappy•12m ago
> I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful.

Interesting. For me I've noticed it tends to do the opposite.

jasonjmcghee•7m ago
If by root urls you mean domains, openai at least supports this.

https://developers.openai.com/api/docs/guides/tools-web-sear...

xpct•3m ago
That is what I meant! Couldn't remember the word 'domain' while I was writing out my comment. Thank you
lo_zamoyski•4m ago
> asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored [...] It always picks its own

...is not the same as claiming...

> LLMs favor LLM-generated passages over human written ones

Here, you're using the same LLM to both produce and judge the resulting work. If anything, I would expect an LLM to tend to prefer its own work given that the same training is producing and judging.

pietz•10m ago
The irony of this article being fully AI generated...

Anyway, it's over for Perplexity. They never had a great a product and the only reason for using them, was when they offered Pro accounts for free. Many people joined. Me included. But with a "meh" product and the general AI business not being very sticky, they lost quite harshly.

I thought they might be able to make money as a search api/index, but this article closed the book.

rcar1046•9m ago
"Sharing a nameserver pair is strong circumstantial evidence of a common Cloudflare account rather than proof of ownership"

-when you read one statement that let's you know to believe no other assertions in the article....

stranded22•10m ago
I paid for perplexity pro for 3 years. I genuinely enjoyed using it and felt it was better overall than ChatGPT etc due to the way it showed sources etc. I liked being able to use different models depending on what I was looking for, and the deep research was helpful.

I think they probably damaged themselves by going for a land grab of user base through freebies. It meant the users weren’t ever going to convert to paid customers, so it was more to show investors that they had a user base. But, with an increased base of users who weren’t paying, it then meant they needed to find either new revenue streams or cheaper ways to provide the service. Unfortunately, it seems they went with the new revenue streams whilst also decreasing the functions paying members were able to access (something I find quite abhorrent- I paid a service level, but then they change what I receive mid-subscription). And then computer - rammed down my throat. One reason I pay for pro is to stop the nagging noise of paid tiers. And instead, they actually created a way of logging in and continually seeing gated functions.

So, after paying them upwards of $400-$500 and being a loyal customer, I walked.

cheesecakegood•6m ago
I’ve personally found that Google’s AI mode is surprisingly capable and almost absurdly fast, although the hallucination rate is somewhat proportionally higher too to match its seeming over-responsiveness. Which means Perplexity in effect doesn’t have anything to differentiate it.

A Note from LWN

https://lwn.net/Articles/1090585/
272•rwky•1h ago•42 comments

Three sites made 215,128 "best software" pages for AI. Perplexity cites them

https://trellner.com/reports/manufactured-sources-behind-ai-recommendations/
48•jakobgreenfeld•47m ago•34 comments

GrapheneOS says Pixel 11 has MTE support after all

https://grapheneos.social/@GrapheneOS/117194007157499435
35•user_7832•46m ago•16 comments

Poisson Disk Sampling

https://stripeacross.com/posts/poisson-disk-sampling/
26•vismit2000•59m ago•1 comments

Biggest dark matter detector spots a single weird particle

https://www.science.org/content/article/world-s-biggest-dark-matter-detector-spots-single-weird-p...
30•randycupertino•1h ago•0 comments

Mistral now trains on user input by default, except on enterprise tier

https://help.mistral.ai/en/articles/455207-can-i-opt-out-of-my-input-or-output-data-being-used-fo...
77•teekert•2h ago•42 comments

Commodore 64 released September 1, 1982

https://dfarq.homeip.net/commodore-64-released-september-1-1982/
208•giuliomagnifico•6h ago•108 comments

Aging Brains Blend Memories Together Instead of Just Forgetting Them

https://studyfinds.com/aging-brains-blend-memories-together-instead-of-forgetting-them-study-finds/
38•mdp2021•1h ago•12 comments

Six curl CVEs after OpenAI and Anthropic came back with zero

https://aisle.com/blog/aisle-discovered-six-curl-cves-after-openai-and-anthropic-found-zero
37•goobreee•1h ago•11 comments

WebLLM: high-performance in-browser LLM inference engine

https://github.com/mlc-ai/web-llm
9•saikatsg•44m ago•1 comments

Exit the Cave

https://turtlespace.blog/p/exit-the-cave
9•akkartik•30m ago•0 comments

Dutch central bank moves share of gold from U.S., Canada to London

https://nltimes.nl/2026/09/02/dutch-central-bank-moves-share-gold-us-canada-london-cites-instability
81•TechTechTech•1h ago•63 comments

A Beginner's Deep Dive Guide to Entra Passkeys

https://emsroute.com/2026/03/19/passkeys-beginners-101/
11•speckx•1h ago•8 comments

The Emergent Symbolic Structure of Artificial Neural Networks

https://arxiv.org/abs/2608.29530
226•schmuhblaster•10h ago•77 comments

Dyson CameraJet: The only toothbrush with a camera and a jet

https://www.dyson.com/discover/news/latest/introducing-camerajet
15•mgh2•59m ago•19 comments

Ending my elixir exploratory writing

https://lucassifoni.info/blog/the-end-of-this-elixir-log/
35•ffin•2h ago•19 comments

LLMs: Intelligence vs. Cost

https://openteams.com/intelligence-vs-cost/
13•theanonymousone•1h ago•4 comments

Telli (YC F24) is hiring engineers and designers [Berlin, on-site]

https://careers.telli.com/
1•sebselassie•6h ago

Check if a file was made with Claude

https://claude.com/check-content
28•frexs•2h ago•18 comments

Quasar 438B: Europe's Leading AI Model

https://multiversecomputing.com/resources/introducing-quasar-438b-europe-s-leading-ai-model
90•amunozo•4h ago•73 comments

Using jq to format JSON on the clipboard

https://chris48s.github.io/blogmarks/posts/2021/jsontidy/
20•stefanvdw1•2h ago•4 comments

The Cables That Connect the World

https://xn--gckvb8fzb.com/the-cables-that-connect-the-world/
16•surprisetalk•1h ago•2 comments

Move in C++ without a std:move

https://andreasfertig.com/blog/2026/09/move-in-cpp-without-a-stdmove/
26•dalvrosa•1d ago•22 comments

Banca Etica Suspends A/I's Account While Condemning the Sanctions Behind It

https://sabot.media/post/banca-etica-statement-english
32•rendx•4h ago•19 comments

Why humanoid robots won't catch up to human workers any time soon

https://www.understandingai.org/p/why-humanoid-robots-wont-catch-up
13•speckx•1h ago•8 comments

Just bury your trash: What if everything we know about recycling is wrong?

https://worksinprogress.co/issue/just-bury-your-trash/
22•magoghm•36m ago•14 comments

It's OK to hardcode feature flags (2025)

https://code.mendhak.com/hardcode-feature-flags/
48•biscuits1•3h ago•27 comments

The Lost Art of Carrying Loads

https://www.carryology.com/insights/the-lost-art-of-carrying-loads/
36•surprisetalk•2d ago•39 comments

A Small Telescope That Surprised Me

https://adfr.io/thoughts/20260831_a_small_telescope_that_surprised_me/
62•speckx•23h ago•52 comments

You Know Who Hates AI? Insurance Claims Adjusters

https://www.wired.com/story/insurance-claims-adjusters-really-hate-ai/
113•joozio•2d ago•84 comments