Remix.run Logo
xpct 8 hours ago

If I recall correctly, there were some papers which suggested that LLMs favor LLM-generated passages over human written ones. I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :) I've also experienced that both Claude and Codex routinely include generated websites when I ask them to search for something. It also doesn't help that the web search tools that OAI and Anthropic have are deeply limiting: can't exclude keywords or domains.

supriyo-biswas 8 hours ago | parent | next [-]

The other day, I remember an article was posted to HN about something, but it came from a company that provides SEO services to companies by doing something like this:

1. For a given company, analyze their target audiences and the questions they are likely to ask LLMs about.

2. For each such question, ask it to each of the major LLMs, and compute the KL divergence between the pages they want to rank for the question vs. the LLM's response.

3. Rewrite the article to minimize said KL divergence.

In effect, they're performing an iterative optimization of some sort that moves the embedding space of their article closer to the question asked to the LLM, and any embedding model or generated responses are going to prefer said responses over others.

I believe we will keep seeing more of this stuff.

SoftTalker 7 hours ago | parent | next [-]

And it's all because of ads. The incentives in an ad-funded internet are just always going to lead to this sort of thing. The most important thing is getting the user to load your page, not actually satisfying their query.

Let's hope the LLM model continues to be paying for credits, because any that move to ad revenue will become useless for real work.

rectang 4 hours ago | parent | next [-]

> Let's hope the LLM model continues to be paying for credits

LLM vendors make this hard because you can't trust them with your session data. Yesterday you were opted out of training, then suddenly today you're opted in.

It's an extension of the idea that they don't need to care about anybody's copyright. They don't care about preserving the security or privacy of customer data, because there is negligible incentive to do so.

For now, there's no substitute but as LLMs get commoditized trusting LLM SAAS vendors becomes an unacceptable business risk.

locknitpicker 7 hours ago | parent | prev [-]

> And it's all because of ads.

Not really. A while ago there was a news piece stating that Israel was behind a series of fake think-tanks with very accessible websites which were created with the express purpose of feeding AI agents with alternative facts aligned with their foreign policy.

If anyone has the link at hand, please post it.

tencentshill 7 hours ago | parent | next [-]

https://www.theguardian.com/world/2026/aug/26/fake-thinktank...

locknitpicker 7 hours ago | parent [-]

Thank you.

Past HN discussions

https://news.ycombinator.com/item?id=49337392 (884 comments)

https://news.ycombinator.com/item?id=49313477

https://news.ycombinator.com/item?id=49447600

trimethylpurine 9 minutes ago | parent | prev | next [-]

[delayed]

xenadu02 5 hours ago | parent | prev | next [-]

Some of us have been predicting this for a while.

People with an axe to grind or states with an agenda are already devoting tremendous effort toward affecting LLM models and it is very difficult to determine real from astroturf for humans let alone an LLM trying to train.

Much like PageRank now that the cat's out of the bag all the current approaches may prove to be useless in the long run.

grumbelbart2 6 hours ago | parent | prev | next [-]

Sure maybe, but the online ad market is USD 400b give or take, which vastly outmatches any such budgets.

kspacewalk2 2 hours ago | parent | prev | next [-]

If by "not really" you mean it's not all because of ads, some governments dabble in it too, you're right. But it's still overwhelmingly because of ads.

csallen 32 minutes ago | parent [-]

It's just incentives in general. Humans are self-motivated creatures, and in the absence of meaningful consequences they'll often do what they're incentivized to do, even if it harms others.

The internet is uniquely devoid of consequences (esp. reputational consequences, social faux pas, etc.) and makes effort expenditure minimal. So you get lots of bad behavior.

I think "ads vs not ads" is maybe the wrong way to model it. Ultimately people are just doing what benefits themselves across every dimension possible.

bjt 7 hours ago | parent | prev | next [-]

There are incentives other than marketing, but I doubt the SEO company referred to by the ancestor comment is planning to do lots of business in the political propaganda space.

SoftTalker 7 hours ago | parent | prev | next [-]

That sounds very plausible too. Propaganda/spin has always been a component of mass media.

astura 5 hours ago | parent | prev | next [-]

>If anyone has the link at hand, please post it.

https://www.theguardian.com/world/2026/aug/26/fake-thinktank...

monster_truck 6 hours ago | parent | prev [-]

Same difference tbh

boilerupnc 7 hours ago | parent | prev [-]

There is a term for the general practice of optimizing responses called GEO - Generative Engine Optimization [0]. A cousin of SEO and equally unsavory in how trust is being eroded through info shaping. Self-discovery by individuals is the victim.

0: https://en.wikipedia.org/wiki/Generative_engine_optimization

cainxinth 6 hours ago | parent | prev | next [-]

Take a random essay and add in a bunch of the phrases that LLMs love like “load-bearing,” “crucial,” structural,” and “woven,” and then submit the original and the edited version to an LLM and ask which is better. It will choose the second one virtually every time. They have ingrained biases that associate those words with good writing and arguments.

lelanthran 3 hours ago | parent | prev | next [-]

> I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful. It always picks its own :)

A better question to ask for each snippet is "Estimate the seniority and competence of the developer who wrote the following code, ignoring bugs that linters or LLMs can catch and focus only on structure, maintainability, logical layout and readability."

It almost always estimates the author of my code as above the author of it's own code.

bastawhiz 2 hours ago | parent | prev | next [-]

I don't have an oai subscription to try, but I'd be interested to know if Codex picks Claude's code over a human's and vice versa.

jasonjmcghee 8 hours ago | parent | prev | next [-]

If by root urls you mean domains, openai at least supports this.

https://developers.openai.com/api/docs/guides/tools-web-sear...

xpct 8 hours ago | parent [-]

That is what I meant! Couldn't remember the word 'domain' while I was writing out my comment. Thank you

dgellow 7 hours ago | parent | prev | next [-]

A bit different, but one thing I’ve seen is models repackaging Reddit slop. Like, it will do a search, find a Reddit thread somewhat related where someone in a comment casually mentioned incorrect information that any human would have dismissed. The model takes that as granted, but expands on it and present it as a well established fact, presented in a very plausible fashion.

In general I don’t find models to be good at evaluating the quality of a source :(

iamacyborg 6 hours ago | parent | next [-]

Newspapers have been doing this for a long time, notably the Metro in London.

The_Blade 7 hours ago | parent | prev [-]

i actively assume it is worse since, for example, spez signed a 60 million dollar deal to give Google access to the firehose. so then if you have niche, highly engaged subreddits infested by AI bots creating posts, then commenting on posts, then being trained on that content... you have Ouroburos eating its own poop, and models have less then zero incentive to evaluate the quality of a source, especially if they are the source

Retr0id 2 hours ago | parent | prev | next [-]

I was giving local models a try recently, I think it was Qwen 3.6 I was trying at the time. I gave it a codebase and just asked it to review it. Its main feedback was that the comments and documentation were excellently written, but they were all Opus 5 slop.

coldtea 8 hours ago | parent | prev | next [-]

>It always picks its own

Makes sense to me, in that its own output would align closer to its own training set

keeda 6 hours ago | parent | prev | next [-]

OMG if this is true, do you realize what this means? The easiest way to do AI SEO is to generate all your content with AI, and we've seen what SEO does to the web...

The Internet is doomed. Time to start some human-only darknets.

sodapopcan 6 hours ago | parent [-]

SEO is what ruined the web AFAIC.

> Time to start some human-only darknets.

I know very little about darknets. How could you ensure that they are human-only?

thatjoeoverthr 7 hours ago | parent | prev | next [-]

> always picks its own

If you hate AI writing enough, this turns AI filters into a kind of humiliation ritual. AI will derank normal business writing for human readers, and uprank inflated, verbose, tic-heavy slop. So you have to put the heavy slop out with your name on it. Really perverse moment.

mistrial9 7 hours ago | parent | next [-]

disagree that there is one kind of ranking and one kind of engine analyzing that ranking; sort of de-facto true that one company does run the ad world; strongly agree that this is a nightmare possibility and directly dystopian

cindyllm 7 hours ago | parent | prev [-]

[dead]

lo_zamoyski 8 hours ago | parent | prev | next [-]

> asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored [...] It always picks its own

...is not the same as claiming...

> LLMs favor LLM-generated passages over human written ones

Here, you're using the same LLM to both produce and judge the resulting work. If anything, I would expect an LLM to tend to prefer its own work given that the same training is producing and judging.

xpct 8 hours ago | parent [-]

It's not intuitive to me for why preference for its own writing would emerge, and during what type of training or tuning.

Perhaps something like: learning to identify what source files it has worked on by the code style alone, because tasks may give human code (public repos, etc) and ask to make changes.

freeone3000 6 hours ago | parent | next [-]

It’s optimizing for good writing. Therefore, it believes its outputs are good. Therefore, it believes inputs that look like its outputs are good.

pixl97 7 hours ago | parent | prev [-]

It would need to be researched, but I wonder if it ends up being something that happens at the token level?

DarmokTanagra 8 hours ago | parent | prev | next [-]

If that were true I would expect to see prose that more closely resembles the "caveman" messages found in the HuggingFace attack than the overly flowery nonsense we see in AI blogspam.

cortesoft 7 hours ago | parent | prev | next [-]

Well, the one it generated is based on how it thought the best way to solve the problem was.

I am sure most humans would pick code written in their style, too.

Wowfunhappy 8 hours ago | parent | prev [-]

> I can consistently reproduce this by asking Claude which code snippet it prefers: the one it generated in a different chat, or one that I refactored for my own needs and find more useful.

Interesting. For me I've noticed it tends to do the opposite.