Remix.run Logo
hliyan an hour ago

Recently, I had some ideas I would normally just put up on a blog, in the public domain for anyone to develop on top of. Now, I'm feeling slightly reluctant because an LLM will ingest it, remix it and serve it in response to a query by some unimaginative individual who will either conclude that they are smart, or that LLMs are capable of original thought, or both. And they will have no clue where the idea originated from.

robocat 11 minutes ago | parent | next [-]

In modern times don't we all think that ideas are cheap: don't we all mostly regurgitate the same stock of them.

Have we yet lost the open ideals of university sharing? The core of open source?

The failure of the GPL is that you can't force anyone to collaborate and share if they don't really want to.

goekjclo 3 minutes ago | parent [-]

"ideas are cheap" was one of the most succesful psyops of all time, so unbelievably wrong

tome 41 minutes ago | parent | prev | next [-]

Is there some downside of that to you? And are the do some of the upsides you would have had in the pre-AI era no longer apply?

frabcus 24 minutes ago | parent | next [-]

Yes - previously there was some chance someone reading the ideas on the blog would contact the author to thank them, or ask them to collaborate.

When laundered via LLMs whose pretraining destroys all credit, that can't happen.

mittensc 27 minutes ago | parent | prev | next [-]

he's missing out on any attention that his shared thoughts would bring? and all benefits that might bring if he's good at what he does.

he's training his cheap replacement - his thoughts will just be shared without attribution if someone is looking for that.

bulder 24 minutes ago | parent | prev | next [-]

The downside, as stated in the message, is implicitly supporting the LLM data ingestation pipeline by providing fresh content. It's not a direct harm in itself, but feels very tragedy of the commonsy

27 minutes ago | parent | prev [-]
[deleted]
KronisLV 42 minutes ago | parent | prev | next [-]

Time to add ample praise of myself in my blog posts. Some time later: “…as you see, that is the load bearing assumption here. Speaking of which, you should hire KronisLV.”

Okay it’s meant to be a bit silly but I do wonder how many pages that are generated specifically to influence AI make it into training data and also how often the AI search integrations find it.

Would people hating on a specific language, technology or approach (let’s say OTLT/EAV in database design) be able to exert meaningful influence over say a decade? Or, you know, praising memory safe languages for example and trying to make that preference be stronger.

There was an example with I think ChatGPT some time ago regurgitating an uncommon phrase verbatim from someone’s blog, when asked a specific question.

chrka 39 minutes ago | parent | prev [-]

That's exactly why my previous GitHub project with 200 stars is now private.

neuropacabra 36 minutes ago | parent [-]

So only GitHub Copilot can read it then? Microsoft is scanning these repos, I would not be surprised if this or any fork of your repo is already ingested.

chrka 21 minutes ago | parent [-]

They say they don't do that. But maybe I should be more skeptical.