Remix.run Logo
yellowapple 4 hours ago

Makes sense to me. AI models are trained on (as large of a subset as possible of) the sum of human knowledge, so their outputs should belong to humanity as a whole.

Really so should all creative works, on the same basis of all creative expression being the product of the society and civilization which fundamentally and inescapably influenced the creator — and they would belong to humanity as a whole, if it wasn't for intellectual property systems demanding the removal of ideas from the commons.

gruez 4 hours ago | parent | next [-]

>Makes sense to me. AI models are trained on (as large of a subset as possible of) the sum of human knowledge, so their outputs should belong to humanity as a whole.

No it doesn't, because that would mean any sort of secondary source shouldn't be eligible for copyright either, eg. encyclopedias, which are basically rehashing "the sum of human knowledge".

solid_fuel 4 hours ago | parent | next [-]

Interesting POV. Remind me - did those encyclopedias pay experts to write their contents, or did they just hoover up every bit of text they could find - regardless of owner - and toss it into a blender?

jujube3 4 hours ago | parent | next [-]

Sure, I'll remind you. The most popular encyclopedia of our time, Wikipedia, doesn't pay most of its contributors (I think they do have some administrative staff or something?). Nor do they pay journalists or authors for the articles and books that they cite.

solid_fuel 4 hours ago | parent [-]

Interesting, interesting. And this wikipedia - it gets its contents by hoovering up the web, then? Or do you maybe want to put on your thinking hat and consider the difference between voluntary contributions and theft?

jujube3 3 hours ago | parent [-]

Yes, Wikipedia gets its content largely by hovering up the web, without the consent of the authors. For example you can cite a New York Times article in Wikipedia, without getting the consent of the NYT author. Wikipedia also "hoovers up" (as you put it) offline sources like books. Again without consent!

Indeed, the need to get consent from the author before reading a published work Isn't A Thing in general, outside some very specific contractural scenarios.

solid_fuel 2 hours ago | parent [-]

Why are you claiming that citations are the same thing as theft without citation? There’s a mile of difference between citing a work and taking it, rewording it, and not crediting or compensating the original author.

gruez 25 minutes ago | parent [-]

> There’s a mile of difference between citing a work and taking it, rewording it, and not crediting or compensating the original author.

So what's wikipedia doing vs what LLMs do? So far as I can tell the only difference is in citations, but:

1. LLMs can be made to cite, eg. if you use google search's AI mode it'll happily provide citations. I doubt that would placate the AI haters though.

2. Outside of academia no one really cares about citations. There's no legal requirement to cite, nor do I think all the people complaining about AI "stealing" other peoples' work are going to be magically placated by the addition of a few citations. Moreover it's unclear whether the concept of citations makes sense in many contexts. If you ask a human programmer how to write fizzbuzz, they'll likely blurt out a solution without providing citations, much like an AI would. Same for most questions people are asking AI about, eg. "gimme a cake recipe", do you really need a citation back to some 18th century cook book?

gruez 4 hours ago | parent | prev [-]

I don't see how "experts" are relevant under OP's framework unless they're did the primary research themselves. Otherwise they're just regurgitating someone else's research. The "oh there's humans involved so that gets pass" excuse doesn't work either, because humans were also involved in training the AI.

solid_fuel 4 hours ago | parent [-]

Please reread the comment you are replying to. I will wait.

—-

Now that you have reread the initial comment, do you think that “experts” was the important part? Or do you think maybe it was the compensation for their work that matters?

gruez 4 hours ago | parent [-]

>Please reread the comment you are replying to. I will wait.

If you're going to post thinly veiled implications that I didn't read your comment, you should be pretty damn sure that you make it look like you read my comment, which it doesn't seem like you did. The second of my comment said:

>The "oh there's humans involved so that gets pass" excuse doesn't work either, because humans were also involved in training the AI.

If you did read it, you sure did a poor job at rebutting it, leaving it unaddressed and preferring to waste words on writing snarky remarks instead.

solid_fuel 3 hours ago | parent [-]

> If you're going to post thinly veiled implications that I didn't read your comment

It was a statement, not an implication.

You still haven't replied to the point in the original comment about compensating the people who do this work, so I think it's quite obvious that you haven't read it.

gruez 3 hours ago | parent [-]

>You still haven't replied to the point in the original comment about compensating the people who do this work, so I think it's quite obvious that you haven't read it.

Issac newton discovers the theory of gravity. He advanced the sum of human knowledge, so fair enough, he should get compensated.

Alice rehashes that and puts it into her encyclopedia, allowing others to learn the theory of gravity.

Bob writes an algorithm for training a chatbot that can produce responses rehashing the theory of gravity, also allowing others to learn the theory of gravity.

Why should Alice be compensated but not Bob? Neither discovered the theory of gravity, so it's not like by funding Alice we're helping discover quantum physics or whatever. It's also not obvious that Alice's work is more valuable. A chatbot interface is often better at teaching someone than a rehashed overview. Of course, you can try to fix this by declaring that human work is valuable and an AI model isn't, by fiat, but that's just a cope and a far cry from the original principle of "trained on [...] human knowledge, so their outputs should belong to humanity"

None of this matters for applying the law, because the law just says only human created works are eligible for copyright protection, but that's not the argument OP was trying to invoke.

Finally none of this actually matters because OP just bites the bullet and says that secondary sources shouldn't be eligible for copyright, period.

yellowapple 3 hours ago | parent | prev [-]

Correct, and if it wasn't already obvious by now I believe that to be an unambiguously good thing.

gruez 3 hours ago | parent [-]

Like, now with the advent of AI, or even before? You might not have much love for encyclopedias, which were mostly replaced by wikipedia, but the "secondary sources don't get copyright protection" would also cover programming books, which roughly speaking are docs rewritten to a cohesive narrative.

yellowapple 3 hours ago | parent [-]

> Like, now with the advent of AI, or even before?

My disdain for intellectual property predates the existence of LLMs by at least a decade.

TheOtherHobbes 4 hours ago | parent | prev | next [-]

It's either/or. Either you reward creators and inventors to keep creating and inventing and keep poverty at bay, or you give everyone enough to avoid poverty whether they work or not and reward c+i in some other way.

What we have now is neither - owners are hugely over-rewarded for owning things and extracting passive value from everyone else, creators and inventors are kinda sorta rewarded sometimes if they're lucky and very much not if they're not. Just like other workers.

"The commons" is not a thing in this model, except in a few small niches.

yellowapple 3 hours ago | parent | next [-]

That's indeed one of many reasons why I'm staunchly pro-UBI.

wakawaka28 3 hours ago | parent | prev [-]

>What we have now is neither - owners are hugely over-rewarded for owning things and extracting passive value from everyone else, creators and inventors are kinda sorta rewarded sometimes if they're lucky and very much not if they're not. Just like other workers.

Creators and inventors are rewarded but obviously they cannot consume the whole pie. The people who invest in creative pursuits eat a lot of losses. People only seem to notice profitable successes, and forget that failures need to be paid for as well.

The same logic also applies to workers. The fact that your labor costs money is a guarantee, and it might not make money at all. We can think of a few examples where the work is directly delivered to consumers with zero marginal overhead, but most work DOES have overhead and liabilities, no matter how simple.

rileymat2 4 hours ago | parent | prev | next [-]

I don't really understand this, I was trained on all the knowledge I was capable of ingesting, but my outputs are largely mine unless there is too much similarity to a copyrighted work. I don't understand why AI would be different. When I do work for hire, my employer owns it, roughly equivalent to me paying Open AI/Anthropic for output.

ranger_danger 4 hours ago | parent [-]

It's not, barring any secondary issues like proof the training data was acquired illegally.

When there is a copyright dispute, it's always a subjective test on the judge/jury's part as to how similar it is in appearance, purpose, etc. if it's not deemed fair use.

9dev 4 hours ago | parent | prev | next [-]

I’m all for this, if we also set up a system to let society pay for artists. And I mean in full: cover for their living expenses; their rent; internet access; materials; everything.

Or how do you suppose art gets created for humanity as a whole to enjoy?

gruez 3 hours ago | parent | next [-]

>And I mean in full: cover for their living expenses; their rent; internet access; materials; everything.

Who decides which artists get subsidized? How would larger projects work? Seeing how often open source volunteer projects implode due to various community drama, this sort of system would basically preclude any sort of big production.

yellowapple 3 hours ago | parent [-]

> Who decides which artists get subsidized?

The easy answer is to just pay literally everyone. This is called a “universal basic income”, and has been repeatedly demonstrated to be a good idea for many reasons besides decoupling creative expression from the need to put food on the table.

yellowapple 3 hours ago | parent | prev [-]

That, too, is something for which I've been advocating for quite some time now.

kulahan 4 hours ago | parent | prev | next [-]

Really gotta force the starving artist motif or what's the point?

ozlikethewizard 4 hours ago | parent [-]

That we all stand on the shoulders of giants? There is no private innovation that isnt built on mountains on publicly shared knowledge and innovation, and usually public funds as well. The idea that private enterprise leads innovation is propaganda, i.e publicly funded academic study paved the way for LLMs, corps just commodotise ideas to make them viable under capitalism.

gruez 4 hours ago | parent [-]

>The idea that private enterprise leads innovation is propaganda, i.e publicly funded academic study paved the way for LLMs, corps just commodotise ideas to make them viable under capitalism.

So if you founded some wildly successful unicorn, I (or the government) can come over and say "nice startup, too bad it runs off of the internet (based off ARPANET), so your startup belongs to the the state now"?

yellowapple 3 hours ago | parent | next [-]

The government can already do that if it so chooses, given that the government is the singular reason why any corporation is able to exist as an independent legal entity in the first place (as opposed to a bunch of individual persons collaborating informally).

ozlikethewizard 4 hours ago | parent | prev [-]

Got a whole wagon of straw to sell you, you're going to need it

iwontberude 4 hours ago | parent | prev [-]

[dead]