| ▲ | Skyy93 an hour ago |
| This is a lobby organisation using only the pieces and bits they like to push their own agenda. "sketchy russian website", how about using some more clear description like: A library for sharing books and articles that should be partly public domain because they were paid for by the public. Only some of the material is copyrighted by authors. However, some of their work is so old that it is not reprinted anyway. But of course such an explanation would not click. I also don't see a problem with statements about making people jobless. Imagine if every robotic or automation company advertised like this: Yeah, you'll buy tons of expensive robots and still rely on expensive labor from real people without any efficiency gains. |
|
| ▲ | quaintdev an hour ago | parent | next [-] |
| It's just sad that this is the top comment on Hacker News. Why are we giving free pass to these tech companies? Why are we trusting these CEOs when they have repeatedly broken laws? Remember Aaron Swartz and the fate he suffered? Why is big tech getting away with so much more? |
| |
| ▲ | TeMPOraL 40 minutes ago | parent | next [-] | | > Remember Aaron Swartz and the fate he suffered? Why is big tech getting away with so much more? Why are you turning him into perpetuum mobile in his grave? Do you really believe Aaron would be arguing against AI companies and for publishing / recording guilds on the grounds of intellectual property claims? No, it's the tech community that did a sudden about-face, and is now all "friendship ended with free access to information and technologies enabling people; now RIAA is my best friend", and this move is as dumb as that meme (https://imgflip.com/memegenerator/137501417/Friendship-ended). | | |
| ▲ | applfanboysbgon 8 minutes ago | parent | next [-] | | Suppose you have three propositions: A: "Information is free" B: "Information is not free" C: "Information is free only for the rich and not free for everyone else, giving the rich a material advantage over everyone else that not only entrenches but accelerates wealth inequality and impedes class mobility" You, or Swartz, are an advocate for A. Why, exactly, do you think that obliges you/Swartz to prefer C over B while A is not true? | | |
| ▲ | simianwords 6 minutes ago | parent [-] | | Sure it’s not free for anyone and both companies and individuals are treated similarly. It’s not like you will be jailed for pirating movies. And neither should OpenAI. What part of this is hard to understand | | |
| ▲ | applfanboysbgon 5 minutes ago | parent [-] | | > It’s not like you will be jailed for pirating movies. We're literally talking in a thread about someone who committed suicide because the US government was hellbent on ruining his life with a felony conviction for piracy. |
|
| |
| ▲ | quaintdev 28 minutes ago | parent | prev [-] | | > Do you really believe Aaron would be arguing against AI companies and for publishing / recording guilds on the grounds of intellectual property claims? That's not the point. All the rules and laws are enforced when its you and me but when it's big tech the laws are treated by these companies as mere instructions. > friendship ended with free access to information and technologies enabling people; now RIAA is my best friend Big tech will enable access to free information and will help people reach new heights. Do you see how wrong that sounds? |
| |
| ▲ | simianwords 9 minutes ago | parent | prev | next [-] | | Aaron was sued for distributing and not for pirating. They are different offences | |
| ▲ | MattGaiser 34 minutes ago | parent | prev | next [-] | | We tend to accept lawbreaking if it is for a product we want. There are several multi billion dollar companies where the founding thesis was “what if we just ignore the law?” | |
| ▲ | Skyy93 an hour ago | parent | prev [-] | | Then you misunderstood my point. I think copyright law should be significantly changed and the current system hurts us all. | | |
| ▲ | ares623 43 minutes ago | parent | next [-] | | Sure. But the fact is they broke current existing laws, with known punishments with precedents. Same as a new law doesn't retroactively punish someone, then a new law shouldn't absolve someone before it's passed. | | |
| ▲ | ben_w 16 minutes ago | parent | next [-] | | They did. They were found guilty. The case I looked at* was a civil case so this was settled out of court before the court imposed a settlement. The law they broke was pirating the materials, not training per se, even though training is what so many people object to: the judge ruled that actually training a model, when the materials you used were ones you otherwise had lawful access to, was not a breach of law. IMO, the laws need to change to reflect what tech can now do. This wouldn't be the first time, copyright law has had to shift several times before as new means of reproduction are created. * the Anthropic one | |
| ▲ | Skyy93 40 minutes ago | parent | prev [-] | | >A copyright is a type of intellectual property that gives its owner the exclusive legal right to copy, distribute, adapt, display, and perform a creative work, usually for a limited time. The more interesting question is IMO if AI training actually falls into one of these cases. You can read a book and also copy it, but you do not do because of the law. However, you have the ability to do so. Is having the ability to do something already forbidden? |
| |
| ▲ | jrflowers 9 minutes ago | parent | prev | next [-] | | > Then you misunderstood my point. No, they were responding to the post defending OpenAI that you wrote. If you meant to communicate something other than “criticism of OpenAI in this context is unwarranted” then it looks like you forgot to do that and wrote something else instead | |
| ▲ | cisc 17 minutes ago | parent | prev [-] | | > the current system hurts us all How? The current system enables the GPL. The GPL protects many open source projects. |
|
|
|
| ▲ | probably_wrong 10 minutes ago | parent | prev | next [-] |
| > I also don't see a problem with statements about making people jobless I do see a problem with a company loudly announcing that they are going to make people's lives miserable purely for profit. Leaving aside that it goes against OpenAI's stated mission ("to ensure that artificial general intelligence benefits all of humanity"), the disdain for the lives they are intentionally trying to ruin makes it a problem. And even if you believe that the transition is inevitable, as it is the case with phasing out combustion engines in cars, anyone reasonable would see that the transition is gradual to give people time to adapt. Instead of doing that, OpenAI is burning cash at astonishing rates, polluting the environment, and killing personal computing with the only aim of being the only ones left atop the ruins. I do see a problem with that. |
|
| ▲ | 4ndrewl an hour ago | parent | prev | next [-] |
| "This is a lobby organisation using only the pieces and bits they like to push their own agenda." Do you have any evidence of them being a lobby organisation (as opposed to OpenAI for example which spends millions of dollars hiring actual lobbyists) |
| |
| ▲ | Skyy93 an hour ago | parent | next [-] | | Simply look at the Wikipedia article? >The group lobbies at the national and state levels on censorship and tax concerns, and it has initiated or supported several major lawsuits in defense of authors' copyrights. | |
| ▲ | TeMPOraL 32 minutes ago | parent | prev [-] | | > Do you have any evidence of them being a lobby organisation Have you looked at their name? |
|
|
| ▲ | trompetenaccoun 41 minutes ago | parent | prev | next [-] |
| It can be a bit confusing due to the terrible style of the article (ironic given the source) but it seems the "sketchy russian website" part is a direct quote by Anthropic's Sam McCandlish. And apparently Dario Amodei referred to it as sketchy as well. I find the brazenness of saying this while running what's arguably the largest copyright theft operation in human history astonishing. If libgen is "sketchy", then what is OpenAI? |
| |
| ▲ | qarl 38 minutes ago | parent | next [-] | | > the largest copyright theft operation in human history Many people think that it was fair use: training is akin to reading, not copying. Especially the courts. | | |
| ▲ | Trusteando 11 minutes ago | parent | next [-] | | No only reading, because the content, style, selection of topics, and more is encoded, written, in the LLM weights, and they obtain profit from them. Noone can compete which copying and pasting (in encoding from) from copyright protected material. | | |
| ▲ | qarl 7 minutes ago | parent [-] | | > No only reading, because the content, style, selection of topics, and more is encoded, written, in the LLM weights Exactly analogous to a human reading the material. |
| |
| ▲ | trompetenaccoun 28 minutes ago | parent | prev [-] | | That's not been legally established, the litigation is ongoing. And if mere downloading and reading of copyrighted material were legal, how come torrent users have been fined for it in the thousands? The law is the law, there can't be different law for corporations with billions in backing. I don't agree with current copyright laws btw and think they should be changed. However, they probably should have lobbied for that before illegally downloading all this material. | | |
| ▲ | qarl 22 minutes ago | parent [-] | | > That's not been legally established 100% of the rulings agree with me. The piracy is not in question. It is unarguably copyright violation. But that's not what anyone means in this context. Training is what everyone means. > The law is the law, there can't be different law for corporations with billions in backing. I didn't say otherwise. That's a straw man. | | |
| ▲ | trompetenaccoun 17 minutes ago | parent [-] | | So we agree they have violated copyright at a much larger scale than LibGen, yet they call LibGen "sketchy" for doing the same thing? Absurd, what exactly are we arguing here? | | |
| ▲ | qarl 12 minutes ago | parent [-] | | > So we agree they have violated copyright at a much larger scale than LibGen No. |
|
|
|
| |
| ▲ | TeMPOraL 35 minutes ago | parent | prev [-] | | They didn't believe it was sketchy. They were just worried that the commentariat on HN will frame it in a dumb, manipulative way like that. Judging by how AI threads look like for the past year, they were absolutely right to be worried. > largest copyright theft operation in human history In fact, you're doing exactly that right here. | | |
| ▲ | probably_wrong 19 minutes ago | parent | next [-] | | The comment you're replying to is citing almost verbatim [1] Microsoft’s director of Applied Science, Brent Hecht, who called OpenAI's data collection practices "the largest theft of labor in human history" in an internal memo. [1] https://techcrunch.com/2026/09/17/microsoft-exec-called-ai-s... | | | |
| ▲ | discreteevent 22 minutes ago | parent | prev | next [-] | | > frame it in a dumb, manipulative way like that
> In fact, you're doing exactly that right here. You're dead fucking right they are doing exactly that. They are saying that what is wrong is wrong. Instead you seem to be making out that AI companies are some kind of victim that has to "worry" about "manipulation". Meanwhile authors are out of a job right now and not by accident. What's up with that? | |
| ▲ | latexr 19 minutes ago | parent | prev [-] | | > They were just worried that the commentariat on HN will frame it in a dumb, manipulative way like that. You’re chastising others for a tone you are yourself employing, and are making monumental assumptions based on a few choice quotes. From the quotes alone you can’t tell if OpenAI thought libgen was sketchy or not. Also, contrary to what you’re claiming, they were wrong. HN in general seems to approve on libgen when used for its purpose of downloading some books on an individual level. The complaint you’re replying to is about what OpenAI did with the data, it has nothing to do with the website they got it from. |
|
|
|
| ▲ | abroszka33 an hour ago | parent | prev | next [-] |
| Any source that this was a library? Even then that would still raise a question if OpenAI is a Russian organisation or not to access that library with good faith. I think they just used a Russian torrent site. |
| |
|
| ▲ | lensecat an hour ago | parent | prev | next [-] |
| "A lobby organisation?" Of course a single author would not be able to afford facing a multi billion dollar company on their own? And the "sketchy russian website" quote is from OpenAI employees themselves? What are you on about? |
| |
| ▲ | Skyy93 an hour ago | parent [-] | | First, it's still a lobbying organization, so it's their job to make exaggerated claims, like a union in a company or any other organization with a political purpose. My first point was to highlight that it's not neutral or news related. It's fine that they have their opinion, but it's also my right to say that they're biased. The second thing underscores my point. They use one line and think they've made a great point because one employee called LibGen sketchy. This site has been around since the 2010s, and it has helped many people do research. It's not just a sketchy website that suddenly appeared and is always doing bad things. I think a more nuanced stance is necessary. | | |
| ▲ | latexr 37 minutes ago | parent [-] | | > They use one line and think they've made a great point because one employee called LibGen sketchy. No, you are using one line from the post to discredit them. The release has more than that and it’s not the only communication they made on this matter nor is there any indication it will be the last, it’s just the current one. | | |
| ▲ | Skyy93 33 minutes ago | parent [-] | | They are having it in the subtitle. In general their whole article is about two main points, first the use of stuff from Libgen, second the points of making people jobless. Three of their points are about the jobless thing two about the LibGen. About LibGen, there might be more discussion - fair. However, the second argument is no real discussion IMO. Why is putting people out of work suddenly a bad thing? Since when do we argue this when talking about automation? | | |
| ▲ | TeMPOraL 29 minutes ago | parent [-] | | There is nothing to discuss about LibGen, really. I don't read this as OpenAI employees even believing LibGen is sketchy. They were worried about optics, because LibGen itself is Russian and does look a bit sketchy, and at the time - much like today - it was easy to make it a headline that makes people pattern-match to "troll farms". (And then Russia invaded Ukraine, turning any association with .ru things into potential corporate suicide.) Really has nothing to do with LibGen or with OpenAI. It's about people being easy to manipulate into believing bullshit, which is a reasonable worry, and the Authors Guild is trying to do that exact thing OpenAI was worried about. | | |
| ▲ | Skyy93 24 minutes ago | parent [-] | | I agree with you, but I acknowledge that some people (not myself) might feel different about copyright. | | |
| ▲ | TeMPOraL 22 minutes ago | parent [-] | | I acknowledge that, but this part isn't even about copyright! The whole "sketchy russian website" bit resolves entirely about being seen as associated or supporting troll farms and Putin. EDIT: look at it this way: no one is calling Internet Archive "a sketchy US website". |
|
|
|
|
|
|
|
| ▲ | latexr 29 minutes ago | parent | prev | next [-] |
| “Sketchy Russian website” is part of a quote by Sam McCandlish (who worked at OpenAI), not the Author’s Guild characterisation. Also, defending it on the basis that some books on libgen are public domain is a poor excuse, like claiming people use The Pirate Bay to download Linux ISOs. Even if some of that is true, we all know that use case is not the popular one. |
|
| ▲ | an hour ago | parent | prev [-] |
| [deleted] |