|
| ▲ | hellohello2 3 hours ago | parent | next [-] |
| But there is evidence, the blog post says: "While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models ." In other words, yes, they had been using ChatGPT, and yes, ChatGPT could very well have trained on their data. Now that there is evidence, we need an investigation: yes or no, was it the case? |
| |
| ▲ | fc417fc802 2 hours ago | parent [-] | | That is not an admission of malfeasance though? As I read it they don't know if anyone fed relevant private documents into the model under an account configured to permit training on user data. If there's more to the story I'd be interested to hear it. | | |
| ▲ | hellohello2 an hour ago | parent [-] | | Of malfeasance no, but they could have easily plagiarized unintentionally. If you commit mansalughter, you still need to explain yourself, even if it was a complete unlucky accident. |
|
|
|
| ▲ | Arodex 5 hours ago | parent | prev | next [-] |
| But the evidence is in the hand of the potential culprit. That's why allegations can be enough to force confiscation and intrusion to get evidence in safe hands before it is destroyed by the accused party. |
| |
| ▲ | fc417fc802 4 hours ago | parent [-] | | Only in the event that there is some reason to suspect them of wrongdoing. Which would generally require evidence. You don't just get to subpoena your neighbor's bank account because "I know he's stealing from me" you need to first present credible evidence that you were stolen from and that he is among the most likely culprits. | | |
| ▲ | Arodex 4 hours ago | parent [-] | | But I can subpoena my neighbours bank account when I see him driving a brand new 500'000$ car and I have a 490'000$ hole in my bank account and he works in the bank where my money is. And when questioned he evades some questions and threatens to destroy my career. Any other argument, fc417fc802? | | |
|
|
|
| ▲ | nulld3v 5 hours ago | parent | prev [-] |
| Nobody except OpenAI knows whether or not OpenAI trained on their data. So the burden remains on OpenAI here. |
| |
| ▲ | tristanj 3 hours ago | parent | next [-] | | Incorrect, Buckmaster and Alpöge can comment if they had the ChatGPT "Improve the model for everyone" setting enabled or disabled. If it was enabled, then their work was included in the training dataset. | | |
| ▲ | opello 2 hours ago | parent [-] | | In order for this to be the strong evidence everyone also has to believe that the setting is absolutely true. That some logging from some piece of the system could not also leak the prompt information in such a way that it could have been included as training data. Perhaps the design of how data is collected for the training dataset is so rigorous as to make this a practical impossibility. But, it's asking a lot without sufficient detail to completely exclude from possibility that one setting is all that could possibly have been absolutely load bearing in deciding if the other researcher's active efforts meaningfully contaminated the internal model. At least, as an ignorant outsider, that's how it seems to me. |
| |
| ▲ | fc417fc802 5 hours ago | parent | prev [-] | | That is an absurd and entirely untenable position that breaks with approximately all western conventions. Only the CIA knows whether or not they're actively covering up reptilian space aliens exerting control over the US government. Therefore the burden of proof remains on the CIA to prove that they are not actively participating in such a scheme. | | |
| ▲ | nulld3v 5 hours ago | parent [-] | | I don't understand, OpenAI can just say: "yes/no we did/did not train on your data". It's not a hard question to answer, and it is a question that OpenAI should be able to answer for all data we feed into ChatGPT. | | |
| ▲ | fc417fc802 4 hours ago | parent [-] | | > It's not a hard question to answer I didn't realize you had insider knowledge about their systems. Do please explain for the class. As I understand it they will only have trained on his data if he consented to it. Do you have evidence that they do otherwise? | | |
| ▲ | Timon3 4 hours ago | parent | next [-] | | This whole discussion is about evidence. That's not proof and it is not certain, but it is evidence pointing into the direction that OpenAI might be doing something that they're strongly incentivized to do. What kind of "evidence" do you see as necessary? | |
| ▲ | hellohello2 3 hours ago | parent | prev [-] | | When someone authors a paper, is it on others to proove the author did not use their work as inspiration? No, it is on the author to give credit where it is due. You guys are acting as if it its legal issue, when it is not. |
|
|
|
|