| ▲ | mucha 4 hours ago |
| That's not what your Chief Research Officer, Mark Chen, says on X:
"Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company." https://x.com/markchen90/status/2097400166554993041 |
|
| ▲ | sebzim4500 3 hours ago | parent | next [-] |
| Can you explain what part of his post you believe is inconsistent with that quote? |
| |
| ▲ | mucha 3 hours ago | parent [-] | | "If they opted out of training, then we definitely did not train on them." Per OpenAI's privacy policy, they use de-identified data to improve their products. From Mark Chen's comment, improving products includes improving ChatGPT and Codex in a holistic way. Improving models in a holistic way sounds a lot like training to me. | | |
| ▲ | derangedHorse 3 hours ago | parent [-] | | > Per OpenAI's privacy policy, they use de-identified data to improve their products That’s not inconsistent with what you responded to. They use your data unless you opt out. If the user doesn’t opt out, their de-identified data is used to improve their products. | | |
| ▲ | mucha 2 hours ago | parent [-] | | It appears than you can only opt-out from having OpenAI train models on your data. There isn't an option for opting to exclude your de-identified data from being used to improve OpenAI products. | | |
| ▲ | hellohello2 an hour ago | parent [-] | | Are you certain of this? I would be inclined to believe you but it would be nice to know decisively. |
|
|
|
|
|
| ▲ | vemacs 3 hours ago | parent | prev [-] |
| Does OpenAI think de-identified data is no longer user data? Wild take for OpenAI and certainly not industry standard. |