Comment by mucha
14 hours ago
"If they opted out of training, then we definitely did not train on them."
Per OpenAI's privacy policy, they use de-identified data to improve their products. From Mark Chen's comment, improving products includes improving ChatGPT and Codex in a holistic way. Improving models in a holistic way sounds a lot like training to me.
> Per OpenAI's privacy policy, they use de-identified data to improve their products
That’s not inconsistent with what you responded to. They use your data unless you opt out. If the user doesn’t opt out, their de-identified data is used to improve their products.
It appears than you can only opt-out from having OpenAI train models on your data. There isn't an option for opting to exclude your de-identified data from being used to improve OpenAI products.
Are you certain of this? I would be inclined to believe you but it would be nice to know decisively.
> Improving models in a holistic way sounds a lot like training to me.
I think that's quite a leap. Using de-indetified data to improve the products is what everyone has been doing since the dawn of web analytics.