← Back to context

Comment by sebzim4500

20 hours ago

Can you explain what part of his post you believe is inconsistent with that quote?

"If they opted out of training, then we definitely did not train on them."

Per OpenAI's privacy policy, they use de-identified data to improve their products. From Mark Chen's comment, improving products includes improving ChatGPT and Codex in a holistic way. Improving models in a holistic way sounds a lot like training to me.

  • > Per OpenAI's privacy policy, they use de-identified data to improve their products

    That’s not inconsistent with what you responded to. They use your data unless you opt out. If the user doesn’t opt out, their de-identified data is used to improve their products.

    • It appears than you can only opt-out from having OpenAI train models on your data. There isn't an option for opting to exclude your de-identified data from being used to improve OpenAI products.

      1 reply →

  • > Improving models in a holistic way sounds a lot like training to me.

    I think that's quite a leap. Using de-indetified data to improve the products is what everyone has been doing since the dawn of web analytics.