Two things to distinguish:
Did any human or agent look at user data as part of the Navier Stokes effort? No.
Do we use user feedback and de-identified data to improve ChatGPT and Codex in a holistic way? Yes. And so does every LLM company.
“we cannot rule out that de-identified data derived from their usage of our products helped improve our models.” i mean props to them for straight coming clean....