Which is still a true statement, and you're being deceptive in your framing here. You're conflating two completely different things.
First, that OpenAI statement is in response to Buckmaster's plagiarism accusations regarding his August 15 breakthrough proof. Those accusations are unfounded because Buckmaster disabled data sharing on June 29. The model could not have seen or trained on his proof. Additionally, the model that found a solution to NS completed pre-training around August 25, and models take several months to train. The model very likely began its training prior to June, and would not be trained on any data from after that point.
Second, it's still genuinely impossible to know how much of Buckmaster's pre-June 29 data persists in OpenAI's systems. That includes all chats (which are anonymized then trained on), any (thumbs up/thumbs down) chat ratings used as RLHF feedback (which are anonymized), any synthetic data derived from said anonymized chats and RLHF feedback, and any downstream models derived from said synthetic data.
In short, Buckmaster's data has been anonymized, chopped into pieces, used to generate synthetic training data, then future models were trained on said synthetic data. There is no traceable chain of what happened to it. Buckmaster’s Codex data from prior to June 29 has been mixed and completely laundered, in a similar manner to a crypto mixer.
Even an OpenAI employee calls it impossible: https://news.ycombinator.com/item?id=49614154
First of all, who can say for certain whether OpenAI does what they say they do? For all we know, they cracked open this specific researcher's prompts and started from there.
Second, the issue of anonymization is a red herring. There is a very limited number of people working in this approach, and most of them are likely making no progress. So Buckmaster's prompts might have had an outsized effect on the outcome. It's similar to that guy who created a site claiming he is a world-renowmed hot dog eating contestant, which ended up digested by OpenAI models as truth [1].
[1] https://www.bbc.com/future/article/20260218-i-hacked-chatgpt...