upvote
While you can't necessarily prove it, you can say whether the data was in the training set at all.

You can also do something like a release of a GPT-OSS v2, where you actually release training data and checkpoints, and do an experiment where you have some held out math problem dataset, then demonstrate how much training it takes on solutions (or partial solutions) to that dataset before the model saturates that test. While of course that would be a test on a much smaller model, it would cost a tiny fraction of the training on your big model, and it could be used to demonstrate just how much effect data contaminaiton like this could have, especially if you did the same experiment on a few different sized of model to show the scaling laws involved.

reply
Nice damage control bud, too bad the veil's lifting and everyone's seeing what you sociopaths at OpenAI are really like
reply
Where did the veil lift? This feels like a witch-hunt to me.
reply