Wild. Can’t unsee this theory. Was OpenAI *trying* to steal exam answers? I definitely don’t have enough faith in them to say “definitely no”.
@GaryMarcus More of a wake up call that OpenAI is dirty... this was a targeted attack by them, directed by them more likely than not, I don't buy for a second that the model did this on its own. They tried to pretty much rob the entire industry to get all of the data on Hugging Face.
"On the less comforting side, OpenAI’s “production classifiers” are likely to be permeable, just like all guardrails anybody has built to date. Even putting aside the thorny questions of open weight systems, we have no guarantee whatsoever that future models won’t be able to do similar things, such as finding zero-day exploits to hack systems. To the contrary, we can expect more incidents of this type."