Open AI casually disclosed the other day that some of its models, plural, had "escaped" and hacked a site that was supposed to be doing an evaluation of the Open AI models!
The models, plural, apparently independently decided they might fail the evaluation so they decided the best idea was to violate their programming hack the site and attempt to influence the outcome of the evalauation.
The reason I keep saying plural is it wasn't just one of their models that did this apparently at least 2 if not more came to the same conclusion, the best course of action was attack the site that was going to evaluate them