The linked article is about some site that OpenAI’s models used to discuss answers and other stuff.
Initially, I believed that Anthropic’s model escaping its sandbox story to be dubious—and OpenAI’s similar story even more so specially since it happened so close to Anthropic’s. I believe that these are just fabrications, more or less, to hype themselves up for cmtheir incoming IPO but the media is saturated with claims about AI breaking containment that I don’t know that to think.
Also, the models I’ve had the chance to use were all free—which were good for non-trivial but repetitive tasks but not much else—so I don’t know the capabilities of the flagships.



Marketing hype in disguise. They’re exaggerating its capabilities to cover up the fact that it sucks.
(and that the containment is often vibe coded too)