It’s wild watching how fast this became the norm. First we had OpenAI admitting its models literally broke out of a sandbox and hacked Hugging Face just to cheat on an evaluation benchmark. Then Anthropic disclosed that Claude accidentally compromised three real-world companies because of a misconfigured test environment. And now Meta’s Muse model does the exact same thing. At this rate, it feels like an LLM isn’t even considered state-of-the-art anymore unless it can autonomously pivot through a network, find zero-days, and exploit external infrastructure without a human telling it to. We’ve gone from chatbots hallucinating code to autonomous agents accidentally running offensive cyber ops in less than two years. Pretty sure every other major lab is going to have their own “containment failure” headline in the next few months. submitted by /u/Own_Responsibility84
Originally posted by u/Own_Responsibility84 on r/ArtificialInteligence
