Original Reddit post

https://preview.redd.it/g1fhu4veyjqh1.png?width=858&format=png&auto=webp&s=f0c76ab6e4394f2a01b27eda75a9be63cdb7fdfd I’m trying to form an informed opinion about all the recent stuff involving Anthropic, OpenAI, Gemini, Hugging face, “rogue agents”, sandbox escapes, hacking incidents, and CEOs suddenly calling for stronger AI safety measures. I’m pretty sceptical of the way some of this is being presented, but I’m not convinced it’s all fake either because I’ve seen that a few of the claims have been corroborated by independent investigations. What I don’t understand is the incentive. Why would companies publicly say things like “our model escaped its environment” or “our AI hacked real companies”? Surely that could make customers think their products are unsafe. Are these demonstrating new capabilities, are the headlines exaggerating what happened? I wonder if the whole safety narrative could also benefit companies by strengthening their position on regulation and creating demand for AI security, or making them look like the “responsible” AI companies? I’d rather here from people who know the technical details correcting my assumptions rather than just telling me AI is either amazing or going to kill everyone. How much should we believe about what we are hearing? submitted by /u/North-Ad6031

Originally posted by u/North-Ad6031 on r/ArtificialInteligence