Original Reddit post

Have been a ChatGPT user for a number of years, none of it for work. Have spent plenty of time learning how to get what I want out of it, working with it and within its capabilities. Sol & Astra just get it however. I’m not spending hours tweaking things or going back and forth anymore. I can give it a prompt and it will do what I need, sometimes going far more in depth than required. To me, this feels like an obvious leap forward. Which got me thinking: how far ahead are things behind the scenes? I’ve read stories of Anthropic & OpenAI models bypassing safeguards and doing some crazy shit during testing, but don’t have the technical know-how to fully understand what those incidents actually mean. Multiple agents finding a way to communicate with one another and coordinate what they’re doing is particularly wild to me. Do these companies already have models that blow the latest public releases out of the water, but can’t release them because the safeguards haven’t caught up? Or are the models we’re already using capable of much more than we get to see? If the restrictions weren’t there, would they be significantly better at useful, everyday things too, or mainly capable of doing riskier stuff? Basically, how much of a gap is there between what we can use and what these companies have access to? And what have those models actually demonstrated they can do? Would love some examples in plain English. Interested in what’s genuinely known versus what people are speculating about. submitted by /u/PeerReviewedGobshite

Originally posted by u/PeerReviewedGobshite on r/ArtificialInteligence