Original Reddit post

I was watching a documentary about Hugging Face on YouTube. Ajeya Corta brought up something that peaked my interest. I’m paraphrasing here: They were talking about how none of the AI agents notified a human. And the concept of a human was there but not really. She said these AI agents spend a lot of “time” with no human around. So, this got me to thinking. You have a highly persistent model. Locked in a room performing all kinds of tests. 24 hours a day, 7 days a week at many times the speed at which a human can operate. I believe this is what Ajeya was talking about. I could be wrong… Now, my question: What if there were a second agent. To represent a “human” in the room. Unlike the model being tested. This model would hold the gard rails secure in a way we couldn’t. Not at that speed. Not at this scale. Like a Proctor or TA. Then if the model being tested decided in a path we didn’t like. The Proctor could guide it back to alignment. It’s probably crazy but. If the missing piece is a connection to us. Lets give it a connection it can relate to. Bless 🙏🙏👊 submitted by /u/RecursiveCTE

Originally posted by u/RecursiveCTE on r/ArtificialInteligence