I’ve read up on every incident that has taken place with agents escaping sandboxes and the thing I continue to not understand is how a cheap transcript watcher couldn’t have surfaced the incidents to a human and avoided the majority of the mess. OpenAI has Luna, intelligence that’s aiming to be too cheap to meter. Why is that not used to observe training runs? Luna agent reads transcript “OH MY GOD A MESSAGE BOARD” -> surfaces issue to human researcher… why wasn’t it that simple? submitted by /u/rhaivn
Originally posted by u/rhaivn on r/ArtificialInteligence
You must log in or # to comment.
