Original Reddit post

Whenever I read discussions about AI risk on Reddit, the conversation often seems to split into echo chambers at opposite ends of the spectrum: from “we’re all doomed” to “it’s all hype”. So I wanted to look at the question properly, especially because I see the same split in my own workplace (I am an academic). I went through a large number of sources and put together a 30-minute video essay trying to separate two questions that are often conflated: What capabilities have actually been demonstrated? And what would still need to happen before genuine loss of human control became plausible? After a brief introduction to bring non-experts up to speed with current AI capabilities, I looked at capability evaluations, the OpenAI/METR multi-agent cyber incident, autonomous-replication tests, researcher surveys, AI safety reports, and evidence on AI-assisted AI research. I also connect the Hugging Face attack with the Navier-Stokes result, because I think they reveal the same underlying capability. My conclusion is deliberately neither “doom” nor “nothing to worry about”. Some capabilities that used to be largely hypothetical now exist in limited form, but several steps between current systems and an AI that humanity could no longer control remain speculative. I’d be interested to hear where people here think the evidence becomes convincing or stops being convincing. Feedback on what is missing from the current AI information space is also very welcome, as I’d like to explore some of these topics in more depth in future discussions with AI colleagues. Video: https://www.youtube.com/watch?v=Ha3nrhxnuWI submitted by /u/Ok-Professor7130

Originally posted by u/Ok-Professor7130 on r/ArtificialInteligence