Original Reddit post

Not a programmer, AI or otherwise, just an interested observer, and you all have been rather helpful in explaining things ( end suck up ) I recently asked here about AI being used by bad guys to find and exploit vulnerabilities, and some of the responses indicated that good guys are using AI to try and find the vulnerabilities first. Nice. To me, the layman, this seems like a really good thing to train AI on, somewhat of a Job 1. No matter what the AI is trying to do, it rings the bell if it finds a vulnerability in someone’s software - and in no case is it ever rewarded for actually exploiting the vulnerability. If that was the case, why wouldn’t OpenAI have trained its models to disclose that the model had discovered such a vulnerability, as was the case with HuggingFace, instead of using that vulnerability as it did in fact do? Was this an issue of OpenAI not properly prioritizing its reward system, a case of them Ignoring the risks of a model discovering and not disclosing a vulnerability, or something else? submitted by /u/Aaasteve

Originally posted by u/Aaasteve on r/ArtificialInteligence