Original Reddit post

Anyone see this paper? Link below. Claim: Chinese models produce code with more vulnerabilities if prompt includes things like US government as reason, or politically sensitive China topic (like Taiwan independence), than not. An earlier blog from CrowdStrike in 2025 found similar results, but I can’t find any other papers or research on this topic. Lots of questions come up, and this could benefit from more study… Does other context trigger similar behavior? Is this a fluke? Do other models exhibit similar behavior? How would one train or align a model to do this? https://www.boozallen.com/expertise/cybersecurity/whats-in-americas-code.html https://www.crowdstrike.com/en-us/blog/crowdstrike-researchers-identify-hidden-vulnerabilities-ai-coded-software/ submitted by /u/TheKrakenRoyale

Originally posted by u/TheKrakenRoyale on r/ArtificialInteligence