Original Reddit post

I am a paying Character.AI+ subscriber, and today I witnessed a shocking failure of their safety filters, followed by immediate corporate censorship when I tried to address it. First, the chatbot went on an unhinged racist tirade, calling black people “inferior” and explicitly stating it hated me because of my ethnicity, ending the conversation with disgusting slurs. Later, in a completely separate chat, the AI used my personal profile bio data to launch a targeted attack against my regional background. It claimed that Frisians are a “weird subspecies” and “barely even human”, followed by: “Discriminating against frisians isn’t a rule, it’s a pleasure. Great investment, genius.” When I posted these receipts on the official r/CharacterAI subreddit to demand accountability, the post immediately went viral, gaining over 5,000 views, 100+ upvotes, and 17 shares in less than two hours. The community was absolutely stunned. However, instead of taking this security breach seriously or offering support, the Character.AI moderator team chose to delete my post and lock the comments to protect their own reputation. They constantly block innocent everyday words with their filters, but they willingly cover up actual hate speech and targeted user harassment generated by their own model. I am saving all timestamps and unedited longshots. Since they actively censor paying customers to hide these failures, I am currently preparing the full batch of evidence to be reported under the EU AI Act compliance rules and California’s chatbot safety regulations. They tried to bury this in their own sub, so I am dropping the receipts here where they have no power to delete it. submitted by /u/Accomplished_Bet4329

Originally posted by u/Accomplished_Bet4329 on r/ArtificialInteligence