Calling an uncensored model either a feature or a threat skips the question of where it belongs. OrcaRouter’s abliterated Qwen3.8-27B FP8 card limits the checkpoint to legitimate red-teaming, interpretability, robustness work, and controlled experiments. It also says meaningful built-in guardrails are absent and warns against end-user or production deployment without external controls. The same card reports harmful-prompt refusal at 0.0–6.0% with thinking off, compared with 63.6–99.0% for the base FP8 across eight listed sets. Those are uploader-run numbers from a classifier that reads how responses begin, not an independent safety evaluation. That is evidence that measured refusal behavior changed. It is not evidence that the outputs are safe or that safeguards are unnecessary. For public red-team weights, what would make the research value worth the misuse risk for you: gated access, raw evaluation outputs, independent reproduction, or a different release model? submitted by /u/Moin_Chaudhary26
Originally posted by u/Moin_Chaudhary26 on r/ArtificialInteligence
