
Using the model starts with inputting a safety policy (covering local laws, organizational rules, and socio-cultural norms), followed by the content to be classified. The model then produces a complete reasoning trace: it determines whether the content complies with the policy and explains its reasoning process—without OpenAI imposing any preset stance.