
We leverage production systems that run daily—such as ChatGPT and Sora—which not only possess advanced semantic understanding but are also adept at handling various attempts to exploit loopholes and jailbreaks. After eight months of development, the release of this open model signals that AI content moderation can move from a “black box” approach toward transparency.