SIRF: A Spec-Internalized Risk Foundation Model for Industrial Content Risk Control
This is a specific tactic for content moderation: bake policy into model weights at init time rather than controlling at inference. It works and is low-latency. The applicability depends entirely on whether your policy is stable and whether you have the annotation infrastructure to ground it. Mainly useful for platforms with mature policy infrastructure.