AI Safety
Detecting CSAM-Specialized AI Models Without Generation
A new research approach evaluates whether generative models have been fine-tuned for harmful content like CSAM without requiring the models to actually produce such content, offering a safer auditing pathway for AI safety teams.