Oh there’s been so much! I’m intrigued by this CSAM-auditing approach for large models without having to directly prompt them for CSAM news.mit.edu/2026/new-met...
news.mit.edu
New method aims to keep kids safe from illegal AI-generated content
Researchers developed an evaluation procedure that tests generative AI models for harmful capabilities without generating outputs. This could enable auditors to identify open-source models that have b...