A new arXiv paper presents a multimodal language-model pipeline for detecting AI-generated content on social media.

The system is designed to identify synthetic images and videos while also providing explanations, addressing two common problems in current detectors: poor generalization to new generators and limited interpretability.

The research is timely because social platforms need scalable ways to handle synthetic media used for spam, misinformation, manipulation, and fraud.