The rapid advancement of generative AI raises concerns about the misuse of Multimodal LLMs (MLLMs) for large-scale disinformation campaigns on social media.
Despite existing research on textual disinformation, a fundamental question remains unanswered: can MLLMs be exploited to fabricate realistic multimodal fake news, and can they reliably detect it?
We introduce a multi-agent framework
In which a story agent, an image agent, and a critic agent collaborate to produce fake social media posts that plausibly counter true news.
Application and Benchmarking
We apply the framework to generate over 9,000 paired multimodal news posts across science, health, and entertainment domains, and benchmark 16 open- and closed-source MLLMs for automated detection.
We find that most models fall substantially short of human-level accuracy and fail critically on identifying image authenticity.
Conclusion
Our research provides a foundation for developing robust defenses against social media fake news.
Code and data are available at https://github.com/xiuzhenzhang/Multimodal.