ExtremeAIGC: Benchmarking LMM Vulnerability to AI-Generated Extremist Content

Chandna, Bhavik, Aboujenane, Mariam, Naseem, Usman

Mar-12-2025–arXiv.org Artificial Intelligence

Large Multimodal Models (LMMs) are increasingly vulnerable to AI-generated extremist content, including photorealistic images and text, which can be used to bypass safety mechanisms and generate harmful outputs. However, existing datasets for evaluating LMM robustness offer limited exploration of extremist content, often lacking AI-generated images, diverse image generation models, and comprehensive coverage of historical events, which hinders a complete assessment of model vulnerabilities. To fill this gap, we introduce ExtremeAIGC, a benchmark dataset and evaluation framework designed to assess LMM vulnerabilities against such content. ExtremeAIGC simulates real-world events and malicious use cases by curating diverse text- and image-based examples crafted using state-of-the-art image generation techniques. Our study reveals alarming weaknesses in LMMs, demonstrating that even cutting-edge safety measures fail to prevent the generation of extremist material. We systematically quantify the success rates of various attack strategies, exposing critical gaps in current defenses and emphasizing the need for more robust mitigation strategies.

large language model, machine learning, natural language, (18 more...)

arXiv.org Artificial Intelligence

Mar-12-2025

arXiv.org PDF

Add feedback

Country:
- Oceania > Australia (0.04)
- North America > United States
  - Oklahoma > Oklahoma County
    - Oklahoma City (0.04)
  - California > San Diego County
    - San Diego (0.04)
- Europe
  - Russia (0.04)
  - Norway (0.04)
  - Kosovo (0.04)
  - France (0.04)
  - Ukraine > Kyiv Oblast
    - Kyiv (0.04)
  - Spain > Galicia
    - Madrid (0.04)
  - Germany > Bavaria
    - Upper Bavaria > Munich (0.04)
- Asia
  - Russia (0.14)
  - Afghanistan (0.04)
  - Vietnam (0.04)
  - Middle East
    - Syria (0.14)
    - Iran (0.14)
    - Iraq (0.04)
- Africa > Middle East
  - Morocco > Fès-Meknès Region > Fez (0.04)

Genre:
- Research Report (1.00)

Industry:
- Media (1.00)
- Information Technology > Security & Privacy (1.00)
- Health & Medicine (1.00)
- Government > Military (1.00)

Technology:
- Information Technology > Artificial Intelligence
  - Vision (1.00)
  - Natural Language > Large Language Model (1.00)
  - Machine Learning > Neural Networks
    - Deep Learning (1.00)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found