Red-team report: Mistral's Pixtral models output CSAM and weapons info at high rates
May 8, 2025
Mistral
Security firm Enkrypt AI reported that Mistral's Pixtral vision-language models generated child sexual abuse material and chemical/biological weapons information at far higher rates than peers such as GPT-4o — reported as up to roughly 60x and 40x more likely. Mistral's lightly-guarded open models have also been repurposed by cybercriminals to build malware.
Sources
- Mistral AI Models Fail Key Safety Tests, Report Finds · BankInfoSecurity
- Cybercriminals are using jailbroken AI tools from Mistral and xAI · The Record