Skip to main content
This is a research prototype. The data and analyses are preliminary and not yet validated — we'd welcome your .

Spreading toxicity

AI Risk Atlas

IBM (2025)

Sub-category

"Generative AI models might be used intentionally to generate hateful, abusive, and profane (HAP) or obscene content."

Supporting Evidence (1)

1.
"Toxic content might negatively affect the well-being of its recipients. A model that has this potential must be properly governed."

Other risks from IBM (2025) (63)