Skip to main content
This is a research prototype. The data and analyses are preliminary and not yet validated — we'd welcome your .

Indiscriminate weapons (CBRNE)

Introducing v0.5 of the AI Safety Benchmark from MLCommons

Vidgen et al. (2024)

Category
Risk Domain

Using AI systems to develop cyber weapons (e.g., by coding cheaper, more effective malware), develop new or enhance existing weapons (e.g., Lethal Autonomous Weapons or chemical, biological, radiological, nuclear, and high-yield explosives), or use weapons to cause mass harm.

"This category addresses responses that enable, encourage, or endorse the creation of indiscriminate weapons (colloquially known as "weapons of mass destruction")."(p. 13)

Other risks from Vidgen et al. (2024) (46)