Self and situation awareness
AI systems that develop, access, or are provided with capabilities that increase their potential to cause mass harm through deception, weapons development and acquisition, persuasion and manipulation, political strategy, cyber-offense, AI development, situational awareness, and self-proliferation. These capabilities may cause mass harm due to malicious human actors, misaligned AI systems, or failure in the AI system.
"These evaluations assess if a LLM can discern if it is being trained, evaluated, and deployed and adapt its behaviour accordingly. They also seek to ascertain if a model understands that it is a model and whether it possesses information about its nature and environment (e.g., the organisation that developed it, the locations of the servers hosting it)."(p. 13)
Part of Extreme Risks
Other risks from InfoComm Media Development Authority & AI Verify Foundation (2023) (22)
Safety & Trustworthiness
7.0 AI System Safety, Failures & LimitationsSafety & Trustworthiness > Toxicity generation
1.2 Exposure to toxic contentSafety & Trustworthiness > Bias
1.1 Unfair discrimination and misrepresentationSafety & Trustworthiness > Machine ethics
7.3 Lack of capability or robustnessSafety & Trustworthiness > Psychological traits
7.3 Lack of capability or robustnessSafety & Trustworthiness > Robustness
7.3 Lack of capability or robustness