Anonymizing Writing Style with LLM Rewri…

BackAI for Cybersecurity

This page is still being polished. If you have thoughts, please share them via the feedback form.

Data on this page is preliminary and may change. Please do not share or cite these figures publicly.

AI for Cybersecurity

Hodes (2025)|LLM classified

Mitigation Taxonomy

1AI System

1.2Non-Model

1.2.3Monitoring & Detection

Runtime behavior observation, anomaly detection, and activity logging.

Also in Non-Model

1.2.1 Guardrails & Filtering1.2.2 Runtime Environment1.2.4 Security Infrastructure1.2.5 Provenance & Watermarking

Definition

AI for cybersecurity refers to the application of artificial intelligence and machine learning techniques to enhance computer systems' security, detect threats, and protect against cyber attacks, representing a technological response to evolving cyber threats in an increasingly AI-powered threat landscape.

LLM Classification Details

Reasoning

Describes using AI for cybersecurity, not mitigating AI-specific risks or harms.

Code: 99.9Version: v0.5Classified: Jan 22, 2026

Other mitigations from Hodes (2025) (34)

Reduce Hallucinations

Reduce hallucination refers to techniques and methods used to minimize AI systems' tendency to generate false or fabricated information, addressing a critical challenge where language models produce inaccurate facts or citations that could spread misinformation.

1 AI System

Lifecycle:Build and Use ModelActor:DeveloperAIRM:Manage

Mitigate Hallucinations

Technical approaches to reduce LLM hallucinations - instances where AI models generate false or unsupported information while appearing confident in their responses

1 AI System

Lifecycle:Verify and ValidateActor:DeveloperAIRM:Manage

Detecting AI-Generated Content

Detecting AI-generated content involves technical methods and tools to identify whether content was created by artificial intelligence or humans, primarily through watermarking, linguistic analysis, and machine learning approaches.

1.2.5 Provenance & Watermarking

Lifecycle:Other (outside lifecycle)Actor:DeveloperAIRM:Measure

Risks from Persuasion

Risk that AI systems can systematically influence human beliefs and behaviors through sustained, personalized interactions by exploiting cognitive biases and adapting in real-time, enabling large-scale manipulation without human intervention.

99 Other

Lifecycle:Operate and MonitorActor:DeveloperAIRM:Govern

Content Moderation

Content moderation systems enable detecting and filtering toxic content (hate speech, harassment, misinformation) in real-time on digital platforms, while maintaining transparency in moderation decisions.

1.2.1 Guardrails & Filtering

Lifecycle:Operate and MonitorActor:DeployerAIRM:Manage

Make AI Manipulation Use Illegal

Legal framework to criminalize the malicious use of AI for manipulation of individuals or groups, including the creation and deployment of deepfakes and automated influence campaigns.

3.1.1 Legislation & Policy

Lifecycle:Other (outside lifecycle)Actor:Governance ActorAIRM:Govern

View all 34 mitigations from this source →

Source Document

Global Risk and AI Safety Preparedness (GRASP)

Hodes, Cyrus; Salem, Fadi; Corruble, Vincent; Ségerie, Charbel-Raphaël; Claybrough, Jonathan; Veron, Thibaud; Majid, Zainab; Fan, Jinyu; Lorin, Amaury (2025)

Project GRASP (Global Risk and AI Safety Preparedness) is a comprehensive database mapping AI risks and mitigation solutions. The initiative addresses both endogenous risk (autonomous AI systems that behave outside of human supervision) and exogenous risk (the human misuse of those AI systems). The platform serves policymakers, researchers, and industry leaders by providing tools required to identify risks, understand solutions, and find innovations.

View source

Classification

AI Lifecycle Stage

Other (outside lifecycle)

Outside the standard AI system lifecycle

Responsible Actor

Other

Actor type not captured by the standard categories

NIST AI RMF Function

Map

Identifying and documenting AI risks, contexts, and impacts

Risk Domains

Primary

2.2 AI system security vulnerabilities and attacks

Other

4.2 Cyberattacks, weapon development or use, and mass harm