Skip to main content
This is a research prototype. The data and analyses are preliminary and not yet validated — we'd welcome your .

Benchmarking (Guideline contamination)

Risk Sources and Risk Management Measures in Support of Standards for General-Purpose AI Systems

Gipiškis et al. (2024)

Sub-category
Risk Domain

Inadequate regulatory frameworks and oversight mechanisms that fail to keep pace with AI development, leading to ineffective governance and the inability to manage AI risks appropriately.

"Guideline contamination refers to scenarios where instructions for the collec- tion, annotation, or use of the dataset are exposed to the model [170]. These instructions may contain explicit data-label pairs that can improve the model’s capabilities for the task."(p. 19)

Supporting Evidence (1)

1.
"For example, for text-based models, this can include prompts used to generate synthetic data, as well as instructions for evaluators on the coverage and method of their evaluations of the model."(p. 19)

Other risks from Gipiškis et al. (2024) (144)