Search Agentik

CtrlK

Glossary · Evaluation & safety

Safety evaluation

A safety evaluation is a test that measures whether an AI model can produce harmful outputs or has dangerous capabilities, such as assisting cyberattacks, used to decide what safeguards are needed before release.

Safety evaluation sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.

Also called Dangerous capability evaluation.