Glossary · Evaluation & safety
Adversarial example
An adversarial example is an input deliberately altered, often in ways humans barely notice, to cause an AI model to make a wrong prediction or produce an unintended output.
Adversarial example sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.
Also called Adversarial attack, Adversarial input.