Glossary · Evaluation & safety
Model extraction
Model extraction is an attack in which someone repeatedly queries a model's API to reconstruct a close copy of the model or recover its behavior, threatening intellectual property and security.
Model extraction sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.
Also called Model stealing.