Glossary · Evaluation & safety
Model inversion attack
A model inversion attack is a privacy attack that uses a model's outputs to reconstruct sensitive information about its training data, such as recovering features of individuals it learned from.
Model inversion attack sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.