Glossary · Evaluation & safety
AI safety
AI safety is the field concerned with preventing AI systems from causing accidental or deliberate harm, covering research on alignment, robustness, misuse prevention, monitoring and the control of increasingly capable models.
AI safety sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.