Search Agentik

CtrlK

Glossary · Evaluation & safety

BLEU

BLEU is an automatic metric that scores machine-generated text, originally translations, by measuring how many word sequences it shares with one or more human reference texts.

BLEU sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.

Also called Bilingual Evaluation Understudy.