Glossary · Evaluation & safety
ROUGE
ROUGE is a family of automatic metrics that evaluate generated text, especially summaries, by measuring the overlap of words and word sequences with human-written reference texts.
ROUGE sits in the Evaluation & safety part of the Agentik {OS} glossary, which defines the words used to build and run AI agent systems.