Which text quality metric specifically evaluates text simplification by comparing model output against both a reference text and the original source input?
Select an answer to reveal the explanation.
Short Explanation and Infographic
Think of it this way: in real-world AI governance, sari is exactly what teams reach for when they need to handle this scenario. SARI is the only standard metric that incorporates the original source sentence in its evaluation. On the exam, remember that this falls squarely under the 4.0 Configure Evaluation and Monitoring domain.
Full explanation below image
Full Explanation
SARI is the only standard metric that incorporates the original source sentence in its evaluation. This three-way comparison between output, reference, and source input makes SARI uniquely suited for evaluating text simplification tasks in watsonx.governance. The correct answer, "SARI", directly addresses the scenario described because it aligns with the specific governance requirement in question. The incorrect options ("BLEU", "ROUGE", "METEOR") may seem plausible but do not satisfy the core requirement. Understanding the distinction between these concepts is critical for IBM watsonx.governance implementations and is frequently tested in the 4.0 Configure Evaluation and Monitoring section of the certification exam.