A community flea-market flyer is generated in Spanish against a human reference translation. Which metric family should score that job?
Select an answer to reveal the explanation.
Short Explanation
Think of a flea-market flyer in Spanish scored against a human translation. BLEU-style n-gram overlap is the usual metric. ROUGE is the summarization name, and a content filter is not an overlap score.
Full Explanation
BLEU is an overlap metric commonly used for translation. ROUGE is the usual summarization name. A content filter is not an overlap score, and automatic metrics do apply to translation-like jobs.