AIF-C01 · Question #372
A company uses Amazon SageMaker AI to generate article summaries in multiple languages. The company needs a metric to evaluate the quality of the summary translations in multiple languages. Which…
The correct answer is B. Bilingual evaluation understudy (BLEU). Bilingual Evaluation Understudy measures how closely a generated translation matches one or more high-quality reference translations, making it well suited for evaluating the quality of multilingual summary translations.
Question
A company uses Amazon SageMaker AI to generate article summaries in multiple languages. The company needs a metric to evaluate the quality of the summary translations in multiple languages. Which evaluation metric will meet these requirements?
Options
- ARecall-Oriented Understudy for Gisting Evaluation (ROUGE)
- BBilingual evaluation understudy (BLEU)
- CArea Under the ROC Curve (AUC)
- DPrecision
How the community answered
(35 responses)- A9% (3)
- B71% (25)
- C17% (6)
- D3% (1)
Explanation
Bilingual Evaluation Understudy measures how closely a generated translation matches one or more high-quality reference translations, making it well suited for evaluating the quality of multilingual summary translations.
Topics
Community Discussion
No community discussion yet for this question.