試験
通常
試験時間: 00:00問題時間: 00:00
問題一覧
問題 29 / 65
Which metric is least suitable for evaluating the naturalness and contextual appropriateness of open-ended conversations?
①BLEU score
②ROUGE score
③Perplexity
④Human subjective evaluation
⑤RMSE
Which metric is least suitable for evaluating the naturalness and contextual appropriateness of open-ended conversations?
当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。