試験

問題 29 / 65

Which metric is least suitable for evaluating the naturalness and contextual appropriateness of open-ended conversations?

BLEU score
ROUGE score
Perplexity
Human subjective evaluation
RMSE

当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。