AWS Certification Exam AI Practitioner (Practice Questions) set17

問題 45 / 65

I want to iteratively improve a chatbot's behavior using feedback obtained from user interactions. Which learning method is optimal?

Use reinforcement learning to learn a policy that maximizes reward by treating a dialogue as an episode.
Use supervised learning to train the utterance–response pairs in the conversation logs on hand just once.
Unsupervised learning to cluster similar utterances.
Improving the rule base by manually adding dialogue rules.
A method that simply generates new responses indiscriminately using a generative model.

当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。