試験

問題 44 / 65

Which design should be given top priority for a real-time chat inference platform to handle sudden traffic spikes such as during Black Friday?

Precompute and cache all possible responses
Prioritize frequently fine-tuning the model to keep it up to date.
Introduce containerization and autoscaling to dynamically scale inference nodes up and down.
Substantially increase the maximum number of tokens per response.
Adopt a policy of shortening responses at the expense of model accuracy.

当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。