AWS Certification Exam AI Practitioner (Practice Questions) set9

問題 19 / 65

Which speech synthesis architecture is recommended when you want to generate natural prosody and high-quality speech but keep the model's latency low?

Classical concatenative TTS (cut-and-paste from a corpus)
Completely rule-based parametric speech synthesis
A simple model that just generates mel spectrograms from text.
Combination of a sequence-to-sequence acoustic model and a lightweight neural vocoder (e.g., Tacotron2 + WaveRNN)
How to apply diffusion models directly to speech waveform generation

当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。