試験

問題 11 / 65

Which architecture is widely adopted as the foundation for multimodal models that handle both text and images?

Convolutional Neural Network (CNN)
Recurrent Neural Network (RNN)
Transformer (Transformer)
Graph Neural Network (GNN)
Hidden Markov Model (HMM)

当サイトでは、ユーザー体験の向上を目的としてCookieを使用しています。サイトの利用を継続することで、Cookieの使用に同意したものとみなされます。