
Google CloudProfessional Machine Learning Engineer
Domain 4Objective 2
Scaling Online Model Serving PROFESSIONAL-MACHINE-LEARNING-ENGINEER Practice Questions (Page 8)
Part of the Serving and scaling models domain, which accounts for ~20% of the PROFESSIONAL-MACHINE-LEARNING-ENGINEER exam.
37questions here
8free pages
14concepts
~20%of the exam
Questions 36–37
- 36
Which hardware is typically best suited for serving a large transformer-based language model with low latency?
Select an answer first - 37
A fintech company serves a fraud-detection model that needs real-time features like transaction velocity. The features are computed in a streaming pipeline and stored in Vertex AI Feature Store. The model is deployed to a Vertex AI endpoint. How should the serving application retrieve features for each prediction request?
Select an answer first
Finished these 2 questions?
Review the revealed explanations, or continue through the curriculum.
No more pagesBack to PROFESSIONAL-MACHINE-LEARNING-ENGINEER
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Google Cloud. “PROFESSIONAL-MACHINE-LEARNING-ENGINEER” is a trademark of its owner, used for identification only.