Examers.io
ExamsOrganizationsHow it worksPricingHelp & FAQ
Google Cloud logo

Google CloudProfessional Machine Learning Engineer

Domain 2Objective 1

Exploring and Preprocessing Data for ML PROFESSIONAL-MACHINE-LEARNING-ENGINEER Practice Questions (Page 2)

Part of the Collaborating within and across teams to manage data and models domain, which accounts for ~16% of the PROFESSIONAL-MACHINE-LEARNING-ENGINEER exam.

24questions here
5free pages
4concepts
~16%of the exam

Questions 6–10

  1. 6expert · hard

    A streaming media company needs to preprocess clickstream data for a real-time personalization model. The data arrives as a stream of 100,000 events per second. The preprocessing requires windowed aggregations (e.g., 'clicks in last 5 minutes per user') and enrichment with a static lookup table of content metadata. The team needs to minimize operational overhead and wants a fully managed solution. What should they use?

    Select an answer first
  2. 7application · medium

    A media company receives user interaction logs as a continuous stream of JSON events. The ML team needs to build a real-time feature for 'average session duration per user' to serve a recommendation model with sub-second latency. The team also needs to ensure that user IDs are hashed before the feature is stored. Which combination of tools should they use?

    Select an answer first
  3. 8foundation · easy

    A dataset contains a column with customers' email addresses. To protect PII during model training, which technique is most appropriate?

    Select an answer first
  4. 9expert · hard

    A company has a 50 TB dataset of user behavior logs in Cloud Storage. The ML team needs to create a set of aggregate features for a recommendation model. The team has a tight budget and wants to minimize compute costs. The features need to be updated daily. What is the most cost-effective approach among the following?

    Select an answer first
  5. 10foundation · easy

    A machine learning engineer is starting a new project and needs to decide how to organize the available data for efficient experimentation. The data consists of a CSV file with customer transactions, a folder of JPEG images, and a set of text documents. Which approach best organizes this data for ML workflows?

    Select an answer first
Finished these 5 questions?

Review the revealed explanations, or continue through the curriculum.

Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Google Cloud. “PROFESSIONAL-MACHINE-LEARNING-ENGINEER” is a trademark of its owner, used for identification only.