
DatabricksCertified Machine Learning Associate
Domain 2Objective 7
Use One-Hot Encoding for Categorical Features MACHINE-LEARNING-ASSOCIATE Practice Questions (Page 1)
Part of the Section 2: Data Processing domain, which makes up ~27% of our current practice bank. Databricks does not publish an official question count, but from its 90-minute exam (~35–60 total, ~9–16 in this domain), expect 1–2 from this objective — we provide 23 practice questions to prepare you well beyond it. (estimate)
23questions here
5free pages
5concepts
Questions 1–5
- 1
A data scientist is building a churn prediction model with a logistic regression classifier. The dataset has a categorical feature 'product_category' with 500 categories. They are concerned about model interpretability and performance. Which approach best balances these concerns?
Select an answer first - 2
A data team is building a Spark ML pipeline for a classification task. They have a categorical column 'region' with 10 categories. They are considering whether to use one-hot encoding or leave the column as a string and rely on the model to handle it. Which statement is correct?
Select an answer first - 3
Which type of column is most appropriate for one-hot encoding?
Select an answer first - 4
In a Spark ML pipeline, which stage typically comes immediately before OneHotEncoder?
Select an answer first - 5
Which statement accurately describes the output of one-hot encoding for a categorical column with three distinct categories?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Databricks. “MACHINE-LEARNING-ASSOCIATE” is a trademark of its owner, used for identification only.