Microsoft Certified:Azure Databricks Data Engineer Associate
Domain 4Objective 1
Design and Implement Data Pipelines DP-750 Practice Questions (Page 4)
Part of the Deploy and maintain data pipelines and workloads domain, which accounts for 30–35% of the DP-750 exam. Microsoft does not publish an official question count, but from its 120-minute exam (~50–80 total, ~15–28 in this domain), expect 4–7 from this objective — we provide 39 practice questions to prepare you well beyond it. (estimate)
39questions here
8free pages
6concepts
30–35%of the exam
Questions 16–20
- 16
A team is building a notebook-based pipeline in Azure Databricks. The pipeline has three tasks: ingest data, clean data, and train a model. The model training task must only run if the cleaning task succeeds. The team also wants to send an alert if any task fails. What should they configure in the job?
Select an answer first - 17
A data pipeline processes a large dataset and performs a join between two tables. The team notices that the join is very slow and wants to improve performance. They also need to ensure that the pipeline can handle occasional data quality issues without failing. Which approach should they take?
Select an answer first - 18
A team is designing a Lakeflow job that processes financial transactions. The job has three tasks: ingest, validate, and aggregate. The validate task checks for duplicate transactions and flags anomalies. The aggregate task computes daily totals. The team wants to ensure that if the validate task finds anomalies, the aggregate task still runs but the results are marked as 'review required'. If the validate task fails due to a system error, the job should stop. How should the team design the task logic?
Select an answer first - 19
When creating a notebook-based pipeline in Databricks Jobs, how do you specify that a notebook task should run after another notebook task?
Select an answer first - 20
A team is building a notebook-based pipeline in a Lakeflow job. The pipeline has three tasks: ingest, clean, and publish. The ingest task reads data from an external source and writes to a staging table. The clean task performs transformations and writes to a final table. The publish task copies the final table to a production location. The team wants to ensure that if the clean task fails, the publish task does not run, but the ingest task should not be retried. They also want to be able to manually retry only the clean task without re-running ingest. What should they do?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Microsoft. “DP-750” is a trademark of its owner, used for identification only.