Examers.io
ExamsOrganizationsHow it worksPricingHelp & FAQ
Databricks logo

DatabricksCertified Data Engineer Professional

Domain 1Objective 2

Building and Testing an ETL Pipeline with Lakeflow Spark Declarative Pipelines, SQL, and Apache Spark on the Databricks Platform DATA-ENGINEER-PROFESSIONAL Practice Questions (Page 5)

Part of the Developing Code for Data Processing using Python and SQL domain, which accounts for 22% of the DATA-ENGINEER-PROFESSIONAL exam. Databricks does not publish an official question count, but from its 120-minute exam (~50–80 total, ~11–18 in this domain), expect 6–9 from this objective — we provide 29 practice questions to prepare you well beyond it. (estimate)

29questions here
6free pages
11concepts
22%of the exam

Questions 21–25

  1. 21foundation · easy

    What is the purpose of the 'auto-optimize' configuration in a Lakeflow pipeline?

    Select an answer first
  2. 22application · medium

    A data engineer needs to automate the deployment of a Lakeflow Spark Declarative Pipeline. The pipeline definition is stored in a Git repository. The engineer wants to update the pipeline in the Databricks workspace whenever a new commit is pushed to the main branch. Which approach should they use?

    Select an answer first
  3. 23application · medium

    A team manages a Lakeflow Spark Declarative Pipeline that runs on a schedule. They need to trigger a one-time run of the pipeline outside the normal schedule after a data backfill. They also want to monitor the run's status and logs. Which approach should they use?

    Select an answer first
  4. 24application · medium

    A data engineer needs to orchestrate an ETL workload that consists of a notebook that ingests data, a SQL query that transforms it, and a final notebook that loads it into a reporting table. The workload must run daily at 2 AM, and if the ingestion task fails, the subsequent tasks should not run. Which approach should be used?

    Select an answer first
  5. 25foundation · easy

    When would you choose to use raw Structured Streaming instead of a Lakeflow Spark Declarative Pipeline?

    Select an answer first
Finished these 5 questions?

Review the revealed explanations, or continue through the curriculum.

Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Databricks. “DATA-ENGINEER-PROFESSIONAL” is a trademark of its owner, used for identification only.