
DatabricksCertified Data Engineer Professional
Domain 2Objective 1
Design and Implement Data Ingestion Pipelines to Efficiently Ingest a Variety of Data Formats Including Delta Lake, Parquet, ORC, AVRO, JSON, CSV, XML, Text and Binary from Diverse Sources Such as Message Buses and Cloud Storage. DATA-ENGINEER-PROFESSIONAL Practice Questions (Page 6)
Part of the Data Ingestion & Acquisition domain, which accounts for 7% of the DATA-ENGINEER-PROFESSIONAL exam. Databricks does not publish an official question count, but from its 120-minute exam (~50–80 total, ~4–6 in this domain), expect 2–3 from this objective — we provide 34 practice questions to prepare you well beyond it. (estimate)
34questions here
7free pages
15concepts
7%of the exam
Questions 26–30
- 26
Which of the following is a common source for Structured Streaming in Databricks?
Select an answer first - 27
Which Spark option allows you to handle malformed JSON records by storing them in a separate column instead of failing the entire read?
Select an answer first - 28
When reading JSON data with nested structures, how does Spark typically represent nested objects?
Select an answer first - 29
What is the data type of the 'content' column when reading binary files with Spark's binaryFile format?
Select an answer first - 30
Which Spark configuration or option controls the target file size when writing DataFrames, helping to avoid too many small files?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Databricks. “DATA-ENGINEER-PROFESSIONAL” is a trademark of its owner, used for identification only.