
DatabricksCertified Data Engineer Professional
Domain 2Objective 1
Design and Implement Data Ingestion Pipelines to Efficiently Ingest a Variety of Data Formats Including Delta Lake, Parquet, ORC, AVRO, JSON, CSV, XML, Text and Binary from Diverse Sources Such as Message Buses and Cloud Storage. DATA-ENGINEER-PROFESSIONAL Practice Questions (Page 7)
Part of the Data Ingestion & Acquisition domain, which accounts for 7% of the DATA-ENGINEER-PROFESSIONAL exam. Databricks does not publish an official question count, but from its 120-minute exam (~50–80 total, ~4–6 in this domain), expect 2–3 from this objective — we provide 34 practice questions to prepare you well beyond it. (estimate)
34questions here
7free pages
15concepts
7%of the exam
Questions 31–34
- 31
Which Spark format is used to read binary files into a DataFrame with columns for the file path, modification time, length, and content?
Select an answer first - 32
Which option in Auto Loader or DataFrameReader captures columns that are present in the data but not in the specified schema, storing them in a separate column?
Select an answer first - 33
Which Spark option allows you to read text files with a custom line delimiter instead of the default newline character?
Select an answer first - 34
A financial services company ingests real-time stock trade data from Kafka into a Delta table using Structured Streaming. The data volume is high, and the team is experiencing backpressure issues, causing the streaming job to fall behind. They need to increase throughput while maintaining exactly-once processing semantics. Which approach should they use?
Select an answer first
Finished these 4 questions?
Review the revealed explanations, or continue through the curriculum.
No more pagesBack to DATA-ENGINEER-PROFESSIONAL
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Databricks. “DATA-ENGINEER-PROFESSIONAL” is a trademark of its owner, used for identification only.