
Dell Data Science Optimize
Domain 2Objective 2
Hive DATA-SCIENCE-OPTIMIZE Practice Questions (Page 3)
Part of the Hadoop Ecosystem and NoSQL domain, which accounts for 15% of the DATA-SCIENCE-OPTIMIZE exam.
31questions here
7free pages
8concepts
15%of the exam
Questions 11–15
- 11
A Hive query performs a large aggregation over a fact table. The team notices that the query runs slowly because it processes rows one at a time. Which Hive feature should they enable to process multiple rows at once and improve CPU efficiency?
Select an answer first - 12
A Hive table stores sensor readings. Each reading has a timestamp, a sensor ID, and a JSON payload with variable fields. The team wants to query specific fields from the JSON payload without parsing it manually in every query. Which Hive data type or feature should they use?
Select an answer first - 13
Which HiveQL statement is used to add a new column to an existing table?
Select an answer first - 14
A Hive query joins a large fact table with a small dimension table (about 50 MB). The query runs as a reduce-side join, causing high network traffic. The team wants to reduce shuffle overhead. Which approach should they take?
Select an answer first - 15
A Hive query is running slower than expected. The team suspects that the execution engine is not using all available resources. They check the YARN configuration and find that the cluster has 100 nodes, but the query is only using 10 containers. Which Hive or YARN setting is most likely to increase parallelism?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Dell Technologies. “DATA-SCIENCE-OPTIMIZE” is a trademark of its owner, used for identification only.