
DatabricksCertified Generative AI Engineer Associate
Domain 6Objective 1
Evaluation Metrics and Model Selection GENERATIVE-AI-ENGINEER-ASSOCIATE Practice Questions (Page 3)
Part of the Section 6: Evaluation and Monitoring domain, which makes up ~17% of our current practice bank. Databricks does not publish an official question count, but from its 90-minute exam (~35–60 total, ~6–10 in this domain), expect 1–2 from this objective — we provide 19 practice questions to prepare you well beyond it. (estimate)
19questions here
4free pages
3concepts
Questions 11–15
- 11
A team is deploying an LLM for a document summarization service. Which metric is most appropriate to monitor the quality of the generated summaries?
Select an answer first - 12
A company deploys an LLM for a customer support email auto-responder. The system must generate responses that are helpful and resolve the customer's issue. The team is considering using a large model with high accuracy but high latency, or a smaller model with lower latency but slightly lower accuracy. The company has a service-level agreement (SLA) that requires responses within 2 seconds. Which model should the team choose?
Select an answer first - 13
A team is evaluating two LLMs for a machine translation task. Model A has a BLEU score of 30, and Model B has a BLEU score of 35. However, human evaluators rate Model A's translations as more fluent and natural. What is the most likely explanation?
Select an answer first - 14
A company is deploying an LLM for a legal document review tool. The tool must extract key clauses and summarize them. The team has two models: a 13B-parameter model that runs on a single GPU and a 70B-parameter model that requires multiple GPUs. The 70B model has higher accuracy on a legal benchmark, but the company has a limited budget for GPU infrastructure. The team also needs to process a high volume of documents daily. What is the most cost-effective approach that maintains acceptable accuracy?
Select an answer first - 15
A team is deploying an LLM-based customer support chatbot. Which metric is most important to monitor to ensure users receive responses quickly?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Databricks. “GENERATIVE-AI-ENGINEER-ASSOCIATE” is a trademark of its owner, used for identification only.