
Certified Tester AI Testing
Domain 5Objective 4
Benchmark Suites for ML CT-AI Practice Questions (Page 3)
Part of the Domain 5: ML Functional Performance Metrics domain, which makes up ~5% of our current practice bank. ISTQB does not publish an official question count, but from its 60-minute exam (~25–40 total, ~1–2 in this domain), expect 1–1 from this objective — we provide 23 practice questions to prepare you well beyond it. (estimate)
23questions here
5free pages
5concepts
Questions 11–15
- 11
A model achieves state-of-the-art results on the SuperGLUE benchmark. However, when deployed in a customer service chatbot, it performs poorly on real user queries. What is the most likely cause?
Select an answer first - 12
A team is developing a new image classification model and wants to compare it with existing models in a standardized way. They decide to use a benchmark suite. What is the primary purpose of using a benchmark suite in this context?
Select an answer first - 13
A team is presenting their model's performance to stakeholders. They claim their model is better than others because it has a higher score on a benchmark. What is a key caveat they should mention?
Select an answer first - 14
Which of the following is a known limitation of benchmark datasets when representing real-world scenarios?
Select an answer first - 15
A model performs well on the CIFAR-10 benchmark but poorly on a custom dataset of satellite images. What is the most likely reason?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by ISTQB. “CT-AI” is a trademark of its owner, used for identification only.