Examers.io
ExamsOrganizationsHow it worksPricingHelp & FAQ
Microsoft logo

Microsoft Certified:AI Agent Builder Associate

Domain 3Objective 1

Evaluate Agent Performance AB-620 Practice Questions (Page 3)

Part of the Test and manage agents domain, which accounts for 20–25% of the AB-620 exam. Microsoft does not publish an official question count, but from its 120-minute exam (~50–80 total, ~10–20 in this domain), expect 5–10 from this objective — we provide 38 practice questions to prepare you well beyond it. (estimate)

38questions here
8free pages
3concepts
20–25%of the exam

Questions 11–15

  1. 11expert · hard

    After an evaluation, an AI agent builder sees that the agent performs well on short queries but poorly on long, complex queries. The test set contains a mix of both. The builder needs to improve the agent's performance on complex queries. What should they do first?

    Select an answer first
  2. 12expert · hard

    A company has an AI agent that provides customer support. They have a test set of 1,000 queries. The evaluation shows 98% accuracy, but a manual review of a random sample reveals that the agent often gives correct answers but in a rude tone. The company has a strict policy that all customer interactions must be polite. The team is under pressure to release the agent. What should they do?

    Select an answer first
  3. 13expert · hard

    A company has an AI agent that provides HR policy answers. They need to evaluate the agent's accuracy and compliance with company policies. They have a test set of 100 queries with verified answers. The team is small and has limited time. Which evaluation approach balances accuracy and effort?

    Select an answer first
  4. 14expert · hard

    An AI agent for a government agency was evaluated, and the results show that the agent is 99% accurate on the test set. However, in production, users report that the agent often gives incorrect information. What is the most likely reason for this discrepancy?

    Select an answer first
  5. 15expert · hard

    A team is building a test set for an AI agent that handles product returns. They have collected 500 real user queries. They want to evaluate the agent's performance on different product categories and return reasons. They also want to ensure the test set is not too large to run frequently. What is the best approach?

    Select an answer first
Finished these 5 questions?

Review the revealed explanations, or continue through the curriculum.

Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by Microsoft. “AB-620” is a trademark of its owner, used for identification only.