
AWSCertified AI Practitioner
Domain 3Objective 4
Task Statement 3.4: Describe Methods to Evaluate FM Performance. AIF-C01 Practice Questions (Page 1)
Part of the Content Domain 3: Applications of Foundation Models domain, which makes up ~32% of our current practice bank. AWS does not publish an official question count, but from its 90-minute exam (~35–60 total, ~11–19 in this domain), expect 3–5 from this objective — we provide 40 practice questions to prepare you well beyond it. (estimate)
40questions here
8free pages
14concepts
Questions 1–5
- 1
What does 'answer faithfulness' mean in the context of RAG evaluation?
Select an answer first - 2
Which benchmark dataset is commonly used to evaluate text generation quality in tasks like summarization?
Select an answer first - 3
A legal research company has built a RAG system that answers questions based on a corpus of court documents. They want to evaluate whether the system's answers are faithful to the retrieved documents and whether the retrieval returns relevant passages. Which combination of metrics should they use?
Select an answer first - 4
How does BERTScore evaluate text generation?
Select an answer first - 5
A company is choosing between two foundation models for a document summarization tool. Model A produces slightly higher quality summaries but has a higher cost per interaction. Model B is cheaper but produces lower quality summaries, leading to more user escalations. The company wants to maximize user satisfaction while controlling costs. Which evaluation approach should they use?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by AWS. “AIF-C01” is a trademark of its owner, used for identification only.