
AWSCertified Generative AI Developer - Professional
Domain 2Objective 2
Task 2.2: Implement Model Deployment Strategies. AIP-C01 Practice Questions (Page 4)
Part of the Content Domain 2: Implementation and Integration domain, which accounts for 26% of the AIP-C01 exam.
25questions here
5free pages
9concepts
26%of the exam
Questions 16–20
- 16
A company is deploying a 70B-parameter LLM on AWS for a real-time chat application. The model requires approximately 140GB of GPU memory. They have a cluster of 8x A100 GPUs (80GB each) and need to minimize inter-GPU communication overhead. Which container-based deployment pattern should they use?
Select an answer first - 17
When balancing performance and resource usage for a GenAI workload, what is a key trade-off to consider?
Select an answer first - 18
Which approach can help balance performance and resource usage for a GenAI workload?
Select an answer first - 19
Which scenario is a typical use case for Amazon SageMaker AI endpoints in a hybrid solution?
Select an answer first - 20
An application requires real-time responses with low latency for a variable number of requests. Which deployment strategy is most appropriate for this workload?
Select an answer first
Finished these 5 questions?
Review the revealed explanations, or continue through the curriculum.
Free Basic Practice is a study aid with revealable answers — not a scored exam. Examers.io is independent and not affiliated with or endorsed by AWS. “AIP-C01” is a trademark of its owner, used for identification only.