Question
A model runs on a GPU endpoint, but GPU utilization is low, CPU utilization is high, and latency is dominated by preprocessing. What should be evaluated?
Flashcard practice
Practice multiple choice certification questions for AWS Certified Machine Learning Engineer - Associate and review the explanation after each answer.
A model runs on a GPU endpoint, but GPU utilization is low, CPU utilization is high, and latency is dominated by preprocessing. What should be evaluated?
Read aloud starts automatically.
Guest checks show correctness and the answer key. Log in to save history and unlock evaluator notes.
i Review the explanation and try similar questions to strengthen your understanding.
Issue reporting
Clara
Search across lessons, syllabus topics, provider capabilities, and certification questions.