AWS Certified AI Practitioner AIF-C01 — Question 369
Topic 1 · Question 369 of 422
Topic 1 · Question 369
An AI Practitioner is using an LLM-as-a-judge in Amazon Bedrock to evaluate the quality of agent responses in a production environment. The AI practitioner wants to apply a built-in metric that assesses how thoroughly the agent responses address all parts of each prompt or question. Which metric will meet these requirements?