Exam AB-620 Topic 1 Question 1 Discussion
Actual exam question for Microsoft's AB-620 exam
Question #: 1
Topic #: 1
Question #: 1
Topic #: 1
Drag and Drop Question
You run the same fixed test set three times in Copilot Studio.
During evaluation, you observe the following:
- The same interaction fails in all three runs.
- Score values range from 0.58 to 0.61.
- The reasoning states that the response partially matches the expected answer.
- The knowledge source that is used is internal documentation.
- No tools are invoked.
You need to determine which conclusions are supported based on the evaluation results.
Which conclusions should you make? To answer, move the appropriate conclusions to the correct evaluation results. You may use each conclusion once, more than once, or not at all. You may need to move the split bar between panes or scroll to view content.
NOTE: Each correct selection is worth one point.

You run the same fixed test set three times in Copilot Studio.
During evaluation, you observe the following:
- The same interaction fails in all three runs.
- Score values range from 0.58 to 0.61.
- The reasoning states that the response partially matches the expected answer.
- The knowledge source that is used is internal documentation.
- No tools are invoked.
You need to determine which conclusions are supported based on the evaluation results.
Which conclusions should you make? To answer, move the appropriate conclusions to the correct evaluation results. You may use each conclusion once, more than once, or not at all. You may need to move the split bar between panes or scroll to view content.
NOTE: Each correct selection is worth one point.

Suggested Answer:

Explanation:
Box 1: A pattern failure across repeated runs.
The correct conclusion concerning evidence of a recurring issue is a) a pattern failure across repeated runs.
When the exact same interaction fails across three repeated runs, it establishes a systemic, predictable pattern rather than an isolated, one-time anomaly.
Box 2: The response partially matches the expected answer.
Based on the evaluation criteria in Microsoft Copilot Studio, the most definitive and direct conclusion that can be made concerning response quality is that the response partially matches the expected answer.
Direct Evidence: The evaluation results explicitly state in the reasoning that the "response partially matches the expected answer." This directly provides a qualitative conclusion regarding the quality of the generated response.
Score Alignment: The semantic similarity or quality score values ranging from 0.58 to 0.61 align with a partial match. In AI evaluation metrics, a perfect match is represented by 1.0, while values in the 0.6x range indicate that the core context was captured but lacked full completeness or exactness.
Box 3: Use of a connected knowledge source during the interaction
Because generative AI outputs are non-deterministic, running the same test set across a dynamic environment can yield varying scores. In Microsoft Copilot Studio, observing scores consistently between 0.58 and 0.61 points to a fundamental limitation with the underlying knowledge source.
The evaluation points to option use of a connected knowledge source during the interaction.
Because no tools are invoked, the agent heavily relies on generative search (retrieval-augmented generation) over internal documentation. The score range and reasoning indicate that the agent successfully retrieves a document but relies on a probabilistic model to paraphrase or assemble the response, which results in a "partial match" rather than an exact extraction.
Incorrect:
In contrast, specific underlying causes require explicit tool invocations or system error logs to diagnose, which are not present here.
Reference:
https://learn.microsoft.com/en-us/microsoft-copilot-studio/analytics-agent-evaluation-results
by Lance at Oct 05, 2026, 11:21 PM
0
0
0
10
Comments
Upvoting a comment with a selected answer will also increase the vote count towards that answer by one. So if you see a comment that you already agree with, you can upvote it instead of posting a new comment.
Report Comment
Commenting
You can sign-up / login (it's free).