QuestionQ14

Test and manage agents

You run multiple evaluation tests for an agent in Copilot Studio before expanding access for users. The tests show the following results:

  • Each evaluation case is reported as meeting or missing the expected response.
  • Some evaluation cases repeatedly fail over several runs.
  • Each run shows the expected response and the response the agent generated.

Provide a conclusion based on the evaluation results.

Community Discussion

No comments yet. Be the first to start the discussion!