A team is preparing a test set to evaluate an agent in Copilot Studio before extending access to more users.
The test set must provide consistent evaluation across repeated runs and concentrate on the agent's most important behaviors.
You need to create a test set that enables repeatable, meaningful evaluation.
In which order should you perform the actions?
Drag & Drop
Community Discussion