Quiz 10 Question 2 of 20

A product team runs a multi-agent research-and-draft pipeline in Microsoft Foundry. They want a continuous improvement loop that scores draft quality on faithfulness and structure at scale, then feeds weak cases into prompt and tool refinements. Human labels exist only for a small golden set. Which approach best meets these goals?

Select an answer to reveal the explanation.

Motivation