Quiz 11 Question 7 of 20

You are building an evaluation dataset for a coding agent that fixes GitHub Issues. The dataset must allow you to objectively measure whether the agent's fix is correct. What makes an evaluation example in this dataset complete and objectively scoreable?

Select an answer to reveal the explanation.

Motivation