Quiz 15 Question 2 of 20

A developer is building a chat interface on top of a Foundry-hosted GPT model. User testing shows that although the model typically finishes a full response in about six seconds, people perceive the assistant as slow and sometimes leave before the answer appears, because the message stays blank until the entire response is ready. Which change would most directly address this perception without changing the model or the total generation time?

Select an answer to reveal the explanation.

Motivation