Quiz 12 Question 5 of 20

A research firm wants a Foundry-hosted assistant to answer questions about individual 300-page technical reports. Each report is far larger than the deployed model's context window, and sending the full document in one prompt causes errors. What approach lets the assistant answer questions grounded in a specific report without exceeding the model's input limit?

Select an answer to reveal the explanation.

Motivation