Quiz 4 Question 12 of 20

A logistics company runs a nightly batch job that pushes roughly 50 million tokens through a Foundry-hosted model every night between 1 a.m. and 4 a.m., with almost no traffic the rest of the day. Operations needs a guaranteed, consistent level of throughput and latency during that fixed nightly window, and cost predictability matters more than minimizing per-request price. Which deployment option best fits this need?

Select an answer to reveal the explanation.

Motivation