Quiz 15 Question 16 of 20

A payments company runs a Foundry-hosted GPT-4o deployment that handles real-time fraud-check calls during checkout, at a steady 50,000 requests per hour, and the team needs predictable, consistent response latency even during peak shopping days. Which deployment type should they choose?

Select an answer to reveal the explanation.

Motivation