A city wants ML inference for permit-image checks but lacks staff to build and operate GPU clusters. Which design best delegates that complexity?
Select an answer to reveal the explanation.
Short Explanation
GPU clusters are a specialty shop most cities do not need to run. Point the permit checker at a managed AWS inference service and skip the rack, the drivers, and the 2 a.m. GPU tickets.
Full Explanation
Delegating advanced ML inference to managed AWS offerings makes the capability accessible without the city operating GPU infrastructure. Branch-office GPU farms, laptop training mandates, and email-based model distribution do not meet professional hosting, security, or scalability expectations.