Quiz 11 Question 14 of 20

An enterprise FAQ multi-agent solution on Azure repeatedly pays for full model completions on near-identical user questions such as “How do I reset my VPN password?” and “What’s the process to reset VPN credentials?” Specialists and the orchestrator share the same grounding corpus. The team wants to cut token cost and latency for semantically equivalent questions while still allowing true novel questions to hit the models. Which caching strategy best fits?

Select an answer to reveal the explanation.

Motivation