A town help-desk Helm chart lists a GPU for the NVIDIA Riva speech containers. A volunteer offers to write a custom CUDA kernel “so Kubernetes can see the device.” What should the associate keep in scope for this conversational experiment deploy?
Select an answer to reveal the explanation.
Short Explanation
Think of the chart as a shipping label that already says “needs a GPU.” The associate’s job is to honor that label for Riva, not to carve a custom CUDA kernel or redesign the cluster fabric. Keep awareness at “speech needs a GPU” and move on.
Full Explanation
For an associate Helm deploy of finished Riva speech containers, GPU awareness means the chart values request a GPU because ASR/TTS expect one. Writing custom kernels, rewriting device plugins, or redesigning NCCL topology is out of NCA-GENM scope. Do not drop speech for a text-only stack just to avoid GPUs.