Perform Operations and Monitoring in a Multicloud Environment
NCP-MCI-7.5 · 75 questions
- A Prism alert reports a CVM health issue on node-3, while the AHV host remains powered on and reachable. The cluster is otherwise stable, and the CVM process is not running on that node. Which condition best explains the alert?
- A Nutanix cluster raises an alert stating that a disk has failed and data resiliency is at risk. How should an administrator classify this alert, and why?
- A Nutanix cluster repeatedly generates a disk capacity alert. An administrator acknowledges the alert from Prism Central each time, but the alert returns and the disk remains low. What action is required to resolve the condition?
- During a planned hardware replacement, several AHV nodes will reboot and temporarily report node-down events. You must avoid paging the operations team while still preserving normal alerting after the window closes. What should you do?
- Your NOC requires every critical cluster alert from Prism to arrive by email and as a webhook POST to an external ticketing system. You need one configuration that can send the same alert to both targets. Which method should you use?
- A VM is powered off unexpectedly in a Prism Central-managed Nutanix AHV cluster. Operations must prove which authenticated user issued the power-off command. Which source should you check first?
- An NCC health check on a four-node AHV cluster reports a Zookeeper quorum issue. While troubleshooting, you need to explain why this matters. Which Nutanix capability is most directly affected by the failed quorum?
- An NCC health check on an AHV cluster reports Cassandra availability concerns. Prism Central shows normal VM power states, but cluster metadata operations are slow. Which component is directly affected?
- During routine monitoring, an NCC health check reports that the Stargate service is unhealthy on one CVM in an AHV cluster. The administrator needs to determine the likely operational impact before restarting the service. Which impact should be assessed first?
- An NCC health check fails for Curator on an upgraded AHV cluster. The admin sees no VM outage but wants to know which Nutanix AOS functions are most directly affected by Curator. Which functions should be investigated first?
- After completing an LCM upgrade on an AHV cluster, an administrator needs to confirm the cluster is healthy before returning workloads to normal operations. Which action should be taken first?
- A Prism alert reports a storage container capacity warning. The administrator wants to validate only that condition without running a broad cluster health sweep. Which action should be taken?
- During an NCC run, most checks return PASS, but a disk health check returns FAIL. How should the administrator interpret this result?
- An NCC health check reports significant time skew between nodes in a production AHV cluster. Before making any configuration changes, why must the administrator correct the skew?
- A Prism alert reports that a storage container has reached its configured capacity threshold. What is the primary operational purpose of this alert?
- During a weekly review, Prism Central shows a replication lag alert for a protection domain protecting several VMs between two Nutanix clusters. What does this alert primarily indicate?
- A four-node AHV cluster using RF2 loses one node to a failed drive. VMs continue to run, and the remaining nodes report normal CPU and network usage, but the cluster health score drops. What best explains the score change?
- During routine monitoring, a failed disk triggers multiple Prism alerts that reference impacted VMs and containers. You want to avoid opening each alert as if it were a separate incident. What should you use to reduce noise while preserving the disk failure as the underlying cause?
- A security reviewer asks for a record of administrative changes to cluster settings over the last 30 days. You must provide evidence without exporting performance data. Which source best satisfies this request?
- In Prism, a VM reports elevated disk latency while network latency stays near zero. CPU and memory metrics are normal, and the AHV host shows no CVM errors. What should an administrator treat as the most likely bottleneck?
- During peak workload hours, a Prism Central performance chart shows a VM with sustained high CPU ready time, while disk latency remains low. Which issue does this metric most directly indicate?
- During performance analysis, an AHV VM shows an inflated balloon driver and increased guest swapping while disk latency and network packet loss remain normal. Which condition most likely explains the VM's degraded performance?
- During a Prism performance review, you see cluster IOPS and storage latency rise sharply at the same time. A single VM’s IOPS graph jumps first, while node health, container capacity, and background jobs remain normal. Which issue best explains the pattern?
- A Prism Central performance graph shows two AHV VMs with nearly the same IOPS, but VM-A has much higher MB/s than VM-B. Latency is acceptable for both. What should you check first to explain the throughput difference?
- A single AHV VM’s application working set has grown over several days. Prism shows IOPS are similar, but the VM is now touching more unique data than before. What is the most likely performance symptom?
- An application team reports slow response times on several AHV VMs. Prism Central shows normal latency, IOPS, throughput, and no active alerts when you look at the last hour. What should you do first to determine whether performance has regressed?
- In a multi-cluster AHV estate, Prism Performance shows one SSD with sustained high read latency, while other disks report normal latency and IOPS. Which action best isolates a hot or degraded disk?
- During a Prism performance review, you notice a cluster's CVM CPU utilization is consistently above 90%, while storage latency has risen sharply. Prism shows many I/O requests waiting for CVM processing. Which cause best explains the latency increase?
- During nightly backups, Prism shows increased I/O latency for AHV VMs and slower protection-domain replication. Network monitoring shows inter-node latency spikes when backup traffic saturates the storage fabric, while disk utilization and CPU remain normal. Which factor most likely explains the degradation?
- A VM on AHV is reported by the guest application as having slow disk response. Prism performance charts show low storage latency and no storage contention. Which action should the administrator take first?
- A Prism performance chart shows a single AHV VM's disk IOPS repeatedly hitting a flat ceiling during business hours, while the CVM's Curator and Stargate metrics remain healthy and other VMs on the same container are not affected. What should you check first to confirm the cause of this flat IOPS pattern?
- During the first 15 minutes of a nightly reporting job, an AHV VM shows high IOPS and low latency. After that window, IOPS drop to a flat ceiling and disk latency rises, while host CPU, CVM CPU, storage container health, and memory pressure remain normal. Which cause should you investigate first?
- In a Nutanix AHV cluster, Prism Central performance charts show high CVM CPU and increasing write latency during a nightly batch job. The VMs use a storage container with compression and deduplication enabled. What should you do to address the performance impact?
- After a disk failure, EC-X rebuild begins. Prism performance shows VM disk latency rise, lower IOPS, and no CPU or memory anomalies on CVMs or VMs. Which performance factor should you investigate first?
- A lab powers on many AHV VMs provisioned from the same golden snapshot. Prism shows a sharp rise in storage read IOPS and read latency, while write latency remains normal. Which issue best explains the spike?
- During business hours, Prism shows AHV hosts with high CPU oversubscription while storage and network metrics remain normal. Users report slow VM response. What should you expect?
- An AHV VM shows low guest CPU utilization, high memory pressure, and rising guest swap activity. Prism Central also reports ballooned memory for the VM, while its allocated memory has not changed. What should you investigate first?
- During nightly backups, a Prism Central performance chart shows no unusual IOPS or latency, yet users report brief slowdowns. The default 24-hour chart smooths the issue. What should you adjust to make the short backup-window spikes visible?
- A multi-node AHV cluster shows 50 TB of free usable capacity in Prism. Capacity reports show consumed capacity has grown by 5 TB per month for the last three months. If no nodes are added and no VMs are deleted or compressed differently, how long is the estimated runway before free capacity is exhausted?
- An administrator checks a storage container in Prism Central. Total capacity is 100 TiB, used capacity is 55 TiB, but free capacity reports only 35 TiB. The capacity view also lists 10 TiB as reserved capacity. Why does free capacity not equal total minus used?
- A Prism Central capacity view shows logical capacity consumption at 95% while physical usable capacity consumption is only 50%. The VMs use thin-provisioned disks and RF2. Which statement explains the difference?
- A Nutanix AHV cluster stores production VMs in a container configured with RF2. The container reports 200 TiB raw capacity, and workloads are growing. An admin forecasts runway and must account for replication. Which calculation should the admin use to estimate baseline usable capacity?
- A Nutanix cluster shows 40% capacity savings from compression and deduplication, but CVM CPU utilization rises after enabling them. When forecasting runway, what should you factor in alongside saved capacity?
- A capacity planner notices that free space on an AHV cluster grows after EC-X is enabled on a container. They need a defensible runway forecast for the next quarter. Which consideration must be included?
- A capacity report shows free space dropping and runway shrinking after a protection policy was changed from 7-day snapshot retention to 30-day snapshot retention. Which action should the administrator take to reduce the capacity consumption caused by this change?
- A Prism Central administrator notices a Nutanix AHV cluster's storage container usage growing faster than expected after many VMs are cloned from the same Windows template. VM sizes and snapshots appear normal in the VM list, but capacity runway shortens. Which action should you take first to identify the cause?
- An administrator is forecasting runway for a growing Nutanix AHV cluster. Raw data is increasing slowly, but the number of managed disks and VMs is climbing. Which factor should be included in the capacity model?
- A Nutanix cluster’s capacity runway shows it will run out of usable capacity in six weeks. What should an administrator do first?
- A four-node AHV cluster reports 92% storage utilization and a six-week runway; CPU and memory are below 45%, and data reduction is already enabled. You must restore capacity without adding unnecessary compute. What should you do?
- During a capacity review, Prism Central shows low usable capacity and a short storage runway. One AHV node has large disks with free space, but the cluster still appears short. Which consideration best explains why the extra disks are not relieving usable capacity?
- A Nutanix administrator is planning capacity for a mixed-size AHV cluster. The cluster must keep normal capacity headroom and still absorb the data from one failed node without exceeding storage capacity limits. Which capacity-planning action directly provides the failure-domain reserve?
- A new application will deploy several AHV VMs with high vCPU and memory reservations. Before placing them, you review cluster capacity. Which planning step prevents placement failures?
- A four-node AHV cluster shows sustained memory pressure in Prism Central, while several VMs have large memory reservations but low guest utilization. When planning capacity, what should the administrator conclude?
- A capacity planner sees 15% free memory on a four-node AHV cluster and wants to place 20 more steady-state VMs without buying hardware. Increasing memory oversubscription would allow placement, but what should the planner do first?
- A DR cluster was sized using only the primary production VM disk allocation. When forecasting storage runway for DR expansion, which DR-specific capacity component must be added?
- Your capacity runway across a production AHV cluster and its backup target has fallen from 90 to 45 days. A change review shows local snapshot retention increased from 7 to 30 days and remote backup retention from 30 to 180 days. Which capacity planning adjustment best explains the shorter runway?
- A four-node AHV cluster reports 70% free storage capacity and a 90-day capacity runway, but production VMs show high storage latency and cannot meet required IOPS. An administrator is asked whether the cluster has a capacity problem. Which conclusion should be reported?
- An automation engineer must call Prism Central REST APIs to generate a weekly capacity report. The script must authenticate without embedding a user password in source control. Which method should the engineer use?
- A Nutanix administrator must identify which Prism Central REST endpoints exist and see sample request bodies before automating VM creation. Which approach should they use?
- An administrator must deliver a daily capacity report for a multi-cluster Nutanix estate without logging into Prism each morning. Which automation method should be used?
- After a planned maintenance window, you must power on 40 AHV VMs so databases start first, application servers next, and web servers last. Repeating the action manually in Prism Central is error-prone. Which approach best automates the sequence?
- A Prism Central alert fires when a production VM's CPU ready exceeds an administrator-defined threshold. The team wants the cluster to automatically pause the VM and notify the owner without manual intervention. Which Nutanix capability should be used to implement this alert-driven remediation?
- Your operations team requires a cluster health check to run automatically every Monday at 02:00. Which Nutanix automation capability should you use?
- Your X-Play playbook triggers on an alert and automatically runs a risky remediation action on a production VM. The administrator must prevent the action from executing until an authorized operator explicitly approves it, while keeping the playbook event-driven. How should the playbook be designed?
- You create an X-Play playbook in Prism Central. Step 1 discovers a VM UUID from a cluster query, and step 2 must use that UUID to power off the VM. How should the discovered identifier be carried between the steps?
- Your team must automate a weekly shutdown and startup of a set of AHV VMs in Prism Central. The workflow should be centralized, supported, and avoid custom scripts or node-level commands. Which approach should you use?
- A Prism Central automation playbook provisions a workload VM by name. During a cleanup window the target VM was deleted, so the playbook fails at the first step because the object no longer exists. Which design change makes the playbook resilient without hiding unrelated failures?
- An X-Play playbook in Prism Central produces automation results that a third-party ticketing system must receive immediately. Which integration should you implement so the ticketing system can create tickets from the playbook output?
- An operator must fetch the current alert status from Prism Central programmatically for a dashboard, without opening the UI. Which action should be used?
- In a Prism Central estate, a Prism alert repeatedly warns that a container is approaching capacity. You want the platform to start a cleanup playbook only when the alert fires, without polling. Which approach uses event-driven automation?
- Your team runs planned maintenance on an AHV host, and Prism Central raises a known maintenance alert. An X-Play automation must mark the alert as handled while leaving the alert history and monitoring intact. Which automation action should it perform?
- A Nutanix Playbook workflow provisions and configures a VM. Operations must receive a completion notice in both the email distribution list and the team collaboration channel as soon as the workflow succeeds. Which configuration should you use?
- An Ansible playbook that creates VMs through the Prism Element REST API works on AOS 6.8 but starts failing after a cluster upgrade to AOS 7.5. The playbook now receives HTTP 404 responses and cannot parse some response fields. What should you verify first?
- An auditor asks which automation changed a VM's power state after a recent change window. You know an X-Play playbook and an external API script could both have done it. Which source should you use to identify the specific playbook run that changed the VM state?
- An administrator needs an automated weekly capacity and health summary for a Nutanix multicloud estate. The summary must combine cluster health and capacity metrics without manual console checks. Which approach meets this requirement?