Nutanix Resource Monitoring

Nutanix Resource Monitoring

Virtualization Monitoring

Nutanix Resource Monitoring

Monitor your entire Nutanix infrastructure from a single page. Covers the Prism environment summary, cluster configuration and resource usage, host hardware health, virtual machine performance, and storage container capacity.

Prism Overview

The Nutanix Prism page provides a summary of the entire Nutanix environment discovered under a Prism instance. Navigate to Virtualization → Nutanix → Nutanix Prism and click any Prism connection to access this view. It shows general Prism information alongside counts of all key entities: clusters, hosts, virtual machines, and storage containers.

SectionWhat it shows
Entity SummaryTotal counts of clusters, hosts, virtual machines, and storage containers discovered under this Prism instance.
Alert SummaryTotal alerts collected from Prism, how many remain unresolved, and a breakdown by Nutanix severity across Critical, Warning, and Info.
Resource HealthEvery discovered resource grouped by type with a health indicator against each one, identifying which specific virtual machine or container is degraded without opening each entity page.
Recent AlertsThe ten most recent unresolved alerts with their severity and age, ensuring teams are immediately informed of potential issues.
AlertsThe full list of alerts collected from Prism, showing severity, title, message, the entity type and name the alert relates to, the originating cluster, and whether the alert has been resolved or acknowledged in Nutanix.
EventsThe activity record collected from Prism. Events are not collected by default, so this view is populated only after Collect Nutanix Events has been enabled on the monitor.
Clusters, Hosts, VMs, Storage ContainersThe resource views described in this article, filtered to the resources belonging to this Prism instance rather than the whole estate.

Note: Cloudmon tracks three independent indicators for every Nutanix resource, and they can legitimately disagree. Status is the operational state reported by Nutanix, such as NORMAL or ONLINE. Health is the health assessment, such as GOOD or CRITICAL. State is the Cloudmon monitoring state, such as clear or down, reflecting whether Cloudmon can poll the resource. A cluster reporting Status NORMAL, State clear, and Health CRITICAL is valid: Nutanix considers it operational, Cloudmon is polling it successfully, and a health check has flagged a condition worth investigating.

Cluster Monitoring

Navigate to Virtualization → Nutanix → Nutanix Cluster and click any cluster to view its detail page. The list page is headed by Top Clusters by CPU Usage and Top Clusters by Memory Usage, ranking clusters by current utilisation. Cloudmon tracks the following for each cluster:

SectionWhat it shows
Resource UtilisationCPU, memory, and storage usage across the cluster, with storage reported as free capacity against total capacity.
Cluster InformationCluster name, UUID, status, architecture, and functions, alongside the Prism Central and AOS versions, whether the running release is on long-term support, and the node count. Useful for confirming clusters are on a consistent release before an upgrade.
Network and ConfigurationExternal and internal subnets, external IP, NFS subnet, name servers, NTP servers, timezone, domain awareness level, and desired redundancy factor. Name servers and NTP servers are the first place to look when diagnosing time drift or resolution failures reported by Nutanix alerts.
Hosts, VMs, and Storage ContainersAll hosts, virtual machines, and storage containers belonging to the cluster, with individual performance tracking available at each entity level.
System MetricsTime-series charts for all collected cluster-level metrics over the selected time range.
Outages and AlarmsA history of availability outages recorded for the cluster, and the alarm triggers and alarm history configured against it.

Host Monitoring

Navigate to Virtualization → Nutanix → Nutanix Host and click any host to view its detail page. The list page is headed by Top Hosts by CPU Usage and Top Hosts by Memory Usage, identifying the most heavily loaded hosts. Cloudmon tracks the following for each host:

SectionWhat it shows
Resource UtilisationProcessor and memory usage on the host, with CPU annotated by processor model and memory reported as the amount used against the total installed.
Host InformationHost name, UUID, and operational status, along with maintenance state and the degraded flag. Maintenance state confirms whether a drop in virtual machine count is planned rather than a fault, and the degraded flag indicates the cluster has detected the node performing abnormally while it remains online.
Network AddressesThe three addresses associated with a Nutanix node: the hypervisor, the Controller VM providing storage services, and the out-of-band management interface.
Hardware DetailsPhysical configuration of the host including processor model, sockets, cores, threads, and total CPU capacity, installed memory, disk count split by solid state and mechanical, and the hypervisor type in use.
Virtual MachinesAll virtual machines running on the host, useful for spotting uneven distribution across a cluster.
System MetricsTime-series charts for CPU and memory usage, input/output operations per second, input/output bandwidth, input/output latency, and network throughput. The input/output charts are each split into total, read, and write series, and network is split into receive and transmit.
Outages and AlarmsA history of availability outages recorded for the host, and the alarm triggers and alarm history configured against it.

Virtual Machine Monitoring

Navigate to Virtualization → Nutanix → Nutanix VM and click any virtual machine to view its detail page. The list page is headed by Top VMs by CPU Usage and Top VMs by Memory Usage. Virtual machines that are powered off report zero utilisation, so these charts reflect running workloads only. Cloudmon tracks the following for each virtual machine:

SectionWhat it shows
Resource UtilisationCPU and memory usage for the virtual machine, annotated with the number of virtual CPUs and the memory allocated to it.
VM InformationVirtual machine name, UUID, description, and power state, along with the virtual processor topology of vCPUs, sockets, and cores per socket, allocated memory, and attached disk count.
Guest and PlacementGuest operating system, network addresses, and the host currently running the virtual machine. Guest details are reported by Nutanix guest tools and show no value when the virtual machine is powered off or the tools are not installed. Host details are empty for a powered-off virtual machine, since it is not placed on a host until it starts.
Power State HandlingWhen a virtual machine is powered off, a banner reading Down Reason: Nutanix vm is in OFF state appears, availability reports zero, and CPU and memory usage report zero. This is the expected result of an intentional shutdown rather than a fault, and health can still report GOOD while the state is down.
System MetricsTime-series charts for CPU and memory usage, input/output operations per second, input/output bandwidth, input/output latency, and network throughput.
Outages and AlarmsA history of availability outages recorded for the virtual machine, and the alarm triggers and alarm history configured against it.

Storage Container Monitoring

Navigate to Virtualization → Nutanix → Nutanix Storage Container and click any container to view its detail page. The list page is headed by Top Containers by Total Capacity, Used Capacity, and Free Capacity, identifying where storage pressure is building before a container fills. Cloudmon tracks the following for each storage container:

SectionWhat it shows
CapacityUsed percentage alongside total, used, and free capacity. Containers drawing from a shared storage pool commonly report the same total capacity, since the pool is the underlying limit rather than the container.
Storage Container InformationContainer name, UUID, and operational status, along with the replication factor. A replication factor of 1 means Nutanix keeps no redundant copy, so a single node failure results in data loss for that container.
Storage EfficiencyWhether compression, deduplication, and erasure coding are enabled, and how much capacity each has reclaimed. A saving of zero on an enabled feature indicates the stored data is not reducing, which is common on containers holding already-compressed data.
System MetricsTime-series charts for input/output operations per second, bandwidth, and latency, each split into total, read, and write. Storage charts track used, capacity, free, and logical values over time, with a tier breakdown covering user used, user capacity, and the mechanical and solid state tiers. Controller input/output charts the operations served by the Controller VM.
Outages and AlarmsA history of availability outages recorded for the container, and the alarm triggers and alarm history configured against it.

Alarms

Alarms can be configured for individual Nutanix clusters, hosts, virtual machines, and storage containers, or at the group level for all Nutanix entities of the same type. The Nutanix Prism connection itself has no alarm configuration. Each alarm is built around a simple IF/THEN model, where you select a metric, set a threshold, and define what happens when it is breached. Learn more.