Azure CVO node shows unhealthy and disks in unknown state after maintenance event
Applies to
- Cloud Volumes ONTAP (CVO)
- ONTAP 9
- Microsoft Azure
Issue
- After an Azure platform maintenance event that caused non-simultaneous (asynchronous) reboots of both nodes in a CVO HA pair, one node (e.g., Cluster-02) reports as “unhealthy” with all local disks showing container type “unknown”.
- Aggregate(s) on the affected node appear offline or in an unknown state, but the partner node and its aggregates remain online.
Example CLI/log output:
Cluster::> cluster show -health false Node Health EligibilityCluster-01 false trueCluster-02 true true
Cluster::> storage disk show
Usable Disk Container Type Container Name OwnerNET-1.2 --- unknown Cluster-01NET-1.23 8.00TB unknown Cluster-01NET-1.34 8.00TB unknown Cluster-01
Cluster::> aggr show Aggregate State #Vols Nodes RAID Statusaggr0_Cluster_01 --- - unknown -aggr0_Cluster_02 online 1 Cluster-02 raid0, normalaggr1 --- - unknown -
Cluster::*> cluster ring show Node UnitName Epoch DB Epoch DB Trnxs Master Online--------- -------- -------- -------- -------- --------- ---------Cluster-01 mgmt 188 188 47621 Cluster-01 masterCluster-01 vldb 146 146 2423 Cluster-02 secondaryCluster-01 vifmgr 215 215 40 Cluster-02 secondaryCluster-01 bcomd 232 232 149 Cluster-02 secondaryCluster-01 crs 144 144 1 Cluster-02 secondaryCluster-02 mgmt 188 188 8415 Cluster-02 masterCluster-02 vldb 146 146 2423 Cluster-02 master
Node UnitName Epoch DB Epoch DB Trnxs Master Online--------- -------- -------- -------- -------- --------- ---------Cluster-02 vifmgr 215 215 40 Cluster-02 masterCluster-02 bcomd 232 232 149 Cluster-02 masterCluster-02 crs 144 144 1 Cluster-02 master
