OTS cluster becomes unreachable due to mgwd resource exhaustion
Applies to
- ONTAP Select
- ONTAP 9
Issue
- A 2-node ONTAP Select cluster became inaccessible. Node1 went down with the following console errors
swap_pager: out of swap spaceswp_pager_getswapspace(11): failed
- EMS logs also showed repeated mgwd process memory exhaustion and allocation failures.
[kern_mgwd:info] ERR: tz_mgmt: timezone error: magic_load() failed: Cannot allocate memory: cannot map `/usr/share/misc/magic.mgc` (Cannot allocate memory)[Node01:mgwd:kern.vm.mmap.return:notice]: mmap(2) by mgwd(pid5373) for size X failed: VMEM limit exceeded, limit 4294967296, error 12.
- Failover logs indicated a takeover due to lost heartbeat, followed by a manual giveback.
