SwitchIfOutErrorsWarn_Alert - AutoSupport 消息
适用于
- ONTAP 9
- 集群网络交换机
- 针对运行状况监控器进程 cshm 呼叫主页:SwitchIfOutErrorsWarn_Alert
事件摘要
在定期运行状况监控期间检测到错误时,会出现此消息。
- 系统健康监视器在监控子系统时会针对检测到的潜在问题创建警报。
- 警报包含有关可能原因的信息以及纠正问题的建议措施。
- 交换机接口"交换机名称/时隙:0 端口:4 10G - 级别"的出站报文错误百分比高于警告阈值。
- 出站数据包错误表示交换机接口在通过集群互连传输流量时遇到错误。
- 集群互连的性能下降可能导致集群不稳定或中断。
验证
AutoSupport 消息
HA Group Notification from Node Name (Health Monitor process cshm: SwitchIfOutErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level]) ALERT
事件日志
event log show -severity * -message-name callhome*
[Node Name Name: mgwd: callhome.hm.alert.major:alert]: Call home for Health Monitor process cshm: SwitchIfOutErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level].命令行
system health alert show -node <node name> -monitor cluster-switch -alert-id SwitchIfOutErrorsWarn_Alert
Node: NetApp-a Monitor: cluster-switch Class of Alert: SwitchIfOutErrorsWarn_Alert Severity of Alert: Major Probable Cause: Threshold_crossed Probable Cause Description: The percentage of outbound packet errors of switch interface "$(cluster_switch_analytics.unique-name)" is above the warning threshold. Possible Effect: Communication between nodes in the cluster might be degraded. Corrective Actions: 1) Migrate any cluster LIF that uses this connection to another port connected to a cluster switch. For example, if cluster LIF "clus1" is on port e0a and the other LIF is on e0b, run the following command to move "clus1" to e0b: "network interface migrate -vserver vs1 -lif clus1 -sourcenode node1 -destnode node1 -dest-port e0b" 2) Replace the network cable with a known-good cable. If errors are corrected, stop. No further action is required. Otherwise, continue to Step 3. 3) Move the network cable to another port on the node (if available). Migrate the cluster LIF to the new port. If errors are corrected, contact technical support to troubleshoot the original node port. Otherwise, continue to Step 4. 4) Move the network cable to another available cluster switch port. Migrate the cluster LIF back to the original port.
解决方法
- 对报告
SwitchIfOutErrorsWarn_Alert的端口执行链路故障排除。 - 尝试重新拔插电缆和 SFP/收发器。
- 检查接线板或中间连接,如有可能,请绕过接线板。
- 将电缆和/或 SFP 更换为已知良好的组件。
- 节点端 - 将受影响的集群 LIF 迁移到运行正常的集群端口,并将连接移至备用节点集群端口(如果可用)。
- 如果错误停止,请调查原始节点端口。(请联系 NetApp 技术支持 或登录 NetApp 支持站点以 创建案例。请参阅本文以获得进一步帮助。)
- 交换机侧 - 将连接移至已知良好的交换机端口。
追加信息