跳转到主内容

SwitchIfInErrorsWarn_Alert - AutoSupport 消息

Views:
148
Visibility:
Public
Votes:
0
Category:
fabric-interconnect-and-management-switches
Specialty:
hw
Last Updated:

适用于

  • ONTAP 9
  • 集群网络交换机
  • 针对运行状况监控进程 cshm 触发 Call home:SwitchIfInErrorsWarn_Alert

事件摘要

在定期运行状况监控期间检测到错误时,会出现此消息。

  • 系统运行状况监视器在监控子系统时会针对检测到的潜在问题创建警报。
  • 警报包含有关可能原因的信息以及纠正问题的建议措施。
  • 交换机接口"交换机名称/时隙:0 端口:4 10G - 级别"的入站数据包错误百分比高于警告阈值。
  • 集群互连中的降级可能导致集群不稳定或可能中断。

验证

AutoSupport 消息

HA Group Notification from Node Name (Health Monitor process cshm: SwitchIfInErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level]) ERROR

事件日志

event log show -severity * -message-name callhome*

[Node Name Name: mgwd: callhome.hm.alert.major:alert]: Call home for Health Monitor process cshm: SwitchIfInErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level].

命令行

system health alert show -node <node name> -monitor cluster-switch -alert-id SwitchIfInErrorsWarn_Alert

            Node: Netapp-a
          Monitor: cluster-switch
       Class of Alert: SwitchIfInErrorsWarn_Alert
     Severity of Alert: Major
       Probable Cause: Threshold_crossed
Probable Cause Description: The percentage of inbound packet errors of switch interface "$(cluster_switch_analytics.unique-name)" is above the warning threshold.
      Possible Effect: Communication between nodes in the cluster might be degraded.
     Corrective Actions: 1) Migrate any cluster LIF that uses this connection to another port connected to a cluster switch.
For example, if cluster LIF "clus1" is on port e0a and the other LIF is on e0b,
run the following command to move "clus1" to e0b:
"network interface migrate -vserver vs1 -lif clus1 -sourcenode node1 -destnode node1 -dest-port e0b"
2) Replace the network cable with a known-good cable.
If errors are corrected, stop. No further action is required.
Otherwise, continue to Step 3.
3) Move the network cable to another port on the node (if available).
Migrate the cluster LIF to the new port.
If errors are corrected, contact technical support to troubleshoot the original node port.
Otherwise, continue to Step 4.
4) Move the network cable to another available cluster switch port.
Migrate the cluster LIF back to the original port.
If errors are corrected, contact technical support to troubleshoot the original switch port.
If errors persist, contact technical support for further assistance.

解决方法

  1. 对报告  SwitchIfInErrorsWarn_Alert: 的端口进行链路故障排除
  2. 尝试重新插拔线缆和/或 SFP。
  3. 确认线缆和/或 SFP 是受支持的部件。
  4. 检查控制器和交换机之间是否存在任何配线架,如果可能,请绕过它。
    1. 如果问题仍然存在,请更换 SFP 和/或线缆。
    2. 如果问题仍然存在,请将连接切换到交换机端已知良好的端口,并检查是否报告警报。
  • 如果问题仍然存在
    • 对于 Broadcom 交换机,请联系 Broadcom 以获得帮助
    • 对于 Cisco 交换机,请联系 Cisco 以获得帮助
    • 对于 NVIDIA 交换机,请联系 NVIDIA 以获得帮助

追加信息

 

NetApp provides no representations or warranties regarding the accuracy or reliability or serviceability of any information or recommendations provided in this publication or with respect to any results that may be obtained by the use of the information or observance of any recommendations provided herein. The information in this document is distributed AS IS and the use of this information or the implementation of any recommendations or techniques herein is a customer's responsibility and depends on the customer's ability to evaluate and integrate them into the customer's operational environment. This document and the information contained herein may be used solely in connection with the NetApp products discussed in this document.