跳转到主内容

Metrocluster IP 中缺少 NS224 磁盘架

Views:
Visibility:
Customer
Votes:
0
Category:
disk-shelves
Specialty:
hw
Last Updated:

适用场景

  • 4节点MCC-IP
  • NS224机架

问题描述

MCC-IP 中的两个集群都会触发自动支持警报,指示:
    • HA组通知(磁盘冗余失败)错误
    • HA组通知(SyncMirror丛失败)警报
    • HA 组通知(文件系统磁盘无响应)错误
    • HA 组通知(健康监视器进程 schm:RaidDegradedMirrorAggrAlert[4c95d23d-1cad-49e3-8b37-0531b7b4a6e6])警报
  • 在 EMS 日志中观察到针对同一磁盘架的多个错误消息:
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3b.21.1.18L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(9000).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3b.21.1.6L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(9000).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3b.21.1.23L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(9000).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3a.21.0.7L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(8999).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Disk device e3a.21.0.7L0: Check Condition: CDB 0xe2:01:0100000000000000:000000400000: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(8285).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.mcc.lunmgr.io.error:debug]: Disk device S/N XXXXXXXXXXXX - CDB 0xe2:01:0100000000000000:000000400000 - (scsi error: command aborted) - Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(DT 8285). (HA status 0x0) - (out_status_flags 0x24)
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3a.21.0.9L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(8999).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3a.21.0.19L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(9000).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.checkCondition:error]: Unknown device e3b.21.1.15L9998: Check Condition: CDB 0x12: Sense Data SCSI:aborted command -  (0xb - 0x90 0x2 0xfc)(8999).
[?]  Thu May 22 01:29:17 +0200 [CLUSTER-A01: scsi_cmdblk_strthr_admin: scsi.cmd.pastTimeToLive:error]: Disk device e3b.21.1.15L0: request failed after try #1: cdb 0xe2:01:0100000000000000:000000400000.
  • 在包含受影响磁盘的 MCC 架的所有四个节点中,系统存储配置从四路径转换为混合路径或多路径
  • 在 SYSCONFIG-A 中不可见,
  • MCC 的两侧都报告丢失的磁盘:

主集群

   RAID group /aggr1_CLUSTERA01_75_TB/plex0/rg1 (partial)

    RAID Disk   Device      HA  SHELF BAY CHAN Pool Type  RPM  Used (MB/blks)   Phys (MB/blks)
    ---------   ------      ------------- ---- ---- ---- ----- --------------   --------------
    dparity   FAILED         N/A             1831170/ -
    parity   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    Raid group is missing 5 disks.

灾难恢复集群

   RAID group /aggr1_CLUSTERB02_75_TB/plex1/rg1 (partial)

    RAID Disk   Device      HA  SHELF BAY CHAN Pool Type  RPM  Used (MB/blks)   Phys (MB/blks)
    ---------   ------      ------------- ---- ---- ---- ----- --------------   --------------
    dparity   FAILED         N/A             1831170/ -
    parity   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    data   FAILED         N/A             1831170/ -
    Raid group is missing 5 disks.

 

 

Sign in to view the entire content of this KB article.

New to NetApp?

Learn more about our award-winning Support

NetApp provides no representations or warranties regarding the accuracy or reliability or serviceability of any information or recommendations provided in this publication or with respect to any results that may be obtained by the use of the information or observance of any recommendations provided herein. The information in this document is distributed AS IS and the use of this information or the implementation of any recommendations or techniques herein is a customer's responsibility and depends on the customer's ability to evaluate and integrate them into the customer's operational environment. This document and the information contained herein may be used solely in connection with the NetApp products discussed in this document.