Hi,
We had a broken disk in our cluster (1.0.5), so I replaced it for a new disk (old disk X343, new disk X426 model). No big deal normally, but this time it's different. After assigning the disk to node 1 it shows status FAILED in the aggregate show:
aggr show-status -aggregate ch_n1_aggr002_FPP
Owner Node:xxxxxxx-001
Aggregate: ch_n1_aggr002_FPP (online, mixed_raid_type, degraded, hybrid) (block checksums)
Plex: /ch_n1_aggr002_FPP/plex0 (online, normal, active, pool0)
RAID Group /ch_n1_aggr002_FPP/plex0/rg0 (degraded, block checksums, raid_dp)
Usable Physical
Position Disk Pool Type RPM Size Size Status
-------- --------------------------- ---- ----- ------ -------- -------- ----------
shared 1.0.4 0 SAS 10000 1.61TB 1.64TB (normal)
shared FAILED - - - 1.61TB 0B (failed)
shared 1.0.6 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.7 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.8 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.9 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.10 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.11 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.12 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.23 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.14 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.13 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.16 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.17 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.18 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.19 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.15 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.21 0 SAS 10000 1.61TB 1.64TB (normal)
shared 1.0.20 0 SAS 10000 1.61TB 1.64TB (normal)
RAID Group /ch_n1_aggr002_FPP/plex0/rg1 (normal, block checksums, raid4) (Storage Pool: ch_n1_n3_sp_gen_001)
Usable Physical
Position Disk Pool Type RPM Size Size Status
-------- --------------------------- ---- ----- ------ -------- -------- ----------
shared 1.0.3 0 SSD - 894.2GB 3.49TB (normal)
shared 1.0.0 0 SSD - 894.2GB 3.49TB (normal)
shared 1.0.2 0 SSD - 894.2GB 3.49TB (normal)
22 entries were displayed.
xxxxxxx-001::> disk show -disk 1.0.5
Disk: 1.0.5
Container Type: shared
Owner/Home: ch-strt-gen-001 / ch-strt-gen-001
DR Home: -
Stack ID/Shelf/Bay: 1 / 0 / 5
LUN: 0
Array: N/A
Vendor: NETAPP
Model: X426_HCBFE1T8A10
Serial Number: 08G6H96Z
UID: 5000CCA0:2C0BCE98:00000000:00000000:00000000:00000000:00000000:00000000:00000000:00000000
BPS: 520
Physical Size: 1.64TB
Position: shared
Checksum Compatibility: block
Aggregate: -
Plex: -
Paths:
LUN Initiator Side Target Side Link
Controller Initiator ID Switch Port Switch Port Acc Use Target Port TPGN Speed I/O KB/s IOPS
------------------ --------- ----- -------------------- -------------------- --- --- ----------------------- ------ ------- ------------ ------------
ch-strt-gen-001 0a 0 N/A N/A AO RDY 5000cca02c0bce9a 34 6 Gb/S 0 0
ch-strt-gen-001 0b 0 N/A N/A AO INU 5000cca02c0bce99 3 12 Gb/S 3 0
ch-strt-gen-003 0a 0 N/A N/A AO RDY 5000cca02c0bce99 3 6 Gb/S 0 0
ch-strt-gen-003 0b 0 N/A N/A AO INU 5000cca02c0bce9a 34 12 Gb/S 1 0
Errors:
-
xxxxxxx-001::> set -privilege advanced
Warning: These advanced commands are potentially dangerous; use them only when directed to do so by NetApp personnel.
Do you want to continue? {y|n}: y
xxxxxxx-001::*> storage disk unfail -disk 1.0.5
Warning: Failed disk "1.0.5" might have aggregate labels and file system data present. In that case, this command will attempt to bring this disk back into the aggregate with which this disk had formerly been associated and preserve
file system data. Are you sure you want to continue with disk unfail? {y|n}: y
Error: command failed: Failed to unfail the disk. Reason: Disk is not currently failed.
I already tried another disk, but same problem.
Is this because it's a different model or something else?
thank for your time
Maurice