mmlsrecoverygroup command

Lists information about IBM Storage Scale RAID recovery groups.

Synopsis

mmlsrecoverygroup [ RecoveryGroupName [-L [--pdisk] ] ] 

Availability

Available on all IBM Storage Scale editions.

Description

The mmlsrecoverygroup command lists information about recovery groups. The command displays various levels of information, depending on the parameters specified.

1. Output values for mmlsrecoverygroup

Descriptions of the output values for the mmlsrecoverygroup command follow.
recovery group
Is the name of the recovery group.
declustered arrays with vdisks
Is the number of declustered arrays with vdisks in this recovery group.
vdisks
Is the number of vdisks in this recovery group.
servers
Is the server pair for the recovery group. The intended primary server is listed first, followed by the intended backup server.

2. Output values for mmlsrecoverygroup RecoveryGroupName

Descriptions of the output values for the mmlsrecoverygroup RecoveryGroupName command follow, as displayed by row, from top to bottom and left to right.

recovery group
Is the name of the recovery group.
declustered arrays with vdisks
Is the number of declustered arrays with vdisks in this recovery group.
vdisks
Is the number of vdisks in this recovery group.
servers
Is the server pair for the recovery group. The intended primary server is listed first, followed by the intended backup server.
declustered array with vdisks
Is the name of the declustered array.
vdisks
Is the number of vdisks in this declustered array.
vdisk
Is the name of the vdisk.
RAID code
Is the RAID code for this vdisk.
declustered array
Is the declustered array for this vdisk.
remarks
Indicates the special vdisk type. Only those vdisks with a dedicated function within the recovery group are indicated here: the log, log tip, log tip backup, and log reserved vdisks. This field is blank for file system NSD vdisks.

3. Output values for mmlsrecoverygroup RecoveryGroupName -L

Descriptions of the output values for the mmlsrecoverygroup RecoveryGroupName -L command follow, as displayed by row, from top to bottom and left to right.

Recovery group section:

recovery group
Is the name of the recovery group.
declustered arrays
Is the number of declustered arrays in this recovery group.
vdisks
Is the number of vdisks in this recovery group.
pdisks
Is the number of pdisks in this recovery group.
format version
Is the recovery group version.

Declustered array section:

declustered array
Is the name of the declustered array.
needs service
Indicates whether this declustered array needs service. A yes value means that disks need to be replaced.
vdisks
Is the number of vdisks in this declustered array.
pdisks
Is the number of pdisks in this declustered array.
spares
The first number of the pair is the amount of spare space that is reserved for rebuilding, expressed as an equivalent number of pdisks. This spare space is allocated equally among all of the pdisks in the declustered array.

The second number of the pair is the number of vdisk configuration data (VCD) replicas that are maintained across all of the pdisks of the declustered array. This is the internal IBM Storage Scale RAID metadata for the recovery group. The number of these VCD spares is set during recovery group creation and should not be changed.

replace threshold
Is the number of pdisks that must fail before the declustered array reports that it needs service and the pdisks are marked for required replacement.
bit error rate (BER)
If the bit error rate of a disk within a DA is predicted to be beyond an expected threshold, that pdisk will be marked as failing.
enable
Means that the bit error rate enforcement is active for this declustered array. This is the default setting for declustered arrays that support bit error rate enforcement.
disable
Means that the bit error rate enforcement is not active for this declustered array. This setting indicates that bit error rate enforcement has been temporarily disabled.
N/A
Means that bit error rate enforcement is not supported on this declustered array. Bit error rate enforcement is permanently disabled.
free space
Is the amount of raw space in this declustered array that is unused and is available for creating vdisks. The pdisk spare space for rebuild has already been removed from this number. The size of the vdisks that can be created using the raw free space depends on the redundancy requirements of the vdisk RAID code: a 4WayReplicated vdisk of size N uses 4N raw space; an 8+3P vdisk of size N uses 1.375N raw space.
scrub duration
Is the length of time in days over which the scrubbing of all vdisks in the declustered array will complete.
background activity
task
Is the task that is being performed on the declustered array.
inactive
Means that there are no vdisks defined or the declustered array is not currently available.
scrub
Means that vdisks are undergoing routine data integrity maintenance.
rebuild-critical
Means that vdisk tracks with no remaining redundancy are being rebuilt.
rebuild-1r
Means that vdisk tracks with one remaining redundancy are being rebuilt.
rebuild-2r
Means that vdisk tracks with two remaining redundancies are being rebuilt.
progress
Is the completion percentage of the current task.
priority
Is the priority given the current task. Critical rebuilds are given high priority; all other tasks have low priority.

Vdisk section:

vdisk
Is the name of the vdisk.
RAID code
Is the RAID code for this vdisk.
declustered array
Is the declustered array in which this vdisk is defined.
vdisk size
Is the usable size this vdisk.
block size
Is the block (track) size for this vdisk.
checksum granularity
Is the buffer granularity at which IBM Storage Scale RAID checksums are maintained for this vdisk. This value is set automatically at vdisk creation. For file system NSD vdisks, this value is 32 KiB. For log, log tip, log tip backup, and log reserved vdisks on Power Systems servers, this value is 4096.
state
Is the maintenance task that is being performed on this vdisk.
ok
Means that the vdisk is being scrubbed or is waiting to be scrubbed. Only one vdisk in a DA is scrubbed at a time.
1/3-deg
Means that tracks with one fault from three redundancies are being rebuilt.
2/3-deg
Means that tracks with two faults from three redundancies are being rebuilt.
1/2-deg
Means that tracks with one fault from two redundancies are being rebuilt.
critical
Means that tracks with no remaining redundancy are being rebuilt.
inactive
Means that the declustered array that is associated with this vdisk is inactive.
remarks
Indicates the special vdisk type. Only those vdisks with a dedicated function within the recovery group are indicated here: the log, log tip, log tip backup, and log reserved vdisks. This field is blank for file system NSD vdisks.

Fault tolerance section:

config data
Is the type of internal IBM Storage Scale RAID recovery group metadata for which fault tolerance is being reported.
rebuild space
Indicates the space that is available for IBM Storage Scale RAID metadata relocation.
declustered array
Is the name of the declustered array.
VCD spares
Is the number of VCD spares that are defined for the declustered array.
actual rebuild spare space
Is the number of pdisks that are currently eligible to hold VCD spares.
remarks
Indicates the effect or limit the VCD spares have on fault tolerance.
config data
Is the type of internal IBM Storage Scale RAID recovery group metadata for which fault tolerance is being reported.
rg descriptor
Indicates the recovery group declustered array and pdisk definitions.
system index
Indicates the vdisk RAID stripe partition definitions.
max disk group fault tolerance
Shows the theoretical maximum fault tolerance for the config data.
actual disk group fault tolerance
Shows the current actual fault tolerance for the config data, given the current state of the recovery group, its pdisks, and its number, type, and size of vdisks.
remarks
Indicates whether the actual fault tolerance is limiting other fault tolerances or is limited by other fault tolerances.
vdisk
Is the name of the vdisk for which fault tolerance is being reported.
max disk group fault tolerance
Shows the theoretical maximum fault tolerance for the vdisk.
actual disk group fault tolerance
Shows the current actual fault tolerance for the vdisk, given the current state of the recovery group, its pdisks, and its number, type, and size of vdisks.
remarks
Indicates why the actual fault tolerance might be different from the maximum fault tolerance.

Server section:

active recovery group server
Is the currently-active recovery group server.
servers
Is the server pair for the recovery group. The intended primary server is listed first, followed by the intended backup server.

4. Output values for mmlsrecoverygroup RecoveryGroupName -L --pdisk

Descriptions of the output values for the mmlsrecoverygroup RecoveryGroupName -L --pdisk command follow, as displayed by row, from top to bottom and left to right.

Recovery group section:

recovery group
Is the name of the recovery group.
declustered arrays
Is the number of declustered arrays in this recovery group.
vdisks
Is the number of vdisks in this recovery group.
pdisks
Is the number of pdisks in this recovery group.
format version
Is the recovery group version.

Declustered array section:

declustered array
Is the name of the declustered array.
needs service
Indicates whether this declustered array needs service. A yes value means that disks need to be replaced.
vdisks
Is the number of vdisks in this declustered array.
pdisks
Is the number of pdisks in this declustered array.
spares
The first number of the pair is the amount of spare space that is reserved for rebuilding, expressed as an equivalent number of pdisks. This spare space is allocated equally among all of the pdisks in the declustered array.

The second number of the pair is the number of VCD replicas that are maintained across all of the pdisks of the declustered array. This is the internal IBM Storage Scale RAID metadata for the recovery group. The number of these VCD spares is set during recovery group creation and should not be changed.

replace threshold
Is the number of pdisks that must fail before the declustered array reports that it needs service and the pdisks are marked for required replacement.
free space
Is the amount of raw space in this declustered array that is unused and is available for creating vdisks. The pdisk spare space for rebuilding has already been removed from this number. The size of the vdisks that can be created using the raw free space depends on the redundancy requirements of the vdisk RAID code: a 4WayReplicated vdisk of size N uses 4N raw space; an 8+3P vdisk of size N uses 1.375N raw space.
scrub duration
Is the length of time in days over which the scrubbing of all vdisks in the declustered array will complete.
background activity
task
Is the task that is being performed on the declustered array.
inactive
Means that there are no vdisks defined or the declustered array is not currently available.
scrub
Means that vdisks are undergoing routine data integrity maintenance.
rebuild-critical
Means that vdisk tracks with no remaining redundancy are being rebuilt.
rebuild-1r
Means that vdisk tracks with only one remaining redundancy are being rebuilt.
rebuild-2r
Means that vdisk tracks with two remaining redundancies are being rebuilt.
progress
Is the completion percentage of the current task.
priority
Is the priority given the current task. Critical rebuilds are given high priority; all other tasks have low priority.

Pdisk section:

pdisk
Is the name of the pdisk.
n. active, total paths
Indicates the number of active (in-use) and total block device paths to the pdisk. The total paths include the paths that are available on the standby server.
declustered array
Is the name of the declustered array to which the pdisk belongs.
free space
Is the amount of raw free space that is available on the pdisk.

Note: A pdisk that has been taken out of service and completely drained of data will show its entire capacity as free space, even though that capacity is not available for use.

user condition
Is the condition of the pdisk from the system administrator's persepective. A "normal" condition requires no attention, while "replaceable" means that the pdisk might be, but is not necessarily required to be, physically replaced.
state, remarks
Is the state of the pdisk. For a description of pdisk states, see the topic Pdisk states in IBM Storage Scale RAID: Administration.

Server section:

active recovery group server
Is the currently-active recovery group server.
servers
Is the server pair for the recovery group. The intended primary server is listed first, followed by the intended backup server.

Parameters

RecoveryGroupName
Specifies the recovery group for which the information is being requested. If no other parameters are specified, the command displays only the information that can be found in the GPFS cluster configuration data.
-L
Displays more detailed runtime information for the specified recovery group.
--pdisk
Indicates that pdisk information is to be listed for the specified recovery group.

Exit status

0
Successful completion.
nonzero
A failure has occurred.

Security

You must have root authority to run the mmlsrecoverygroup command.

The node on which the command is issued must be able to execute remote shell commands on any other node in the cluster without the use of a password and without producing any extraneous messages. For additional details, see the following IBM Storage Scale RAID: Administration topic: Requirements for administering IBM Storage Scale RAID.

Examples

  1. The following command example shows how to list all the recovery groups in the GPFS cluster:
    mmlsrecoverygroup
    The system displays output similar to the following:
     
                         declustered
                         arrays with
     recovery group        vdisks     vdisks  servers
     ------------------  -----------  ------  -------
     BB1RGL                        4       8  c45f01n01-ib0.gpfs.net,c45f01n02-ib0.gpfs.net
     BB1RGR                        3       7  c45f01n02-ib0.gpfs.net,c45f01n01-ib0.gpfs.net
    
  2. The following command example shows how to list the basic non-runtime information for recovery group 000DE37BOT:
    mmlsrecoverygroup 000DE37BOT
    
    The system displays output similar to the following:
    
                         declustered
                         arrays with
     recovery group        vdisks     vdisks  servers
     ------------------  -----------  ------  -------
     BB1RGL                        4       8  c45f01n01-ib0.gpfs.net,c45f01n02-ib0.gpfs.net
    
     declustered array
         with vdisks     vdisks
     ------------------  ------
     DA1                      3
     DA2                      3
     NVR                      1
     SSD                      1
    
                                             declustered
     vdisk               RAID code              array     remarks
     ------------------  ------------------  -----------  -------
     BB1RGLDATA1         8+3p                DA1
     BB1RGLDATA2         8+3p                DA2
     BB1RGLMETA1         4WayReplication     DA1
     BB1RGLMETA2         4WayReplication     DA2
     lhome_BB1RGL        4WayReplication     DA1          log
     ltbackup_BB1RGL     Unreplicated        SSD
     ltip_BB1RGL         2WayReplication     NVR
     reserved1_BB1RGL    4WayReplication     DA2
    
  3. The following command example shows how to display the runtime status of recovery group BB01L:
    mmlsrecoverygroup BB01L -L
    
    The system displays output similar to the following:
    
                        declustered                     current       allowable
     recovery group       arrays     vdisks  pdisks  format version format version
     -----------------  -----------  ------  ------  -------------- --------------
     BB01L                        1       5      49  5.1.2.0        5.1.2.0
    
     declustered   needs                            replace                               scrub       background activity
        array     service  vdisks  pdisks  spares  threshold  BER      trim  free space  duration  task   progress  priority
     -----------  -------  ------  ------  ------  ---------  -------  ----  ----------  --------  -------------------------
     DA1          no            5      49    2,27          2  enable   no       339 TiB   14 days  inactive    56%  low
    
                                             declustered                           checksum
     vdisk               RAID code              array     vdisk size  block size  granularity  state remarks
     ------------------  ------------------  -----------  ----------  ----------  -----------  ----- -------
     RG001LOGHOME        4WayReplication     DA1             176 GiB      2 MiB      4096      ok    log
     RG001VS002          3WayReplication     DA1            7148 GiB      2 MiB     32 KiB     ok
     RG001VS001          8+2p                DA1              50 TiB     16 MiB     32 KiB     ok
     RG001VS006          8+3p                DA1              63 GiB     16 MiB     32 KiB     ok
     RG001VS005          8+3p                DA1              63 GiB     16 MiB     32 KiB     ok
    
     config data         declustered array   spare space    remarks
     ------------------  ------------------  -------------  -------
     rebuild space       DA1                 44 pdisk
    
     config data         disk group fault tolerance         remarks
     ------------------  ---------------------------------  -------
     rg descriptor       4 pdisk                            limiting fault tolerance
     system index        4 pdisk                            limited by rg descriptor
    
     vdisk               disk group fault tolerance         remarks
     ------------------  ---------------------------------  -------
     RG001LOGHOME        3 pdisk
     RG001VS002          2 pdisk
     RG001VS001          2 pdisk
     RG001VS006          3 pdisk
     RG001VS005          3 pdisk
    
     active recovery group server                     servers
     -----------------------------------------------  -------
     c145f03n03.gpfs.net                              c145f03n03.gpfs.net,c145f03n04.gpfs.net
     rebuild space       DA1                 44 pdisk                                             
    
     config data         disk group fault tolerance         remarks
     ------------------  ---------------------------------  -------
     rg descriptor       4 pdisk                            limiting fault tolerance
     system index        4 pdisk                            limited by rg descriptor
    
     vdisk               disk group fault tolerance         remarks
     ------------------  ---------------------------------  -------
     RG001LOGHOME        3 pdisk                                              
     RG001VS002          2 pdisk                                              
     RG001VS001          2 pdisk                                              
     RG001VS006          3 pdisk                                              
     RG001VS005          3 pdisk                                              
    
     active recovery group server                     servers
     -----------------------------------------------  -------
     c145f03n03.gpfs.net                              c145f03n03.gpfs.net,c145f03n04.gpfs.net 
    For more information, see the following IBM Storage Scale RAID: Administration topic: Determining pdisk-group fault-tolerance.
  4. The following example shows how to include pdisk information for BB01L:
    mmlsrecoverygroup BB01L -L --pdisk
    The system displays output similar to the following:
    
                        declustered                     current       allowable
     recovery group       arrays     vdisks  pdisks  format version format version
     -----------------  -----------  ------  ------  -------------- --------------
     BB01L                        1       3      52  5.1.2.0        5.1.2.0       
    
     declustered   needs                            replace                    scrub             background activity
        array     service  vdisks  pdisks  spares  threshold  BER      trim  free space   duration  task   progress  priority
     -----------  -------  ------  ------  ------  ---------  ----     -----  ----------  --------  -------------------------
     DA1          no            3      52    2,27          2  enable   no       360 TiB     0 day   scrub       45%  low   
    
                         n. active,   declustered              state,
     pdisk               total paths     array     free space  remarks
     -----------------   -----------  -----------  ----------  -------
     e1s007                2,  4      DA1          7472 GiB    ok   
     e1s008                2,  4      DA1          7464 GiB    ok   
     e1s009                2,  4      DA1          7472 GiB    ok   
     e1s010                2,  4      DA1          7456 GiB    ok   
    .
    .
    .
    active recovery group server                     servers
     -----------------------------------------------  -------
     c145f03n04.gpfs.net                              c145f03n04.gpfs.net,c145f03n03.gpfs.net

See also

Location

/usr/lpp/mmfs/bin