メインコンテンツへスキップ

SwitchIfOutErrorsWarn_Alert - AutoSupport メッセージ

Views:
Visibility:
Public
Votes:
0
Category:
fabric-interconnect-and-management-switches
Specialty:
hw
Last Updated:

環境

  • ONTAP 9
  • クラスター ネットワーク スイッチ
  • ヘルス モニター プロセス cshm のコールホーム:SwitchIfOutErrorsWarn_Alert

イベント サマリ

このメッセージは、定期的な健康状態監視中にエラーが検出された場合に発生します。

  • システム健全性モニターは、サブシステムの監視中に検出された潜在的な問題に対してアラートを生成します。
  • アラートには、考えられる原因に関する情報と、問題を解決するための推奨される対策が含まれています。
  • スイッチインターフェイス「Switch Name/Slot: 0 Port: 4 10G - Level」の送信パケットエラー率が警告しきい値を超えています。
  • 送信パケットエラーは、スイッチインターフェイスがクラスタ インターコネクトを介してトラフィックを送信中にエラーが発生していることを示しています。
  • クラスタ インターコネクトの劣化は、クラスターの不安定化や、場合によってはシステム障害を引き起こす可能性があります。

検証

AutoSupportメッセージ

HA Group Notification from Node Name (Health Monitor process cshm: SwitchIfOutErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level]) ALERT

イベントログ

event log show -severity * -message-name callhome*

[Node Name Name: mgwd: callhome.hm.alert.major:alert]: Call home for Health Monitor process cshm: SwitchIfOutErrorsWarn_Alert[Node Name/Slot: 0 Port: 4 10G - Level].
Command Line

system health alert show -node <node name> -monitor cluster-switch -alert-id SwitchIfOutErrorsWarn_Alert

Node: NetApp-a Monitor: cluster-switch Class of Alert: SwitchIfOutErrorsWarn_Alert Severity of Alert: Major Probable Cause: Threshold_crossed Probable Cause Description: The percentage of outbound packet errors of switch interface "$(cluster_switch_analytics.unique-name)" is above the warning threshold. Possible Effect: Communication between nodes in the cluster might be degraded. Corrective Actions: 1) Migrate any cluster LIF that uses this connection to another port connected to a cluster switch. For example, if cluster LIF "clus1" is on port e0a and the other LIF is on e0b, run the following command to move "clus1" to e0b: "network interface migrate -vserver vs1 -lif clus1 -sourcenode node1 -destnode node1 -dest-port e0b" 2) Replace the network cable with a known-good cable. If errors are corrected, stop. No further action is required. Otherwise, continue to Step 3. 3) Move the network cable to another port on the node (if available). Migrate the cluster LIF to the new port. If errors are corrected, contact technical support to troubleshoot the original node port. Otherwise, continue to Step 4. 4) Move the network cable to another available cluster switch port. Migrate the cluster LIF back to the original port.

解決策

  1. SwitchIfOutErrorsWarn_Alertを報告しているポートのリンクトラブルシューティングを実行します。
  2. ケーブルとSFP/トランシーバーを抜き差ししてみてください。
  3. パッチパネルや中間接続がないか確認し、可能であればバイパスしてください。
  4. ケーブルおよび/またはSFPを、正常に動作することが確認されている部品と交換してください。
  5. ノード側 - 影響を受けているクラスタLIFを正常なクラスタポートに移行し、接続を代替ノードクラスタポートに移動します(利用可能な場合)。
    • エラーが停止した場合は、元のノードポートを調査してください。(Contact NetApp テクニカルサポート またはログインして NetApp Support Site で ケースを作成してください。詳細については、この記事を参照してください。)
  6. スイッチ側 - 接続を正常なスイッチ ポートに移動します。
    • エラーが止まったら、元のスイッチ ポートを調査し、スイッチログを収集して、スイッチサポートに連絡してください。 
      • Broadcom製スイッチについては、Broadcom にお問い合わせください
      • Cisco製スイッチについては、Cisco にお問い合わせください
      • NVIDIA製スイッチについては、NVIDIA にお問い合わせください

    追加情報

     

    NetApp provides no representations or warranties regarding the accuracy or reliability or serviceability of any information or recommendations provided in this publication or with respect to any results that may be obtained by the use of the information or observance of any recommendations provided herein. The information in this document is distributed AS IS and the use of this information or the implementation of any recommendations or techniques herein is a customer's responsibility and depends on the customer's ability to evaluate and integrate them into the customer's operational environment. This document and the information contained herein may be used solely in connection with the NetApp products discussed in this document.
    • この記事は役に立ちましたか?