Hi, I got many critical alerts in ceph dashboard. Meanwhile the cluster shows health ok status. See attached screenshot for detail. My questions are, are they real alerts? How to get rid of them? Thanks Ben
Hi Ben, It looks like you forgot to attach the screenshots. Regards, Nizam On Wed, Jun 21, 2023, 12:23 Ben <ruidong.gao@gmail.com> wrote:
Hi,
I got many critical alerts in ceph dashboard. Meanwhile the cluster shows health ok status.
See attached screenshot for detail. My questions are, are they real alerts? How to get rid of them?
Thanks Ben _______________________________________________ ceph-users mailing list -- ceph-users@ceph.io To unsubscribe send an email to ceph-users-leave@ceph.io
Hi Ben, also if some alerts are noisy, we have option in dashboard to silence those alerts. Also, can you provide the list of critical alerts that you see? On Wed, 21 Jun 2023 at 12:48, Nizamudeen A <nia@redhat.com> wrote:
Hi Ben,
It looks like you forgot to attach the screenshots.
Regards, Nizam
On Wed, Jun 21, 2023, 12:23 Ben <ruidong.gao@gmail.com> wrote:
Hi,
I got many critical alerts in ceph dashboard. Meanwhile the cluster shows health ok status.
See attached screenshot for detail. My questions are, are they real alerts? How to get rid of them?
Thanks Ben _______________________________________________ ceph-users mailing list -- ceph-users@ceph.io To unsubscribe send an email to ceph-users-leave@ceph.io
_______________________________________________ ceph-users mailing list -- ceph-users@ceph.io To unsubscribe send an email to ceph-users-leave@ceph.io
attached screenshot was filtered out. Here it is partially: name Severity Group Duration Summary CephadmDaemonFailed critical cephadm 30 seconds A ceph daemon manged by cephadm is down CephadmPaused warning cephadm 1 minute Orchestration tasks via cephadm are PAUSED CephadmUpgradeFailed critical cephadm 30 seconds Ceph version upgrade has failed CephDaemonCrash critical generic 1 minute One or more Ceph daemons have crashed, and are pending acknowledgement CephDeviceFailurePredicted waming osd 1 minute Device(s) predicted to fail soon CephDeviceFailurePrediction TooHigh critical osd 1 minute Too many devices are predicted to fail, unable to resolve CephDeviceFailureRelocationincomplete warning osd 1 minute Device failure is predicted, but unable to relocate data CephFilesystemDamaged critical mds 1 minute CephFS filesystem is damaged. CephFilesystemDegraded critical mds 1 minute CephFS filesystem is degraded CephFilesystemFailureNoStandby critical mds 1 minute MDS daemon failed, no further standby available Meanwhile the cluster status is green ok. What should we do for this? Thanks, Ben
participants (3)
-
Ankush Behl
-
Ben
-
Nizamudeen A