After cluster enters healthy state mgr should re-check stray daemons, a lot of activities are on hold while cluster is in warning state. In the event it does not disappear after cluster is healthy than mgr restart should help. Kind regards, Nino On Fri, Jun 16, 2023 at 10:24 PM Nicola Mori <mori@fi.infn.it> wrote:
The osd daemon finally disappeared without further intervention. I guess I should have had more patience and wait the purge process to finish. Thanks to everybody who helped.
Nicola
Il 15 giugno 2023 15:02:16 CEST, Nicola Mori <mori@fi.infn.it> ha scritto:
I have been able to (sort-of) fix the problem by removing the problematic
OSD, zapping the disk and starting a new OSD. The new OSD is backfilling, but now the problem is that some parts of Ceph are still waiting for the OSD removal, and the OSD (despite not running anymore on the host) is seen as a stray daemon:
# ceph health detail HEALTH_WARN 1 stray daemon(s) not managed by cephadm [WRN] CEPHADM_STRAY_DAEMON: 1 stray daemon(s) not managed by cephadm stray daemon osd.34 on host balin not managed by cephadm
# ceph osd tree | grep 34 34 hdd 1.81940 osd.34 down 0 1.00000
# ceph orch osd rm status OSD HOST STATE PGS REPLACE FORCE ZAP DRAIN
STARTED AT
34 balin done, waiting for purge 0 False True False
# ceph orch osd rm stop osd.34 Unable to find OSD in the queue: osd.34
I restarted the mgrs but it didn't help. Any suggestion?
Nicola
-- Nicola Mori, Ph.D. INFN sezione di Firenze Via Bruno Rossi 1, 50019 Sesto F.no (Italy) +390554572660 mori@fi.infn.it _______________________________________________ ceph-users mailing list -- ceph-users@ceph.io To unsubscribe send an email to ceph-users-leave@ceph.io