Hi, Am 21.11.25 um 5:11 AM schrieb gagan tiwari:
We are getting an increasing number of RX errors on the ports on some of the nodes in the cluster including mon , mds and osd nodes
In an attempt to fix it , we plan to replace SFPs / cables on these nodes.
I think I will need to temp remove those nodes one by one from the cluster and replace the network cables / SFPs and add them back to cluster
So, plz let me know the best / safest way to proceed.
With the cephadm orchestrator you can use ceph orch host maintenance enter HOSTNAME do your stuff ceph orch host maintenance exit HOSTNAME for each host in a row. Without the orchestrator you can use ceph osd set-group noout HOSTNAME do your stuff ceph osd unset-group noout HOSTNAME for each host. But never two hosts at the same time. Regards -- Robert Sander Linux Consultant Heinlein Consulting GmbH Schwedter Str. 8/9b, 10119 Berlin https://www.heinlein-support.de Tel: +49 30 405051 - 0 Fax: +49 30 405051 - 19 Amtsgericht Berlin-Charlottenburg - HRB 220009 B Geschäftsführer: Peer Heinlein - Sitz: Berlin