Hi, On 17/03/2021 22:26, Andrew Walker-Brown wrote:
How have folks implemented getting email or snmp alerts out of Ceph? Getting things like osd/pool nearly full or osd/daemon failures etc. I'm afraid we used our existing Nagios infrastructure for checking HEALTH status, and have a script that runs daily to report on failed OSDs.
Our existing metrics infrastructure is collectd/graphite/grafana so we have dashboards and so on, but as far as I'm aware the Octopus dashboard only supports prometheus, so we're a bit stuck there :-( Regards, Matthew -- The Wellcome Sanger Institute is operated by Genome Research Limited, a charity registered in England with number 1021457 and a company registered in England with number 2742969, whose registered office is 215 Euston Road, London, NW1 2BE.