Hi Janek, On Tue, Aug 6, 2019 at 11:25 AM Janek Bevendorff <janek.bevendorff@uni-weimar.de> wrote:
Here are tracker tickets to resolve the issues you encountered:
https://tracker.ceph.com/issues/41140 https://tracker.ceph.com/issues/41141
The fix has been merged into master and will be backported soon. I've also done testing in a large cluster to confirm the issue you found. Using multiple processes to create files as fast as possible in a single client reliably reproduced the issue. The MDS cannot recall capabilities fast enough when the internal upkeep thread ran every 5 seconds. Moving the cache trimming and capability recall to a separate thread running every second resolved the issue. -- Patrick Donnelly, Ph.D. He / Him / His Senior Software Engineer Red Hat Sunnyvale, CA GPG: 19F28A586F808C2402351B93C3301A3E258DD79D