== replay-vbr test 12a: lost data due to missed REMOTE client during replay ========================================================== 19:23:49 (1713482629) Starting client: oleg240-client.virtnet: -o user_xattr,flock oleg240-server@tcp:/lustre /mnt/lustre2 mdt.lustre-MDT0000.commit_on_sharing=0 UUID 1K-blocks Used Available Use% Mounted on lustre-MDT0000_UUID 1414116 1924 1285764 1% /mnt/lustre[MDT:0] lustre-MDT0001_UUID 1414116 1704 1285984 1% /mnt/lustre[MDT:1] lustre-OST0000_UUID 3833116 1524 3605496 1% /mnt/lustre[OST:0] lustre-OST0001_UUID 3833116 1524 3605496 1% /mnt/lustre[OST:1] filesystem_summary: 7666232 3048 7210992 1% /mnt/lustre total: 25 open/close in 0.17 seconds: 150.80 ops/second total: 25 open/close in 0.17 seconds: 149.82 ops/second Stopping client oleg240-client.virtnet /mnt/lustre2 (opts:) Failing mds1 on oleg240-server Stopping /mnt/lustre-mds1 (opts:) on oleg240-server 19:23:55 (1713482635) shut down Failover mds1 to oleg240-server mount facets: mds1 Starting mds1: -o localrecov /dev/mapper/mds1_flakey /mnt/lustre-mds1 oleg240-server: oleg240-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all pdsh@oleg240-client: oleg240-server: ssh exited with exit code 1 Started lustre-MDT0000 19:24:10 (1713482650) targets are mounted 19:24:10 (1713482650) facet_failover done - unlinked 0 (time 1713482719 ; total 0 ; last 0) total: 25 unlinks in 1 seconds: 25.000000 unlinks/second - unlinked 0 (time 1713482720 ; total 0 ; last 0) total: 25 unlinks in 0 seconds: inf unlinks/second Can't lstat /mnt/lustre/d12a.replay-vbr/f12a.replay-vbr: No such file or directory