[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-8.fc42 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 650369426 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2400.000 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE2439 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE22D5 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 002295 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE2349 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE23D9 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2411 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe22d5-0xbffe2348] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe22d4] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe2349-0xbffe23d8] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe23d9-0xbffe2410] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2411-0xbffe2438] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524580K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001011] APIC: Switch to symmetric I/O mode setup [ 0.003299] x2apic enabled [ 0.004000] Switched APIC routing to physical x2apic. [ 0.004000] kvm-guest: setup PV IPIs [ 0.004000] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.004000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 0.004000] Calibrating delay loop (skipped) preset value.. 4800.00 BogoMIPS (lpj=2400000) [ 0.004012] pid_max: default: 32768 minimum: 301 [ 0.005131] LSM: Security Framework initializing [ 0.006055] Yama: becoming mindful. [ 0.007034] SELinux: Initializing. [ 0.008077] *** VALIDATE selinux *** [ 0.016333] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.021456] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.022158] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.023138] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.024109] *** VALIDATE tmpfs *** [ 0.026144] *** VALIDATE proc *** [ 0.027245] *** VALIDATE cgroup *** [ 0.028018] *** VALIDATE cgroup2 *** [ 0.029208] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.031120] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.032010] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.033031] Spectre V2 : User space: Vulnerable [ 0.034008] Speculative Store Bypass: Vulnerable [ 0.037000] debug: unmapping init [mem 0xffffffffb2e59000-0xffffffffb2e60fff] [ 0.038892] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.039723] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.040024] ... version: 2 [ 0.041012] ... bit width: 48 [ 0.042010] ... generic registers: 4 [ 0.043014] ... value mask: 0000ffffffffffff [ 0.044012] ... max period: 00007fffffffffff [ 0.045011] ... fixed-purpose events: 3 [ 0.046010] ... event mask: 000000070000000f [ 0.048219] rcu: Hierarchical SRCU implementation. [ 0.050476] smp: Bringing up secondary CPUs ... [ 0.051618] x86: Booting SMP configuration: [ 0.052030] .... node #0, CPUs: #1 #2 #3 [ 0.056492] smp: Brought up 1 node, 4 CPUs [ 0.058012] smpboot: Max logical packages: 1 [ 0.059012] smpboot: Total of 4 processors activated (19200.00 BogoMIPS) [ 0.138807] node 0 deferred pages initialised in 77ms [ 0.141011] devtmpfs: initialized [ 0.142228] x86/mm: Memory block size: 128MB [ 0.144711] gcov: version magic: 0x41383552 [ 0.146097] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.147094] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.148250] pinctrl core: initialized pinctrl subsystem [ 0.149180] [ 0.149775] ************************************************************* [ 0.150012] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.151011] ** ** [ 0.152010] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.153016] ** ** [ 0.154014] ** This means that this kernel is built to expose internal ** [ 0.155012] ** IOMMU data structures, which may compromise security on ** [ 0.156015] ** your system. ** [ 0.157015] ** ** [ 0.158013] ** If you see this message and you are not debugging the ** [ 0.159020] ** kernel, report this immediately to your vendor! ** [ 0.160011] ** ** [ 0.161018] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.162015] ************************************************************* [ 0.163674] NET: Registered protocol family 16 [ 0.164420] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.165117] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.166068] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.167454] cpuidle: using governor menu [ 0.185897] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.188545] PCI: Using configuration type 1 for base access [ 0.191131] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.200057] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.201020] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.203046] cryptd: max_cpu_qlen set to 1000 [ 0.205740] ACPI: Added _OSI(Module Device) [ 0.207012] ACPI: Added _OSI(Processor Device) [ 0.209013] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.211012] ACPI: Added _OSI(Processor Aggregator Device) [ 0.216370] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.226339] ACPI: Interpreter enabled [ 0.228065] ACPI: PM: (supports S0 S3 S4 S5) [ 0.230011] ACPI: Using IOAPIC for interrupt routing [ 0.232115] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.235483] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.247663] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.250045] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.254021] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.257086] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.263612] acpiphp: Slot [2] registered [ 0.265250] acpiphp: Slot [5] registered [ 0.267251] acpiphp: Slot [6] registered [ 0.269261] acpiphp: Slot [7] registered [ 0.271147] acpiphp: Slot [8] registered [ 0.273128] acpiphp: Slot [9] registered [ 0.275136] acpiphp: Slot [10] registered [ 0.276163] acpiphp: Slot [3] registered [ 0.278095] acpiphp: Slot [4] registered [ 0.280197] acpiphp: Slot [11] registered [ 0.281112] acpiphp: Slot [12] registered [ 0.282115] acpiphp: Slot [13] registered [ 0.285428] acpiphp: Slot [14] registered [ 0.288135] acpiphp: Slot [15] registered [ 0.290107] acpiphp: Slot [16] registered [ 0.292287] acpiphp: Slot [17] registered [ 0.294106] acpiphp: Slot [18] registered [ 0.296114] acpiphp: Slot [19] registered [ 0.298129] acpiphp: Slot [20] registered [ 0.300127] acpiphp: Slot [21] registered [ 0.302114] acpiphp: Slot [22] registered [ 0.304119] acpiphp: Slot [23] registered [ 0.306118] acpiphp: Slot [24] registered [ 0.308112] acpiphp: Slot [25] registered [ 0.311139] acpiphp: Slot [26] registered [ 0.313138] acpiphp: Slot [27] registered [ 0.315112] acpiphp: Slot [28] registered [ 0.316116] acpiphp: Slot [29] registered [ 0.318119] acpiphp: Slot [30] registered [ 0.320121] acpiphp: Slot [31] registered [ 0.322098] PCI host bridge to bus 0000:00 [ 0.324020] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.327026] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.330024] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.333023] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.336025] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.340034] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.342176] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.346000] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.349577] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.359658] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.365013] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.369020] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.371019] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.374020] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.377711] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.380813] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.384043] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.389101] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.395015] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.412008] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.418017] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.425443] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.433013] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.441016] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.462017] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.474242] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.481025] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.489023] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.503019] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.515052] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.522013] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.530018] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.551016] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.565398] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.572014] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.582019] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.601015] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.612387] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.620014] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.628013] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.645016] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.678814] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.689016] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.702018] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.719014] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.731878] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.735383] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.738503] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.741415] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.744227] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.751079] iommu: Default domain type: Passthrough [ 0.753348] SCSI subsystem initialized [ 0.756150] ACPI: bus type USB registered [ 0.758123] usbcore: registered new interface driver usbfs [ 0.761114] usbcore: registered new interface driver hub [ 0.764114] usbcore: registered new device driver usb [ 0.766330] pps_core: LinuxPPS API ver. 1 registered [ 0.768013] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.773063] PTP clock support registered [ 0.777097] EDAC MC: Ver: 3.0.0 [ 0.779108] PCI: Using ACPI for IRQ routing [ 0.780675] NetLabel: Initializing [ 0.782014] NetLabel: domain hash size = 128 [ 0.784011] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.787100] NetLabel: unlabeled traffic allowed by default [ 0.790130] vgaarb: loaded [ 0.791282] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.795018] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.801857] clocksource: Switched to clocksource kvm-clock [ 0.928577] VFS: Disk quotas dquot_6.6.0 [ 0.931701] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 0.936415] *** VALIDATE ramfs *** [ 0.938197] *** VALIDATE hugetlbfs *** [ 0.942206] pnp: PnP ACPI init [ 0.944881] pnp: PnP ACPI: found 6 devices [ 0.966094] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 0.970666] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 0.973020] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 0.975157] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 0.977485] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 0.979643] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 0.983548] NET: Registered protocol family 2 [ 0.987239] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 0.993177] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 0.996917] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 1.002790] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 1.009615] TCP: Hash tables configured (established 65536 bind 65536) [ 1.013624] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 1.018492] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.022426] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.026725] NET: Registered protocol family 1 [ 1.030489] RPC: Registered named UNIX socket transport module. [ 1.034062] RPC: Registered udp transport module. [ 1.036170] RPC: Registered tcp transport module. [ 1.038268] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 1.041181] NET: Registered protocol family 44 [ 1.043150] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 1.045520] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 1.047869] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 1.050570] PCI: CLS 0 bytes, default 64 [ 1.052518] Unpacking initramfs... [ 2.581181] debug: unmapping init [mem 0xffff9c4ffcc54000-0xffff9c4ffffbffff] [ 2.585514] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 2.588301] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 2.591227] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 3.082613] Initialise system trusted keyrings [ 3.084480] Key type blacklist registered [ 3.086338] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 3.096447] zbud: loaded [ 3.100262] *** VALIDATE nfs *** [ 3.101881] *** VALIDATE nfs4 *** [ 3.104254] pstore: using deflate compression [ 3.108519] Platform Keyring initialized [ 3.215382] NET: Registered protocol family 38 [ 3.217182] Key type asymmetric registered [ 3.219130] Asymmetric key parser 'x509' registered [ 3.221204] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 3.224223] io scheduler mq-deadline registered [ 3.226327] io scheduler kyber registered [ 3.227977] io scheduler bfq registered [ 3.230261] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 3.233983] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 3.237531] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 3.241177] ACPI: Power Button [PWRF] [ 3.247143] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 3.255864] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 3.275187] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 3.285437] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 3.303767] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 3.332895] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 3.361874] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 3.367189] Non-volatile memory driver v1.3 [ 3.368449] Linux agpgart interface v0.103 [ 3.403371] virtio_blk virtio1: [vda] 134016 512-byte logical blocks (68.6 MB/65.4 MiB) [ 3.406286] vda: detected capacity change from 0 to 68616192 [ 3.421423] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 3.423782] vdb: detected capacity change from 0 to 1073741824 [ 3.439201] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.442064] vdc: detected capacity change from 0 to 2621440000 [ 3.458422] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.461240] vdd: detected capacity change from 0 to 2621440000 [ 3.476794] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.479393] vde: detected capacity change from 0 to 4294967296 [ 3.493802] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.495983] vdf: detected capacity change from 0 to 4294967296 [ 3.502852] libphy: Fixed MDIO Bus: probed [ 3.512516] usbcore: registered new interface driver usbserial_generic [ 3.515206] usbserial: USB Serial support registered for generic [ 3.517642] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 3.523385] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 3.525109] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 3.527850] mousedev: PS/2 mouse device common for all mice [ 3.531665] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 3.532200] rtc_cmos 00:05: RTC can wake from S4 [ 3.540321] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 3.540623] rtc_cmos 00:05: registered as rtc0 [ 3.545574] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 3.548825] intel_pstate: CPU model not supported [ 3.549785] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 3.555533] hid: raw HID events driver (C) Jiri Kosina [ 3.557744] usbcore: registered new interface driver usbhid [ 3.559573] usbhid: USB HID core driver [ 3.561458] drop_monitor: Initializing network drop monitor service [ 3.563620] Initializing XFRM netlink socket [ 3.565486] NET: Registered protocol family 10 [ 3.568358] Segment Routing with IPv6 [ 3.569713] NET: Registered protocol family 17 [ 3.571779] mpls_gso: MPLS GSO support [ 3.576861] RAS: Correctable Errors collector initialized. [ 3.580443] AVX version of gcm_enc/dec engaged. [ 3.583296] AES CTR mode by8 optimization enabled [ 3.674309] sched_clock: Marking stable (3674287175, 0)->(4665808288, -991521113) [ 3.678557] registered taskstats version 1 [ 3.680407] Loading compiled-in X.509 certificates [ 3.700844] zswap: loaded using pool lzo/zbud [ 3.753064] Key type big_key registered [ 3.767742] Key type encrypted registered [ 3.769780] ima: No TPM chip found, activating TPM-bypass! [ 3.772812] ima: Allocated hash algorithm: sha1 [ 3.776057] ima: No architecture policies found [ 3.778713] evm: Initialising EVM extended attributes: [ 3.781224] evm: security.selinux [ 3.783837] evm: security.ima [ 3.785260] evm: security.capability [ 3.786860] evm: HMAC attrs: 0x1 [ 3.790090] rtc_cmos 00:05: setting system clock to 2026-04-30 18:51:56 UTC (1777575116) [ 3.799164] debug: unmapping init [mem 0xffffffffb3e03000-0xffffffffb3ffffff] [ 3.803631] debug: unmapping init [mem 0xffffffffb2b82000-0xffffffffb2e58fff] [ 3.813077] Write protecting the kernel read-only data: 28672k [ 3.817730] debug: unmapping init [mem 0xffffffffb1203000-0xffffffffb13fffff] [ 3.820691] debug: unmapping init [mem 0xffffffffb1b14000-0xffffffffb1bfffff] [ 3.869505] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 3.878795] systemd[1]: Detected virtualization kvm. [ 3.880292] systemd[1]: Detected architecture x86-64. [ 3.881822] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 3.906839] systemd[1]: No hostname configured. [ 3.908127] systemd[1]: Set hostname to . [ 3.909757] random: systemd: uninitialized urandom read (16 bytes read) [ 3.912588] systemd[1]: Initializing machine ID from random generator. [ 4.077675] random: systemd: uninitialized urandom read (16 bytes read) [ 4.083208] systemd[1]: Listening on Journal Socket (/dev/log). [ OK ] Listening on Journal Socket (/dev/log). [ 4.091161] random: systemd: uninitialized urandom read (16 bytes read) [ 4.097572] systemd[1]: Reached target Slices. [ OK ] Reached target Slices. [ 4.104055] systemd[1]: Reached target Swap. [ OK ] Reached target Swap. [ OK ] Reached target Initrd Root Device. [ OK ] Listening on udev Control Socket. [ OK ] Listening on Journal Socket. Starting Create list of required st…ce nodes for the current kernel... Starting Journal Service... [ OK ] Reached target Timers. [ OK ] Reached target Local File Systems. Starting Create Volatile Files and Directories... [ OK ] Started Memstrack Anylazing Service. Starting Apply Kernel Variables... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Paths. [ OK ] Listening on udev Kernel Socket. [ OK ] Reached target Sockets. Starting Setup Virtual Console... [ OK ] Reached target Local Encrypted Volumes. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Create Volatile Files and Directories. [ OK ] Started Apply Kernel Variables. [ OK ] Started Setup Virtual Console. Starting dracut cmdline hook... Starting Create Static Device Nodes in /dev... [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started Journal Service. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 4.902140] device-mapper: uevent: version 1.0.3 [ 4.904557] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. Starting dracut initqueue hook... [ 5.754823] virtio_net virtio0 ens2: renamed from eth0 [ 5.805601] random: fast init done [ 5.841860] scsi host0: ata_piix [ 5.860499] scsi host1: ata_piix [ 5.863086] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 5.869335] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 10.319924] random: crng init done [ 10.328037] random: 7 urandom warning(s) missed due to ratelimiting [ 11.548362] dracut-initqueue[589]: RTNETLINK answers: File exists Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. Mounting /sysroot... [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. [ 13.311833] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped target Timers. [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Initrd Root Device. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Sockets. [ OK ] Stopped target System Initialization. [ OK ] Stopped target Swap. [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Local File Systems. [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. Stopping udev Kernel Device Manager... [ OK ] Stopped target Slices. [ OK ] Stopped target Paths. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Closed udev Kernel Socket. [ OK ] Closed udev Control Socket. Starting Cleanup udevd DB... [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 17.074274] printk: systemd: 26 output lines suppressed due to ratelimiting [ 18.052915] SELinux: Disabled at runtime. [ 18.216452] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 18.237200] systemd[1]: Detected virtualization kvm. [ 18.241071] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 20.087409] systemd[1]: initrd-switch-root.service: Succeeded. [ 20.095815] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 20.107753] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 20.112808] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 20.120567] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 20.132568] systemd[1]: Starting Journal Service... Starting Journal Service... [ 20.146847] systemd[1]: Stopped target Switch Root. [ OK ] Stopped target Switch Root. [ OK ] Created slice User and Session Slice. [ OK ] Reached target Slices. [ OK ] Listening on RPCbind Server Activation Socket. [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. [ OK ] Created slice system-getty.slice. Starting Remount Root and Kernel File Systems... [ OK ] Stopped target Initrd File Systems. Starting Create list of required st…ce nodes for the current kernel... [ OK ] Started Forward Password Requests to Wall Directory Watch. [ OK ] Listening on udev Control Socket. Mounting Huge Pages File System... [ OK ] Listening on initctl Compatibility Named Pipe. [ OK ] Created slice system-sshd\x2dkeygen.slice. [ OK ] Created slice system-serial\x2dgetty.slice. [ OK ] Listening on udev Kernel Socket. Mounting POSIX Message Queue File System... Activating swap /dev/disk/by-label/SWAP... [ OK ] Reached target RPC Port Mapper. [ 20.629922] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS [ OK ] Listening on Process Core Dump Socket. Starting Apply Kernel Variables... Starting udev Coldplug all Devices... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Reached target Paths. Mounting Kernel Debug File System... [ OK ] Stopped target Initrd Root File System. [ OK ] Reached target rpc_pipefs.target. [ OK ] Started Journal Service. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Mounted Huge Pages File System. [ OK ] Mounted POSIX Message Queue File System. [ OK ] Activated swap /dev/disk/by-label/SWAP. [ OK ] Started Apply Kernel Variables. [ OK ] Mounted Kernel Debug File System. [ OK ] Reached target Swap. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... Starting Flush Journal to Persistent Storage... [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Reached target Local File Systems (Pre). Mounting /home/green/git/lustre-release... Mounting /mnt... Starting udev Kernel Device Manager... [ OK ] Mounted /mnt. [ 21.626490] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Coldplug all Devices. [ OK ] Started udev Kernel Device Manager. [ 22.679677] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 22.720756] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 23.270456] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 23.379281] EDAC sbridge: Ver: 1.1.2 [* ] A start job is running for Configur…-only root support (7s / no limit) [** ] A start job is running for Configur…-only root support (7s / no limit) [*** ] A start job is running for Configur…-only root support (8s / no limit) [ *** ] A start job is running for Configur…-only root support (8s / no limit) [ *** ] A start job is running for Configur…-only root support (9s / no limit)[ 29.647965] Key type dns_resolver registered [ ***] A start job is running for Configur…-only root support (9s / no limit) [ **] A start job is running for Configur…only root support (10s / no limit)[ 30.458739] NFS: Registering the id_resolver key type [ 30.461725] Key type id_resolver registered [ 30.469226] Key type id_legacy registered [ *] A start job is running for Configur…only root support (10s / no limit) [ OK ] Started Configure read-only root support. [ OK ] Reached target Local File Systems. Starting Mark the need to relabel after reboot... Starting Rebuild Dynamic Linker Cache... Starting Create Volatile Files and Directories... Starting Load/Save Random Seed... [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Create Volatile Files and Directories. [ OK ] Started Load/Save Random Seed. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Started daily update of the root trust anchor for DNSSEC. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. Starting Restore /run/initramfs on shutdown... Starting Login Service... [ OK ] Started D-Bus System Message Bus. Starting Network Manager... [ OK ] Reached target sshd-keygen.target. [ OK ] Started irqbalance daemon. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ OK ] Started dnf makecache --timer. [ OK ] Reached target Timers. [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Login Service. [ OK ] Started Network Manager. [ OK ] Reached target Network. Starting Dynamic System Tuning Daemon... Starting GSSAPI Proxy Daemon... Starting OpenSSH server daemon... Starting Network Manager Wait Online... Starting Hostname Service... [ OK ] Started OpenSSH server daemon. [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... [ OK ] Started Permit User Sessions. [ OK ] Started Serial Getty on ttyS0. [ OK ] Started Getty on tty1. [ OK ] Started Command Scheduler. [ OK ] Started Serial Getty on ttyS1. [ OK ] Reached target Login Prompts. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting Crash recovery kernel arming... Starting Notify NFS peers of a restart... Starting System Logging Service... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg416-server login: [ 98.122751] libcfs: loading out-of-tree module taints kernel. [ 98.153471] Key type ._llcrypt registered [ 98.155704] Key type .llcrypt registered [ 98.263376] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_hostid [ 124.974234] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing load_modules_local [ 126.442882] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 126.470417] alg: No test for adler32 (adler32-zlib) [ 128.048400] Lustre: Lustre: Build Version: 2.17.52_53_gcd13cf3 [ 129.066376] LNet: Added LNI 192.168.204.116@tcp [8/256/0/180] [ 130.855193] Key type lgssc registered [ 132.723069] Lustre: Echo OBD driver; http://www.lustre.org/ [ 151.046692] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 191.109077] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing load_modules_local [ 203.544425] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 203.569363] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 203.720017] hrtimer: interrupt took 10205929 ns [ 204.996251] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 205.035809] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 205.165028] Lustre: lustre-MDT0000: new disk, initializing [ 205.344240] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 205.366315] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 210.243652] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 224.670420] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 224.772328] Lustre: 6510:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 224.821536] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 224.826890] Lustre: Skipped 1 previous similar message [ 224.940320] Lustre: lustre-MDT0001: new disk, initializing [ 225.033898] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 225.053246] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 225.063074] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 230.108291] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 235.357852] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 245.223669] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 245.463804] Lustre: lustre-OST0000: new disk, initializing [ 245.467519] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 245.526384] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 249.531774] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 249.545780] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 249.651477] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 253.034870] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 268.598821] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 268.774139] Lustre: lustre-OST0001: new disk, initializing [ 268.781964] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 268.894601] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 274.984385] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 274.999456] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 275.051826] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 275.231970] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 288.016555] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 295.112882] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 307.078861] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing check_logdir /tmp/testlogs/ [ 312.323254] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing yml_node [ 316.964956] Lustre: DEBUG MARKER: Client: 2.17.52.53 [ 320.134204] Lustre: DEBUG MARKER: MDS: 2.17.52.53 [ 322.662971] Lustre: DEBUG MARKER: OSS: 2.17.52.53 [ 324.462639] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-dual ============----- Thu Apr 30 14:57:14 EDT 2026 [ 341.766556] Lustre: DEBUG MARKER: excepting tests: 14b 21b [ 343.679709] Lustre: DEBUG MARKER: skipping tests SLOW=no: 21b [ 345.475118] Lustre: DEBUG MARKER: === replay-dual: start setup 14:57:35 (1777575455) === [ 350.868420] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing check_config_client /mnt/lustre [ 373.059580] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 377.283264] Lustre: 13233:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 381.900398] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 388.240651] Lustre: DEBUG MARKER: === replay-dual: finish setup 14:58:18 (1777575498) === [ 390.264957] Lustre: DEBUG MARKER: == replay-dual test 0a: expired recovery with lost client ========================================================== 14:58:20 (1777575500) [ 399.979507] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 405.011334] Lustre: Failing over lustre-MDT0000 [ 405.638991] Lustre: server umount lustre-MDT0000 complete [ 408.037884] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 408.056653] Lustre: Skipped 2 previous similar messages [ 412.664902] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 412.679549] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 7 previous similar messages [ 417.820056] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 417.854107] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 6 previous similar messages [ 422.904113] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 422.925057] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 5 previous similar messages [ 423.394651] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777575520/real 1777575520] req@ffff9c4f44442a00 x1863922736032768/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777575536 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 423.429307] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 427.093549] LDISKFS-fs (dm-0): 10 truncates cleaned up [ 427.099965] LDISKFS-fs (dm-0): recovery complete [ 427.132217] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 428.026529] LustreError: 12251:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 428.047628] LustreError: 12251:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 7 previous similar messages [ 433.639667] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 433.660462] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 9 previous similar messages [ 433.952313] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 434.175194] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 438.751096] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 439.294613] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 541.500746] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 541.509263] Lustre: 14761:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 037efc13-9d1e-4122-9f07-c7756406aff9@192.168.204.16@tcp [ 541.521105] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 541.543595] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 541.550744] Lustre: Skipped 2 previous similar messages [ 541.553939] Lustre: 14761:0:(ldlm_lib.c:2933:target_recovery_thread()) too long recovery - read logs [ 541.565283] LustreError: dumping log to /tmp/lustre-log.1777575654.14761 [ 541.711650] Lustre: lustre-MDT0000: Recovery over after 1:47, of 3 clients 2 recovered and 1 was evicted. [ 541.770119] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:65) [ 541.775154] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:65) [ 565.476670] Lustre: DEBUG MARKER: == replay-dual test 0b: lost client during waiting for next transno ========================================================== 15:01:15 (1777575675) [ 572.419188] Lustre: lustre-MDT0000: haven't heard from client 037efc13-9d1e-4122-9f07-c7756406aff9 (at 192.168.204.16@tcp) in 31 seconds. I think it's dead, and I am evicting it. exp ffff9c4f42aaa000, cur 1777575685 deadline 1777575684 last 1777575654 [ 574.054431] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 576.353596] Lustre: Failing over lustre-MDT0000 [ 576.839608] Lustre: server umount lustre-MDT0000 complete [ 577.505791] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 577.512070] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 577.526216] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 577.526680] Lustre: Skipped 4 previous similar messages [ 577.544713] LustreError: 9477:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 3 previous similar messages [ 593.887111] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777575690/real 1777575690] req@ffff9c4f45326a00 x1863922736111104/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777575706 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 593.906132] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 593.925802] LustreError: 15961:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 593.948446] LustreError: 15961:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 18 previous similar messages [ 599.647748] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 599.649756] LDISKFS-fs (dm-0): recovery complete [ 599.675444] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 604.452538] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 604.500041] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 607.258146] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 609.188870] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 609.781261] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 622.547043] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:54 [ 627.705063] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:49 [ 632.841056] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:44 [ 637.946701] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:39 [ 643.066318] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:34 [ 643.066951] Lustre: lustre-MDT0001: haven't heard from client 037efc13-9d1e-4122-9f07-c7756406aff9 (at 192.168.204.16@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff9c5077481000, cur 1777575755 deadline 1777575754 last 1777575654 [ 653.311646] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:24 [ 653.343219] Lustre: Skipped 1 previous similar message [ 673.793419] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:03 [ 673.820485] Lustre: Skipped 3 previous similar messages [ 677.500111] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 677.509048] Lustre: 16498:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 30216cdd-bf64-49be-9081-72a5328b6b83@ [ 677.517636] Lustre: lustre-MDT0000: disconnecting 2 stale clients [ 708.067338] Lustre: lustre-MDT0001: haven't heard from client 6f228d2b-20c2-4a4a-b03f-3fc54484b363 (at 192.168.204.16@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff9c50774c8000, cur 1777575820 deadline 1777575819 last 1777575719 [ 709.626582] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 2 evicted) to recover in 1:08 [ 709.657039] Lustre: Skipped 6 previous similar messages [ 776.187243] Lustre: lustre-MDT0000: Denying connection for new client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp), waiting for 4 known clients (1 recovered, 1 in progress, and 2 evicted) to recover in 0:02 [ 776.216908] Lustre: Skipped 12 previous similar messages [ 778.500107] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 778.503860] Lustre: 16498:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 6f228d2b-20c2-4a4a-b03f-3fc54484b363@192.168.204.16@tcp [ 778.521015] Lustre: 16498:0:(genops.c:1622:class_disconnect_stale_exports()) Skipped 1 previous similar message [ 778.538261] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 778.548507] Lustre: 16498:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 778.579573] Lustre: 16498:0:(ldlm_lib.c:2933:target_recovery_thread()) too long recovery - read logs [ 778.581930] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 778.605683] Lustre: Skipped 2 previous similar messages [ 778.607929] LustreError: dumping log to /tmp/lustre-log.1777575891.16498 [ 778.776204] Lustre: lustre-MDT0000: Recovery over after 2:51, of 4 clients 1 recovered and 3 were evicted. [ 778.835247] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:97) [ 778.835574] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:97) [ 790.224602] Lustre: DEBUG MARKER: == replay-dual test 1: |X| simple create ================= 15:04:59 (1777575899) [ 799.529168] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 801.910305] Lustre: Failing over lustre-MDT0000 [ 802.313964] Lustre: server umount lustre-MDT0000 complete [ 802.874898] LustreError: 15961:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 802.898986] LustreError: 15961:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 13 previous similar messages [ 804.323823] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 804.328901] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 821.727411] Lustre: 3665:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777575918/real 1777575918] req@ffff9c50435e9880 x1863922736208640/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777575934 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 821.760604] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 826.062202] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 826.067584] LDISKFS-fs (dm-0): recovery complete [ 826.091708] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 831.970036] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c505182f800 x1863922736217088/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 832.307025] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 832.369875] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 834.015269] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 837.003717] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 837.620605] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 837.821166] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 837.880626] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:129) [ 837.883699] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:129) [ 844.646254] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 846.381629] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 855.620298] Lustre: DEBUG MARKER: == replay-dual test 2: |X| mkdir adir ==================== 15:06:05 (1777575965) [ 863.773700] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 865.698666] Lustre: Failing over lustre-MDT0000 [ 865.895925] Lustre: server umount lustre-MDT0000 complete [ 868.328293] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 868.346050] Lustre: Skipped 6 previous similar messages [ 868.348795] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 868.367895] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 46 previous similar messages [ 884.704401] Lustre: 3663:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777575981/real 1777575981] req@ffff9c4f42dc2300 x1863922736245632/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777575997 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 884.756376] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 889.523190] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 889.527557] LDISKFS-fs (dm-0): recovery complete [ 889.541102] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 894.947343] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c506fb1d880 x1863922736254976/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 895.331262] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 895.418915] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 896.004052] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 900.118635] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 900.601837] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 900.609914] Lustre: Skipped 3 previous similar messages [ 900.677808] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 900.723801] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:161) [ 900.729679] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:161) [ 907.861764] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 909.356762] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 918.434884] Lustre: DEBUG MARKER: == replay-dual test 3: |X| mkdir adir, mkdir adir/bdir === 15:07:07 (1777576027) [ 926.795246] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 929.360494] Lustre: Failing over lustre-MDT0000 [ 929.718939] Lustre: server umount lustre-MDT0000 complete [ 931.301028] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 931.312642] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 931.348815] Lustre: Skipped 2 previous similar messages [ 948.195723] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576044/real 1777576044] req@ffff9c4f43c44000 x1863922736286592/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576060 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 948.222775] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 951.784482] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 951.787143] LDISKFS-fs (dm-0): recovery complete [ 951.804286] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 957.411539] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f43c44000 x1863922736294656/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 957.742980] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 958.453647] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 962.280823] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 963.070818] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 963.094238] Lustre: Skipped 3 previous similar messages [ 963.220161] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 963.281853] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:193) [ 963.282792] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:193) [ 969.844484] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 971.391894] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 981.038306] Lustre: DEBUG MARKER: == replay-dual test 4: |X| mkdir adir (-EEXIST), mkdir adir/bdir ========================================================== 15:08:10 (1777576090) [ 988.454764] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 990.127869] Lustre: Failing over lustre-MDT0000 [ 990.465588] Lustre: server umount lustre-MDT0000 complete [ 993.761221] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 993.780712] Lustre: Skipped 2 previous similar messages [ 998.883799] LustreError: 6523:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 998.904583] LustreError: 6523:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 83 previous similar messages [ 1009.631130] Lustre: 3665:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576106/real 1777576106] req@ffff9c507d997800 x1863922736328832/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576122 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1009.664273] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1011.916461] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 1011.919411] LDISKFS-fs (dm-0): recovery complete [ 1011.929457] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1020.146143] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1020.149626] Lustre: Skipped 1 previous similar message [ 1020.194163] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1021.589907] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1024.384721] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1025.521588] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1025.529435] Lustre: Skipped 3 previous similar messages [ 1025.654955] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1025.711850] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:225) [ 1025.719118] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:225) [ 1031.835403] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1033.238543] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1041.136491] Lustre: DEBUG MARKER: == replay-dual test 5: open, unlink |X| close ============ 15:09:11 (1777576151) [ 1049.437375] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1051.731949] Lustre: Failing over lustre-MDT0000 [ 1052.114268] Lustre: server umount lustre-MDT0000 complete [ 1056.227609] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1056.241076] Lustre: Skipped 3 previous similar messages [ 1072.017341] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576168/real 1777576168] req@ffff9c4f43c45500 x1863922736368768/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576184 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1072.050928] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1075.187183] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1075.189670] LDISKFS-fs (dm-0): recovery complete [ 1075.196490] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1081.312677] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f43c47480 x1863922736377344/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1081.573871] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1082.866300] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1085.888885] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1087.085561] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1087.129179] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:257) [ 1087.129978] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:257) [ 1093.385728] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1094.940986] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1102.759874] Lustre: DEBUG MARKER: == replay-dual test 6: open1, open2, unlink |X| close1 [fail mds1] close2 ========================================================== 15:10:12 (1777576212) [ 1110.425648] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1112.490344] Lustre: Failing over lustre-MDT0000 [ 1112.551675] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1112.575039] Lustre: Skipped 3 previous similar messages [ 1112.593081] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 1112.600877] Lustre: Skipped 2 previous similar messages [ 1112.800155] Lustre: server umount lustre-MDT0000 complete [ 1134.047814] Lustre: 3664:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576230/real 1777576230] req@ffff9c505d954700 x1863922736406144/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576246 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1134.109670] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1136.295340] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1136.303694] LDISKFS-fs (dm-0): recovery complete [ 1136.312913] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1144.295911] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f43c45180 x1863922736415232/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1144.697088] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1145.858358] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1149.061721] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1149.950749] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1149.963228] Lustre: Skipped 7 previous similar messages [ 1150.044180] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1150.078251] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:289) [ 1150.081338] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:289) [ 1156.780892] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1158.469929] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1167.644082] Lustre: DEBUG MARKER: == replay-dual test 8: replay of resent request ========== 15:11:17 (1777576277) [ 1175.903742] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1177.012781] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1177.019233] LustreError: 12251:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f42fbe680 x1863922707874944/t38654705670(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:525/0 lens 512/448 e 0 to 0 dl 1777576300 ref 1 fl Interpret:/200/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1193.525575] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnecting [ 1193.595125] Lustre: 6518:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c4f42f00700 x1863922707874944/t38654705670(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:542/0 lens 512/2880 e 0 to 0 dl 1777576317 ref 1 fl Interpret:/202/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1197.911295] Lustre: Failing over lustre-MDT0000 [ 1198.150823] Lustre: server umount lustre-MDT0000 complete [ 1201.125284] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1201.162295] Lustre: Skipped 3 previous similar messages [ 1216.480427] Lustre: 3665:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576313/real 1777576313] req@ffff9c505dac5f80 x1863922736453120/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576329 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1216.541941] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1221.239210] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1221.243526] LDISKFS-fs (dm-0): recovery complete [ 1221.256561] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1226.734927] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c04840 [ 1227.054736] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1227.060805] Lustre: Skipped 2 previous similar messages [ 1227.107533] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1229.304860] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1231.341339] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1232.528134] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 1232.570973] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:321) [ 1232.571824] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:321) [ 1239.457703] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1241.124728] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1249.885687] Lustre: DEBUG MARKER: == replay-dual test 9: resending a replayed create ======= 15:12:39 (1777576359) [ 1257.800126] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1260.419987] Lustre: Failing over lustre-MDT0000 [ 1260.647957] Lustre: server umount lustre-MDT0000 complete [ 1263.074529] LustreError: 6524:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1263.092953] LustreError: 6524:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 152 previous similar messages [ 1284.455251] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1284.462445] LDISKFS-fs (dm-0): recovery complete [ 1284.472037] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1288.680919] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c50510d2d80 x1863922736498688/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1293.936826] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1294.329577] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1294.333567] Lustre: Skipped 8 previous similar messages [ 1294.356049] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1294.364631] LustreError: 31190:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c5080bbb800 x1863922707892736/t42949672962(42949672962) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:639/0 lens 528/448 e 0 to 0 dl 1777576414 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1307.173901] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnected, waiting for 3 clients in recovery for 1:27 [ 1307.369859] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:353) [ 1307.376380] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:353) [ 1312.112623] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1313.840876] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1323.622293] Lustre: DEBUG MARKER: == replay-dual test 10: resending a replayed unlink ====== 15:13:53 (1777576433) [ 1331.505978] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1334.320559] Lustre: Failing over lustre-MDT0000 [ 1334.580479] Lustre: server umount lustre-MDT0000 complete [ 1335.276604] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1335.289638] Lustre: Skipped 5 previous similar messages [ 1350.560926] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576447/real 1777576447] req@ffff9c4f44eec700 x1863922736531840/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576463 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1350.593574] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1350.603323] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1350.619912] LustreError: Skipped 1 previous similar message [ 1356.775303] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1356.778322] LDISKFS-fs (dm-0): recovery complete [ 1356.786177] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1360.877434] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c05487 [ 1361.345500] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1361.360957] Lustre: Skipped 1 previous similar message [ 1363.457411] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1363.463764] Lustre: Skipped 1 previous similar message [ 1365.447805] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1366.560163] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1366.574174] LustreError: 33136:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c507bb4e680 x1863922707911808/t47244640260(47244640260) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:712/0 lens 528/448 e 0 to 0 dl 1777576487 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1378.835143] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1378.944959] Lustre: lustre-MDT0000: Recovery over after 0:15, of 3 clients 3 recovered and 0 were evicted. [ 1378.960149] Lustre: Skipped 1 previous similar message [ 1379.050298] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:385) [ 1379.069111] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:385) [ 1383.251353] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1384.918833] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1394.512989] Lustre: DEBUG MARKER: == replay-dual test 11: both clients timeout during replay ========================================================== 15:15:04 (1777576504) [ 1402.034161] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1404.804555] Lustre: Failing over lustre-MDT0000 [ 1404.896404] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1404.911779] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 1405.142318] Lustre: server umount lustre-MDT0000 complete [ 1427.150482] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1427.154512] LDISKFS-fs (dm-0): recovery complete [ 1427.168816] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1433.588735] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c505daf3100 x1863922736580352/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1438.447426] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1439.298390] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1439.302353] LustreError: 35079:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f43c9bb80 x1863922707930496/t51539607554(51539607554) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:28/0 lens 528/448 e 0 to 0 dl 1777576558 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1444.184775] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1451.549788] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnected, waiting for 3 clients in recovery for 1:27 [ 1451.686485] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:417) [ 1451.689888] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:417) [ 1454.111493] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 7 sec [ 1461.616681] Lustre: DEBUG MARKER: == replay-dual test 12: open resend timeout ============== 15:16:11 (1777576571) [ 1468.723709] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1471.566202] Lustre: Failing over lustre-MDT0000 [ 1471.886375] Lustre: server umount lustre-MDT0000 complete [ 1472.483267] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1495.458308] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1495.461540] LDISKFS-fs (dm-0): recovery complete [ 1495.472815] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1503.135258] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1503.145503] Lustre: Skipped 3 previous similar messages [ 1503.199195] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1503.210462] Lustre: Skipped 1 previous similar message [ 1507.502203] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1508.448662] Lustre: *** cfs_fail_loc=302, val=2147483648*** [ 1523.760388] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnected, waiting for 3 clients in recovery for 1:24 [ 1523.919263] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:449) [ 1523.926047] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:449) [ 1531.402806] Lustre: DEBUG MARKER: == replay-dual test 13: close resend timeout ============= 15:17:21 (1777576641) [ 1540.028804] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1543.190119] Lustre: Failing over lustre-MDT0000 [ 1543.680145] Lustre: server umount lustre-MDT0000 complete [ 1567.630266] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1567.633527] LDISKFS-fs (dm-0): recovery complete [ 1567.641648] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1575.083459] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1575.404831] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1575.414434] Lustre: Skipped 16 previous similar messages [ 1575.477373] Lustre: *** cfs_fail_loc=115, val=2147483648*** [ 1590.833796] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnected, waiting for 3 clients in recovery for 1:24 [ 1591.001042] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:481) [ 1591.002196] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:481) [ 1600.564130] Lustre: DEBUG MARKER: SKIP: replay-dual test_14b skipping ALWAYS excluded test 14b [ 1602.484265] Lustre: DEBUG MARKER: == replay-dual test 15a: timeout waiting for lost client during replay, 1 client completes ========================================================== 15:18:32 (1777576712) [ 1611.934535] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1615.141205] Lustre: Failing over lustre-MDT0000 [ 1615.311324] Lustre: server umount lustre-MDT0000 complete [ 1616.354093] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1616.378478] Lustre: Skipped 15 previous similar messages [ 1632.746592] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777576729/real 1777576729] req@ffff9c505d9c7480 x1863922736689920/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777576745 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1632.834711] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 1632.856966] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1632.875717] LustreError: Skipped 3 previous similar messages [ 1639.436632] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1639.454653] LDISKFS-fs (dm-0): recovery complete [ 1639.471304] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1642.476475] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c505ff2ad80 x1863922736698240/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1643.359475] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1643.368549] Lustre: Skipped 3 previous similar messages [ 1646.716672] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1713.501032] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1713.516283] Lustre: 40612:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 2a76c88c-539d-4bf1-a132-9d72bf6dbae4@ [ 1713.544966] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1714.359925] Lustre: lustre-MDT0000: Recovery over after 1:11, of 3 clients 2 recovered and 1 was evicted. [ 1714.370660] Lustre: Skipped 3 previous similar messages [ 1714.425180] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:495 to 0x280000401:513) [ 1714.426500] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:494 to 0x2c0000401:513) [ 1719.219792] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1721.363659] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1732.756573] Lustre: DEBUG MARKER: == replay-dual test 15c: remove multiple OST orphans ===== 15:20:42 (1777576842) [ 1742.122328] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1882.331579] Lustre: Failing over lustre-MDT0000 [ 1882.705392] Lustre: server umount lustre-MDT0000 complete [ 1883.639341] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1883.641078] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1883.672121] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 223 previous similar messages [ 1908.709451] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1908.713566] LDISKFS-fs (dm-0): recovery complete [ 1908.743283] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1909.855149] LustreError: 42481:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 1909.860200] LustreError: 42481:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c505d912a00 x1863922736821888/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1777577022 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1909.894870] LustreError: 42481:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 1910.280585] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c236a0 [ 1910.851084] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1910.876763] Lustre: Skipped 2 previous similar messages [ 1917.071333] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 1982.500175] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1982.505565] Lustre: 42514:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 17e21611-9e2a-465d-b1f9-f27be9c69cf8@ [ 1982.516208] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1982.688338] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:495 to 0x280000401:1537) [ 1982.693125] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:494 to 0x2c0000401:1537) [ 1987.827358] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1989.474845] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1998.436858] Lustre: DEBUG MARKER: == replay-dual test 16: fail MDS during recovery (3571) == 15:25:08 (1777577108) [ 2008.206818] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2011.371283] Lustre: Failing over lustre-MDT0000 [ 2011.814139] Lustre: server umount lustre-MDT0000 complete [ 2013.665718] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2036.831440] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2036.840398] LDISKFS-fs (dm-0): recovery complete [ 2036.850527] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2039.776883] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c507d951180 x1863922736885888/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2040.252338] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2040.259460] Lustre: Skipped 3 previous similar messages [ 2045.763265] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2070.713055] Lustre: Failing over lustre-MDT0000 [ 2070.724790] LustreError: 44825:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 2070.737474] Lustre: 44367:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2070.746122] Lustre: 44367:0:(ldlm_lib.c:1899:abort_req_replay_queue()) @@@ aborted: req@ffff9c4f472f4700 x1863922710411648/t0(73014444033) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:658/0 lens 528/0 e 2 to 0 dl 1777577188 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 2070.765727] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 2070.783178] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 2070.795333] LustreError: 44367:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9c5077489880 x1863922736909568/t0(0) o1000->lustre-MDT0001-osp-MDT0000@0@lo:24/4 lens 336/33016 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'tgt_recover_0.0' uid:0 gid:0 projid:4294967295 [ 2070.804121] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.16@tcp (stopping) [ 2070.807443] LustreError: 44367:0:(llog_osd.c:1177:llog_osd_next_block()) lustre-MDT0001-osp-MDT0000: can't read llog block from log [0x240000401:0x1:0x0] offset 32768: rc = -5 [ 2070.835497] LustreError: 44367:0:(llog.c:870:llog_process_thread()) lustre-MDT0001-osp-MDT0000 retry remote llog process [ 2070.843449] LustreError: 44367:0:(fid_request.c:213:seq_client_alloc_seq()) cli-cli-lustre-MDT0001-osp-MDT0000: Cannot allocate new meta-sequence: rc = -5 [ 2070.851491] LustreError: 44367:0:(fid_request.c:316:seq_client_alloc_fid()) cli-cli-lustre-MDT0001-osp-MDT0000: Can't allocate new sequence: rc = -5 [ 2071.133554] Lustre: server umount lustre-MDT0000 complete [ 2090.324697] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2096.625646] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c2797b [ 2096.651447] Lustre: MGC192.168.204.116@tcp: Connection restored to 0@lo (at 0@lo) [ 2096.659891] Lustre: Skipped 15 previous similar messages [ 2102.279623] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2167.502110] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2167.513862] Lustre: 45283:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 0ca4f65d-86ac-4dd4-8e29-b7db8166485d@ [ 2167.541947] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2168.362569] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1550 to 0x280000401:1569) [ 2168.364044] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1551 to 0x2c0000401:1569) [ 2173.425397] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2175.917464] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2186.985029] Lustre: DEBUG MARKER: == replay-dual test 17: fail OST during recovery (3571) == 15:28:17 (1777577297) [ 2196.545156] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 2199.025643] Lustre: Failing over lustre-OST0000 [ 2199.236624] Lustre: server umount lustre-OST0000 complete [ 2199.526992] LustreError: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 2199.531087] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2199.564613] Lustre: Skipped 14 previous similar messages [ 2225.143677] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 2225.151861] LDISKFS-fs (dm-2): recovery complete [ 2225.166970] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2226.178766] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 2226.185129] Lustre: Skipped 3 previous similar messages [ 2232.615516] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2257.610980] Lustre: Failing over lustre-OST0000 [ 2257.628401] LustreError: 47670:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 2257.657983] Lustre: 47122:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2257.676461] Lustre: 47122:0:(ldlm_lib.c:2389:target_recovery_overseer()) Skipped 2 previous similar messages [ 2257.684243] LustreError: 47122:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 2257.693046] Lustre: lustre-OST0000: Recovery over after 0:31, of 4 clients 0 recovered and 4 were evicted. [ 2257.703462] Lustre: Skipped 3 previous similar messages [ 2257.964747] Lustre: server umount lustre-OST0000 complete [ 2279.118909] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2279.392142] Lustre: 3662:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777577339/real 1777577339] req@ffff9c4f472f5880 x1863922736989056/t0(0) o400->lustre-OST0000-osc-MDT0001@0@lo:28/4 lens 224/224 e 3 to 1 dl 1777577392 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 2279.467065] Lustre: 3662:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2287.827321] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2351.500258] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 2351.507940] Lustre: 48111:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client 5c782a87-5d5a-43a9-94dd-9b22a97f1ba0@ [ 2351.515885] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 2356.147710] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 2357.923198] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 2368.623968] Lustre: DEBUG MARKER: == replay-dual test 18: ldlm_handle_enqueue succeeds on evicted export (3822) ========================================================== 15:31:18 (1777577478) [ 2373.115326] LustreError: 6519:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b sleeping for 40000ms [ 2413.151284] LustreError: 6519:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b awake [ 2429.476210] Lustre: DEBUG MARKER: == replay-dual test 19: resend of open request =========== 15:32:19 (1777577539) [ 2438.295222] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2439.883950] Lustre: *** cfs_fail_loc=157, val=2147483648*** [ 2439.890464] LustreError: 6519:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c507d9c9880 x1863922710538880/t0(0) o101->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:348/0 lens 576/688 e 0 to 0 dl 1777577633 ref 1 fl Interpret:/600/0 rc 0/0 job:'createmany.0' uid:0 gid:0 projid:0 [ 2525.719248] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnecting [ 2529.643961] Lustre: Failing over lustre-MDT0000 [ 2530.051490] Lustre: server umount lustre-MDT0000 complete [ 2530.845398] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2530.870888] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 112 previous similar messages [ 2531.808278] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2548.194295] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2548.213414] LustreError: Skipped 3 previous similar messages [ 2552.598227] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 2552.601449] LDISKFS-fs (dm-0): recovery complete [ 2552.610070] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2557.414670] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c29582 [ 2557.878528] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2557.892851] Lustre: Skipped 4 previous similar messages [ 2562.771406] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2563.165024] Lustre: 50648:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2563.484784] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1601) [ 2563.486285] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1601) [ 2571.702658] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2573.840524] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2583.412408] Lustre: DEBUG MARKER: == replay-dual test 20: recovery time is not increasing == 15:34:53 (1777577693) [ 2591.498570] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2594.244417] Lustre: Failing over lustre-MDT0000 [ 2594.278183] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 2594.579305] Lustre: server umount lustre-MDT0000 complete [ 2618.797568] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2618.808925] LDISKFS-fs (dm-0): recovery complete [ 2618.822856] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2625.006769] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c29b4e [ 2629.755563] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2766.500585] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2766.510885] Lustre: 52478:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client de416d69-a833-42dd-9a20-dc04e6b131e3@ [ 2766.526538] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2766.567928] Lustre: 52478:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2766.579684] Lustre: 52478:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 6 previous similar messages [ 2766.646405] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2766.650203] Lustre: Skipped 15 previous similar messages [ 2766.711253] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1633) [ 2766.713854] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1603 to 0x280000401:1633) [ 2772.595324] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2774.645319] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2786.584730] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2788.907242] Lustre: Failing over lustre-MDT0000 [ 2789.297286] Lustre: server umount lustre-MDT0000 complete [ 2812.118799] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2812.123261] LDISKFS-fs (dm-0): recovery complete [ 2812.135453] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2815.974423] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c505d9c4000 x1863922737261184/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2815.990952] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 1 previous similar message [ 2816.361380] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2816.366104] Lustre: Skipped 5 previous similar messages [ 2821.151469] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 2959.502090] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2959.517508] Lustre: 54148:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 3f7c39cc-4eb9-43d0-8203-e7c625b5ddcb@ [ 2959.546433] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2959.630383] Lustre: 54148:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2959.651301] Lustre: 54148:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 2959.711379] Lustre: lustre-MDT0000: Recovery over after 2:20, of 3 clients 2 recovered and 1 was evicted. [ 2959.718451] Lustre: Skipped 3 previous similar messages [ 2959.802455] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1665) [ 2959.805211] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1665) [ 2964.572930] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2966.132176] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2975.633158] Lustre: DEBUG MARKER: == replay-dual test 21a: commit on sharing =============== 15:41:25 (1777578085) [ 2984.542414] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2986.780464] Lustre: Failing over lustre-MDT0000 [ 2987.101575] Lustre: server umount lustre-MDT0000 complete [ 2990.562336] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2990.581819] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2990.583899] Lustre: Skipped 14 previous similar messages [ 2990.606596] LustreError: Skipped 1 previous similar message [ 3006.416096] Lustre: 3663:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777578103/real 1777578103] req@ffff9c507d8be300 x1863922737340288/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777578119 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3006.471741] Lustre: 3663:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 3011.967627] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3011.976516] LDISKFS-fs (dm-0): recovery complete [ 3011.997507] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3017.423727] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3017.429474] Lustre: Skipped 4 previous similar messages [ 3021.792116] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3157.500606] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3157.509992] Lustre: 56066:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 4a355c74-5285-4563-aa1f-9123fde32cc3@ [ 3157.541431] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3157.680557] Lustre: 56066:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3157.698226] Lustre: 56066:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3157.786119] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1697) [ 3157.786861] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1697) [ 3171.192904] Lustre: DEBUG MARKER: SKIP: replay-dual test_21b skipping SLOW test 21b [ 3173.012804] Lustre: DEBUG MARKER: == replay-dual test 22a: c1 lfs mkdir -i 1 dir1, M1 drop reply [ 3174.650948] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3174.663498] LustreError: 15961:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c505da48e00 x1863922710662016/t4294967341(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:327/0 lens 560/448 e 0 to 0 dl 1777578367 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3177.436951] Lustre: Failing over lustre-MDT0001 [ 3177.684730] Lustre: server umount lustre-MDT0001 complete [ 3178.988762] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3179.001803] LustreError: 6518:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 149 previous similar messages [ 3198.788520] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3199.502819] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3199.514997] Lustre: Skipped 3 previous similar messages [ 3203.499785] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3204.671404] Lustre: 6518:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c505dac5880 x1863922710662016/t4294967341(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:357/0 lens 560/2880 e 0 to 0 dl 1777578397 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3212.834266] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3214.748778] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3225.925570] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3228.994525] Lustre: Failing over lustre-MDT0000 [ 3229.416978] Lustre: server umount lustre-MDT0000 complete [ 3246.035874] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3246.048308] LustreError: Skipped 3 previous similar messages [ 3254.668978] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3254.671855] LDISKFS-fs (dm-0): recovery complete [ 3254.708714] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3255.968578] LustreError: 58970:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 3255.983250] LustreError: 58970:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c4f4881c000 x1863922737459072/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1777578368 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3256.013749] LustreError: 58970:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 3256.334494] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f476d6d80 x1863922737462528/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3262.019959] Lustre: 59004:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3262.136893] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1729) [ 3262.146252] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1729) [ 3262.898332] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3271.355896] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3273.187465] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3283.437731] Lustre: DEBUG MARKER: == replay-dual test 22b: c1 lfs mkdir -i 1 d1, M1 drop reply [ 3284.635048] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3284.639880] LustreError: 9477:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f44c32680 x1863922710702848/t8589934617(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:437/0 lens 560/448 e 0 to 0 dl 1777578477 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3287.346407] Lustre: Failing over lustre-MDT0000 [ 3287.530323] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3287.539775] Lustre: Skipped 2 previous similar messages [ 3287.638040] Lustre: server umount lustre-MDT0000 complete [ 3288.036649] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3288.050692] LustreError: Skipped 1 previous similar message [ 3291.343315] LustreError: 6504:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1777578404 with bad export cookie 8790466873432125427 [ 3291.346698] Lustre: Failing over lustre-MDT0001 [ 3291.368968] LustreError: 6504:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 3291.707341] Lustre: server umount lustre-MDT0001 complete [ 3311.864388] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3312.234804] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3317.153979] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c505da4b100 x1863922737499008/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3323.002851] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3323.205880] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3324.061376] Lustre: 60756:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c4f42e73b80 x1863922710702848/t8589934617(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:476/0 lens 560/2880 e 0 to 0 dl 1777578516 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3324.097882] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:36 to 0x280000400:65) [ 3324.102472] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:36 to 0x2c0000400:65) [ 3328.811983] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1761) [ 3328.811737] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1761) [ 3333.883301] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3335.924816] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3337.855981] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3349.134416] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3351.892705] Lustre: Failing over lustre-MDT0000 [ 3352.442127] Lustre: server umount lustre-MDT0000 complete [ 3375.358498] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3375.360761] LDISKFS-fs (dm-0): recovery complete [ 3375.372309] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3385.495817] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3386.347760] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 3386.355536] Lustre: Skipped 21 previous similar messages [ 3386.431144] Lustre: 62832:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3386.441392] Lustre: 62832:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3386.541889] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1793) [ 3386.542903] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1793) [ 3393.969942] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3395.767518] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3407.835576] Lustre: DEBUG MARKER: == replay-dual test 22c: c1 lfs mkdir -i 1 d1, M1 drop update [ 3409.852277] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3409.859904] LustreError: 60808:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f42fbf480 x1863922737570432/t107374182411(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:493/0 lens 2520/4320 e 0 to 0 dl 1777578533 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3413.308284] Lustre: Failing over lustre-MDT0000 [ 3414.058093] Lustre: server umount lustre-MDT0000 complete [ 3437.331762] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3442.662209] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f48656300 x1863922737581312/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3442.732822] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 1 previous similar message [ 3443.328096] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 3443.333727] Lustre: Skipped 6 previous similar messages [ 3448.465490] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1825) [ 3448.476363] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1825) [ 3450.287224] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3460.329637] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3462.318542] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3472.768104] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3475.714102] Lustre: Failing over lustre-MDT0000 [ 3475.940327] Lustre: server umount lustre-MDT0000 complete [ 3498.462534] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3498.464955] LDISKFS-fs (dm-0): recovery complete [ 3498.477326] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3510.077341] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3510.340477] Lustre: 65800:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3510.354119] Lustre: 65800:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3510.505678] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1857) [ 3510.505989] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1857) [ 3519.420994] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3521.716916] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3532.289913] Lustre: DEBUG MARKER: == replay-dual test 22d: c1 lfs mkdir -i 1 d1, M1 drop update [ 3537.461450] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3537.465456] LustreError: 8410:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c5073dc2c50 x1863922737652736/t115964117002(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:621/0 lens 2520/4320 e 0 to 0 dl 1777578661 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3542.099837] Lustre: Failing over lustre-MDT0000 [ 3542.403111] Lustre: server umount lustre-MDT0000 complete [ 3547.015870] LustreError: 6503:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1777578659 with bad export cookie 8790466873432133092 [ 3547.018497] Lustre: Failing over lustre-MDT0001 [ 3547.033123] LustreError: 6503:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 3547.048111] LustreError: 66924:0:(ldlm_resource.c:1172:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x78:0x0].0xf7117594 (ffff9c4f4607cb00) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3547.071741] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.16@tcp (stopping) [ 3551.211900] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 3551.223659] Lustre: Skipped 2 previous similar messages [ 3553.557925] Lustre: server umount lustre-MDT0001 complete [ 3575.233667] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3575.236088] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3575.504873] LustreError: 67634:0:(llog.c:1646:llog_backup()) MGC192.168.204.116@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3575.511295] Lustre: 67634:0:(mgc_request_server.c:768:mgc_llog_local_copy()) MGC192.168.204.116@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3592.160805] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c4f482caa00 x1863922737663360/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3592.193299] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 2 previous similar messages [ 3597.511574] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3597.767212] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3598.430016] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 3598.441895] Lustre: Skipped 8 previous similar messages [ 3598.482111] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1889) [ 3598.483039] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1889) [ 3598.675363] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 3598.682277] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 3598.805186] Lustre: 67641:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c4f46c3b800 x1863922710793600/t12884901939(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:751/0 lens 560/2880 e 0 to 0 dl 1777578791 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3608.509607] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3610.609355] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3613.273178] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3625.514604] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3628.190971] Lustre: Failing over lustre-MDT0000 [ 3628.508804] Lustre: server umount lustre-MDT0000 complete [ 3629.036308] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3629.063704] Lustre: Skipped 35 previous similar messages [ 3645.407289] Lustre: 3663:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777578741/real 1777578741] req@ffff9c4f48653100 x1863922737698304/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777578757 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3645.447886] Lustre: 3663:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 21 previous similar messages [ 3652.926514] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3652.930507] LDISKFS-fs (dm-0): recovery complete [ 3652.951609] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3656.686599] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c2de30 [ 3658.431644] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3658.440869] Lustre: Skipped 9 previous similar messages [ 3662.384707] Lustre: 69729:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3662.394397] Lustre: 69729:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3662.541436] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1921) [ 3662.545676] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1921) [ 3662.639873] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3671.350045] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3674.002930] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3685.445985] Lustre: DEBUG MARKER: == replay-dual test 23a: c1 rmdir d1, M1 drop reply and fail, client2 mkdir d1 ========================================================== 15:53:14 (1777578794) [ 3686.923080] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3686.925044] LustreError: 68481:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f46f00700 x1863922710842752/t17179869210(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:84/0 lens 496/456 e 0 to 0 dl 1777578879 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3691.565759] Lustre: Failing over lustre-MDT0001 [ 3692.376929] Lustre: server umount lustre-MDT0001 complete [ 3711.978391] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3717.222279] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3717.769883] Lustre: 67642:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c5072f9df80 x1863922710842752/t17179869210(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:115/0 lens 496/2888 e 0 to 0 dl 1777578910 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3717.772522] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:129) [ 3717.773313] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:129) [ 3726.203715] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3728.050552] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3739.189711] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3742.047128] Lustre: Failing over lustre-MDT0000 [ 3742.710405] Lustre: server umount lustre-MDT0000 complete [ 3769.697208] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3769.699318] LDISKFS-fs (dm-0): recovery complete [ 3769.708421] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3780.068838] LustreError: 60808:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3780.094818] LustreError: 60808:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 341 previous similar messages [ 3790.941643] Lustre: 72678:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3790.957785] Lustre: 72678:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3791.150203] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1953) [ 3791.151713] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1953) [ 3791.405721] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3800.728776] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3803.145079] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3814.434683] Lustre: DEBUG MARKER: == replay-dual test 23b: c1 rmdir d1, M1 drop reply and fail M0/M1, c2 mkdir d1 ========================================================== 15:55:24 (1777578924) [ 3816.388552] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3816.395541] LustreError: 67642:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c4f42869880 x1863922710883584/t21474836483(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:214/0 lens 496/456 e 0 to 0 dl 1777579009 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3821.405760] Lustre: Failing over lustre-MDT0000 [ 3821.549874] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3821.558094] Lustre: Skipped 4 previous similar messages [ 3821.919752] Lustre: server umount lustre-MDT0000 complete [ 3822.056220] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3822.070284] LustreError: Skipped 3 previous similar messages [ 3826.867934] LustreError: 6504:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1777578939 with bad export cookie 8790466873432140358 [ 3826.877076] LustreError: 6504:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 3826.879830] Lustre: Failing over lustre-MDT0001 [ 3827.155526] Lustre: server umount lustre-MDT0001 complete [ 3847.964953] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3848.058557] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3852.256773] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c507c514a80 x1863922737824640/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3852.798223] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3852.809706] Lustre: Skipped 11 previous similar messages [ 3856.986712] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3857.796237] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3858.079835] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:161) [ 3858.088369] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:161) [ 3858.114677] Lustre: 74475:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c4f48652d80 x1863922710883584/t21474836483(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:255/0 lens 496/2888 e 0 to 0 dl 1777579050 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3862.903071] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1985) [ 3862.903118] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1985) [ 3866.698338] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3868.535477] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3870.295144] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3880.527927] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3883.144246] Lustre: Failing over lustre-MDT0000 [ 3883.563485] Lustre: server umount lustre-MDT0000 complete [ 3902.367190] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3902.377563] LustreError: Skipped 8 previous similar messages [ 3907.723262] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3907.730597] LDISKFS-fs (dm-0): recovery complete [ 3907.755869] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3918.075101] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3918.586346] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2017) [ 3918.593636] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2017) [ 3927.301637] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3929.789976] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3940.643156] Lustre: DEBUG MARKER: == replay-dual test 23c: c1 rmdir d1, M0 drop update reply and fail M0, c2 mkdir d1 ========================================================== 15:57:30 (1777579050) [ 3942.518933] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3942.530923] LustreError: 8409:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c506fb25c00 x1863922737896576/t137438953491(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:271/0 lens 1984/4320 e 0 to 0 dl 1777579066 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3946.593272] Lustre: Failing over lustre-MDT0000 [ 3947.003509] Lustre: server umount lustre-MDT0000 complete [ 3967.856983] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3980.496319] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 3981.398713] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2049) [ 3981.402918] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2049) [ 3988.749014] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3990.679567] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4002.102593] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4004.654465] Lustre: Failing over lustre-MDT0000 [ 4005.063926] Lustre: server umount lustre-MDT0000 complete [ 4030.168595] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4030.172893] LDISKFS-fs (dm-0): recovery complete [ 4030.180336] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4032.415097] LustreError: 79477:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 4032.426459] LustreError: 79477:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c5072e8e300 x1863922737941760/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1777579145 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4032.454838] LustreError: 79477:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 4037.009614] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4038.127996] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 4038.147492] Lustre: Skipped 43 previous similar messages [ 4038.189889] Lustre: 79510:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4038.211320] Lustre: 79510:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 17 previous similar messages [ 4038.387623] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2081) [ 4038.391934] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2081) [ 4044.681705] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4046.853353] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4059.182572] Lustre: DEBUG MARKER: == replay-dual test 23d: c1 rmdir d1, M0 drop update reply and fail M0/M1, c2 mkdir d1 ========================================================== 15:59:28 (1777579168) [ 4065.158869] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 4065.165945] LustreError: 8409:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c506e4b2d80 x1863922737974784/t146028888081(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:393/0 lens 1984/4320 e 0 to 0 dl 1777579188 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 4069.556963] Lustre: Failing over lustre-MDT0000 [ 4070.060786] Lustre: server umount lustre-MDT0000 complete [ 4074.816622] LustreError: 6505:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1777579187 with bad export cookie 8790466873432147617 [ 4074.824602] Lustre: Failing over lustre-MDT0001 [ 4074.860989] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.16@tcp (stopping) [ 4081.739809] Lustre: server umount lustre-MDT0001 complete [ 4103.748091] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4104.031042] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4104.272279] LustreError: 81327:0:(llog.c:1646:llog_backup()) MGC192.168.204.116@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 4104.280663] Lustre: 81327:0:(mgc_request_server.c:768:mgc_llog_local_copy()) MGC192.168.204.116@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 4120.524268] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 4120.536183] Lustre: Skipped 11 previous similar messages [ 4126.624326] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4126.890859] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4129.226159] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:193) [ 4129.244937] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:193) [ 4129.293127] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2113) [ 4129.297357] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2113) [ 4129.430207] Lustre: 81356:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c506f3d5500 x1863922710965376/t25769803783(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:526/0 lens 496/2888 e 0 to 0 dl 1777579321 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 4136.249317] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4138.190701] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4140.302947] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4151.212443] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4153.836620] Lustre: Failing over lustre-MDT0000 [ 4154.040778] Lustre: server umount lustre-MDT0000 complete [ 4179.333134] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4179.338083] LDISKFS-fs (dm-0): recovery complete [ 4179.348125] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4181.791148] LustreError: 83413:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 4181.805384] LustreError: 83413:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c50815d4700 x1863922738024832/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1777579294 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4181.826497] LustreError: 83413:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 4187.053202] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4187.922829] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2115 to 0x2c0000401:2145) [ 4187.927016] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2115 to 0x280000401:2145) [ 4195.577953] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4197.414893] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4207.514294] Lustre: DEBUG MARKER: == replay-dual test 24: reconstruct on non-existing object ========================================================== 16:01:57 (1777579317) [ 4208.734780] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 4208.738741] LustreError: 81356:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c5050809880 x1863922711005824/t154618822673(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:606/0 lens 488/456 e 0 to 0 dl 1777579401 ref 1 fl Interpret:/200/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4294.218540] Lustre: lustre-MDT0000: Client 3f6f4915-1adb-4bd0-b523-4ca667be7c42 (at 192.168.204.16@tcp) reconnecting [ 4294.253231] Lustre: 82154:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c5046ec9f80 x1863922711005824/t154618822673(0) o36->3f6f4915-1adb-4bd0-b523-4ca667be7c42@192.168.204.16@tcp:691/0 lens 488/3152 e 0 to 0 dl 1777579486 ref 1 fl Interpret:/202/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4302.610474] Lustre: DEBUG MARKER: == replay-dual test 25: replay|resend ==================== 16:03:32 (1777579412) [ 4304.912922] Lustre: *** cfs_fail_loc=304, val=0*** [ 4307.108671] Lustre: Failing over lustre-OST0000 [ 4307.230622] Lustre: server umount lustre-OST0000 complete [ 4308.453129] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4308.468258] Lustre: Skipped 35 previous similar messages [ 4326.766084] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4328.170931] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 4328.186643] Lustre: Skipped 10 previous similar messages [ 4328.911291] Lustre: lustre-OST0000: Recovery over after 0:01, of 4 clients 4 recovered and 0 were evicted. [ 4328.923157] Lustre: Skipped 12 previous similar messages [ 4333.898605] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4342.054470] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4343.827984] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4354.534214] Lustre: DEBUG MARKER: == replay-dual test 26: dbench and tar with mds failover ========================================================== 16:04:24 (1777579464) [ 4366.288664] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4370.234923] Lustre: DEBUG MARKER: test_26 fail mds1 1 times [ 4372.471458] Lustre: Failing over lustre-MDT0000 [ 4372.521256] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.16@tcp (stopping) [ 4372.534375] Lustre: Skipped 4 previous similar messages [ 4372.978817] Lustre: server umount lustre-MDT0000 complete [ 4381.182104] LustreError: 82910:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.16@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4381.213340] LustreError: 82910:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 219 previous similar messages [ 4392.927160] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1777579489/real 1777579489] req@ffff9c507b8c2a00 x1863922738176896/t0(0) o400->MGC192.168.204.116@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1777579505 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4392.955508] Lustre: 3666:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 16 previous similar messages [ 4395.501546] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4395.506329] LDISKFS-fs (dm-0): recovery complete [ 4395.519048] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4403.178834] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x79fe030c28c35abd [ 4407.555110] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4408.815126] Lustre: 87144:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4408.836841] Lustre: 87144:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 17 previous similar messages [ 4410.589472] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2170 to 0x2c0000401:2209) [ 4410.608673] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2171 to 0x280000401:2209) [ 4416.604791] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4418.695362] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4431.167529] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4434.930798] Lustre: DEBUG MARKER: test_26 fail mds2 2 times [ 4436.735837] Lustre: Failing over lustre-MDT0001 [ 4436.937136] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.16@tcp (stopping) [ 4443.340939] Lustre: server umount lustre-MDT0001 complete [ 4469.975624] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4469.980090] LDISKFS-fs (dm-1): recovery complete [ 4470.003525] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4470.931785] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 4470.959837] Lustre: Skipped 9 previous similar messages [ 4476.479058] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4479.320378] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:264 to 0x2c0000400:289) [ 4479.320590] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:263 to 0x280000400:289) [ 4486.909831] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4489.215191] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4502.154046] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4506.284347] Lustre: DEBUG MARKER: test_26 fail mds1 3 times [ 4508.761229] Lustre: Failing over lustre-MDT0000 [ 4508.779923] LustreError: 89841:0:(ldlm_resource.c:1172:ldlm_resource_complain()) lustre-MDT0001-osp-MDT0000: namespace resource [0x2400032e2:0x1a:0x0].0x0 (ffff9c5047beb100) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 4508.814520] LustreError: 89841:0:(ldlm_resource.c:1172:ldlm_resource_complain()) Skipped 1 previous similar message [ 4508.910460] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.16@tcp (stopping) [ 4508.917757] Lustre: Skipped 6 previous similar messages [ 4511.714120] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 4511.730443] LustreError: Skipped 4 previous similar messages [ 4515.125554] Lustre: server umount lustre-MDT0000 complete [ 4532.688908] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4532.719171] LustreError: Skipped 5 previous similar messages [ 4542.353919] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4542.357561] LDISKFS-fs (dm-0): recovery complete [ 4542.370573] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4542.955844] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c506f3d6300 x1863922738479744/t0(0) o250->MGC192.168.204.116@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4542.986354] LustreError: 3662:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 7 previous similar messages [ 4547.907113] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4552.331788] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2290 to 0x280000401:2305) [ 4552.336259] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2290 to 0x2c0000401:2305) [ 4556.673237] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4558.815131] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4618.803134] Lustre: DEBUG MARKER: == replay-dual test 28: lock replay should be ordered: waiting after granted ========================================================== 16:08:47 (1777579727) [ 4639.226441] Lustre: Failing over lustre-OST0000 [ 4639.719486] Lustre: lustre-OST0000: Not available for connect from 0@lo (stopping) [ 4639.727177] Lustre: Skipped 7 previous similar messages [ 4641.577909] Lustre: server umount lustre-OST0000 complete [ 4663.874213] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4666.318114] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4666.322113] Lustre: lustre-OST0000-osc-MDT0001: Connection restored to 0@lo (at 0@lo) [ 4666.334296] Lustre: Skipped 27 previous similar messages [ 4670.696433] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4678.871103] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4680.818580] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4691.657834] Lustre: DEBUG MARKER: == replay-dual test 29: replay vs update with the same xid ========================================================== 16:10:01 (1777579801) [ 4692.986633] Lustre: DEBUG MARKER: SKIP: replay-dual test_29 needs >= 2 clients [ 4694.791707] Lustre: DEBUG MARKER: == replay-dual test 30: layout lock replay is not blocked on IO ========================================================== 16:10:04 (1777579804) [ 4698.145215] Lustre: Failing over lustre-MDT0000 [ 4698.651530] Lustre: server umount lustre-MDT0000 complete [ 4718.681794] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4728.166257] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 4728.173896] Lustre: Skipped 7 previous similar messages [ 4732.345959] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4733.554035] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2334 to 0x280000401:2369) [ 4733.554867] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2332 to 0x2c0000401:2369) [ 4741.384962] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4743.593293] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4752.744452] Lustre: DEBUG MARKER: == replay-dual test 31: deadlock on file_remove_privs and occupied mod rpc slots ========================================================== 16:11:02 (1777579862) [ 4757.010372] Lustre: Failing over lustre-OST0000 [ 4757.153625] Lustre: server umount lustre-OST0000 complete [ 4775.708634] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4782.786191] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4790.655025] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4792.345380] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL [ 4803.582648] Lustre: DEBUG MARKER: == replay-dual test 32: gap in update llog shouldn't break recovery ========================================================== 16:11:53 (1777579913) [ 4804.878158] Lustre: *** cfs_fail_loc=131d, val=10*** [ 4805.494363] Lustre: *** cfs_fail_loc=131d, val=2*** [ 4805.497328] Lustre: Skipped 7 previous similar messages [ 4806.600281] Lustre: *** cfs_fail_loc=131d, val=4294967280*** [ 4806.609518] Lustre: Skipped 17 previous similar messages [ 4809.260123] Lustre: Failing over lustre-MDT0001 [ 4809.614477] Lustre: server umount lustre-MDT0001 complete [ 4813.919702] Lustre: Failing over lustre-MDT0000 [ 4815.889021] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.16@tcp (stopping) [ 4815.898342] Lustre: Skipped 5 previous similar messages [ 4816.370628] Lustre: server umount lustre-MDT0000 complete [ 4824.802600] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4825.505960] Lustre: *** cfs_fail_loc=131d, val=4294967266*** [ 4825.509303] Lustre: Skipped 13 previous similar messages [ 4830.718001] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4838.957853] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4839.094556] Lustre: *** cfs_fail_loc=131d, val=4294967262*** [ 4839.097099] Lustre: Skipped 3 previous similar messages [ 4844.398800] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4844.627401] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:353) [ 4844.628914] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:353) [ 4844.650052] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2409 to 0x280000401:2465) [ 4844.674342] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2332 to 0x2c0000401:2401) [ 4856.553687] Lustre: DEBUG MARKER: == replay-dual test 33: Check for OBD_INCOMPAT_MULTI_RPCS in last_rcvd after abort_recovery ========================================================== 16:12:46 (1777579966) [ 4864.099491] Lustre: Failing over lustre-MDT0001 [ 4864.768179] Lustre: server umount lustre-MDT0001 complete [ 4884.665835] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4889.067903] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4895.299948] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount REPLAY_WAIT mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4897.129454] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in REPLAY_WAIT state after 0 sec [ 4898.153794] Lustre: lustre-MDT0001: Aborting client recovery [ 4898.158859] LustreError: 99109:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 4898.164810] Lustre: 98588:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 4898.180738] Lustre: 98588:0:(ldlm_lib.c:2389:target_recovery_overseer()) Skipped 2 previous similar messages [ 4898.187742] Lustre: 98588:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 82f10074-fcd2-426c-871a-51c563f8b7cd@ [ 4898.206064] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 4898.232400] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 4898.259977] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 4898.363612] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:385) [ 4898.363681] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:385) [ 4902.818490] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4904.756775] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4909.497770] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4916.757941] Lustre: Failing over lustre-MDT0001 [ 4916.981561] Lustre: server umount lustre-MDT0001 complete [ 4918.752376] Lustre: lustre-MDT0001-osp-MDT0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4918.762326] Lustre: Skipped 28 previous similar messages [ 4924.757524] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4929.921687] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug -1 all [ 4930.582118] Lustre: lustre-MDT0001: Recovery over after 0:03, of 2 clients 2 recovered and 0 were evicted. [ 4930.603286] Lustre: Skipped 9 previous similar messages [ 4930.730768] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:417) [ 4930.731452] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:417) [ 4938.065409] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4940.106986] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4944.267352] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4955.622593] Lustre: DEBUG MARKER: == replay-dual test complete, duration 4629 sec ========== 16:14:25 (1777580065) [ 4958.107781] Lustre: DEBUG MARKER: === replay-dual: start cleanup 16:14:27 (1777580067) === [ 4971.416561] Lustre: DEBUG MARKER: === replay-dual: finish cleanup 16:14:41 (1777580081) === [ 4973.297089] Lustre: Failing over lustre-MDT0000 [ 4973.556167] Lustre: server umount lustre-MDT0000 complete [ 4981.737461] LustreError: 97903:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4981.760165] LustreError: 97903:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 220 previous similar messages [ 5000.691517] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5005.305371] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 5005.313263] Lustre: Skipped 10 previous similar messages [ 5006.665284] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 5145.500173] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 5145.503427] Lustre: 102030:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 82f10074-fcd2-426c-871a-51c563f8b7cd@ [ 5145.521447] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 5145.621921] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2409 to 0x280000401:2497) [ 5145.622925] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2332 to 0x2c0000401:2433) [ 5150.649821] Lustre: DEBUG MARKER: oleg416-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 5152.020646] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5159.906477] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 5159.926332] Lustre: Skipped 3 previous similar messages [ 5163.989269] Lustre: server umount lustre-MDT0000 complete [ 5172.033739] LustreError: 6505:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1777580284 with bad export cookie 8790466873432315533 [ 5172.046283] LustreError: MGC192.168.204.116@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 5172.047583] LustreError: 6505:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 5 previous similar messages [ 5172.061903] LustreError: Skipped 3 previous similar messages [ 5172.497693] Lustre: server umount lustre-MDT0001 complete [ 5192.968292] Lustre: server umount lustre-OST0000 complete [ 5211.821627] Lustre: server umount lustre-OST0001 complete [ 5230.899549] Lustre: DEBUG MARKER: oleg416-server.virtnet: executing unload_modules_local [ 5234.037080] Key type lgssc unregistered [ 5234.436680] LNet: 104839:0:(lib-ptl.c:967:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 5234.447549] LNetError: 104839:0:(acceptor.c:252:lnet_acceptor_remove_socket()) Interface ens2 not found [ 5234.469673] LNet: Removed LNI 192.168.204.116@tcp [ 5235.549194] Key type .llcrypt unregistered [ 5235.550608] Key type ._llcrypt unregistered