[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-8.fc42 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 490349981 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2400.000 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE2439 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE22D5 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 002295 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE2349 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE23D9 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2411 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe22d5-0xbffe2348] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe22d4] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe2349-0xbffe23d8] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe23d9-0xbffe2410] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2411-0xbffe2438] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524580K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001012] APIC: Switch to symmetric I/O mode setup [ 0.003322] x2apic enabled [ 0.004010] Switched APIC routing to physical x2apic. [ 0.005016] kvm-guest: setup PV IPIs [ 0.008316] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.009000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 0.009021] Calibrating delay loop (skipped) preset value.. 4800.00 BogoMIPS (lpj=2400000) [ 0.010015] pid_max: default: 32768 minimum: 301 [ 0.011138] LSM: Security Framework initializing [ 0.012071] Yama: becoming mindful. [ 0.013042] SELinux: Initializing. [ 0.014107] *** VALIDATE selinux *** [ 0.022496] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.027844] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.028156] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.029199] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.031102] *** VALIDATE tmpfs *** [ 0.032426] *** VALIDATE proc *** [ 0.034198] *** VALIDATE cgroup *** [ 0.035011] *** VALIDATE cgroup2 *** [ 0.036300] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.037176] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.039011] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.040032] Spectre V2 : User space: Vulnerable [ 0.041007] Speculative Store Bypass: Vulnerable [ 0.044079] debug: unmapping init [mem 0xffffffff97a59000-0xffffffff97a60fff] [ 0.046225] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.047740] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.048026] ... version: 2 [ 0.049018] ... bit width: 48 [ 0.050018] ... generic registers: 4 [ 0.051013] ... value mask: 0000ffffffffffff [ 0.052018] ... max period: 00007fffffffffff [ 0.053016] ... fixed-purpose events: 3 [ 0.054012] ... event mask: 000000070000000f [ 0.055328] rcu: Hierarchical SRCU implementation. [ 0.057579] smp: Bringing up secondary CPUs ... [ 0.058540] x86: Booting SMP configuration: [ 0.059022] .... node #0, CPUs: #1 #2 #3 [ 0.062534] smp: Brought up 1 node, 4 CPUs [ 0.064010] smpboot: Max logical packages: 1 [ 0.065017] smpboot: Total of 4 processors activated (19200.00 BogoMIPS) [ 0.149019] node 0 deferred pages initialised in 82ms [ 0.152512] devtmpfs: initialized [ 0.154295] x86/mm: Memory block size: 128MB [ 0.157933] gcov: version magic: 0x41383552 [ 0.162295] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.163085] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.164302] pinctrl core: initialized pinctrl subsystem [ 0.165215] [ 0.165808] ************************************************************* [ 0.166013] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.167011] ** ** [ 0.168014] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.169013] ** ** [ 0.170019] ** This means that this kernel is built to expose internal ** [ 0.171012] ** IOMMU data structures, which may compromise security on ** [ 0.172013] ** your system. ** [ 0.173014] ** ** [ 0.174011] ** If you see this message and you are not debugging the ** [ 0.175017] ** kernel, report this immediately to your vendor! ** [ 0.176013] ** ** [ 0.177014] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.178015] ************************************************************* [ 0.179846] NET: Registered protocol family 16 [ 0.181445] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.184069] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.187066] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.191473] cpuidle: using governor menu [ 0.193751] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.195467] PCI: Using configuration type 1 for base access [ 0.197171] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.206162] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.209025] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.212278] cryptd: max_cpu_qlen set to 1000 [ 0.215252] ACPI: Added _OSI(Module Device) [ 0.217029] ACPI: Added _OSI(Processor Device) [ 0.219015] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.221016] ACPI: Added _OSI(Processor Aggregator Device) [ 0.226273] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.232253] ACPI: Interpreter enabled [ 0.233040] ACPI: PM: (supports S0 S3 S4 S5) [ 0.235015] ACPI: Using IOAPIC for interrupt routing [ 0.236098] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.239406] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.249421] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.252043] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.255026] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.258081] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.264224] acpiphp: Slot [2] registered [ 0.266109] acpiphp: Slot [5] registered [ 0.267118] acpiphp: Slot [6] registered [ 0.269128] acpiphp: Slot [7] registered [ 0.270114] acpiphp: Slot [8] registered [ 0.271134] acpiphp: Slot [9] registered [ 0.273113] acpiphp: Slot [10] registered [ 0.274126] acpiphp: Slot [3] registered [ 0.276108] acpiphp: Slot [4] registered [ 0.278086] acpiphp: Slot [11] registered [ 0.279211] acpiphp: Slot [12] registered [ 0.280057] acpiphp: Slot [13] registered [ 0.281113] acpiphp: Slot [14] registered [ 0.283109] acpiphp: Slot [15] registered [ 0.285110] acpiphp: Slot [16] registered [ 0.286095] acpiphp: Slot [17] registered [ 0.287094] acpiphp: Slot [18] registered [ 0.289102] acpiphp: Slot [19] registered [ 0.290112] acpiphp: Slot [20] registered [ 0.292116] acpiphp: Slot [21] registered [ 0.293148] acpiphp: Slot [22] registered [ 0.295112] acpiphp: Slot [23] registered [ 0.297119] acpiphp: Slot [24] registered [ 0.299121] acpiphp: Slot [25] registered [ 0.301123] acpiphp: Slot [26] registered [ 0.303130] acpiphp: Slot [27] registered [ 0.305093] acpiphp: Slot [28] registered [ 0.307116] acpiphp: Slot [29] registered [ 0.308108] acpiphp: Slot [30] registered [ 0.310146] acpiphp: Slot [31] registered [ 0.312060] PCI host bridge to bus 0000:00 [ 0.314021] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.317013] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.319026] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.322022] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.325025] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.327025] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.329180] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.332066] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.335977] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.345015] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.349516] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.352016] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.355020] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.356029] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.358527] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.362936] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.365034] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.368745] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.375017] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.389016] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.394019] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.398466] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.405015] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.412023] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.433020] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.445372] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.451014] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.456014] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.471018] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.482181] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.487020] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.494019] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.518020] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.532011] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.539014] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.547014] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.567017] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.576937] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.584015] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.590018] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.619019] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.630260] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.638017] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.648012] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.667015] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.679461] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.681267] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.683375] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.686366] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.688202] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.692178] iommu: Default domain type: Passthrough [ 0.695450] SCSI subsystem initialized [ 0.697147] ACPI: bus type USB registered [ 0.698098] usbcore: registered new interface driver usbfs [ 0.700079] usbcore: registered new interface driver hub [ 0.701071] usbcore: registered new device driver usb [ 0.703174] pps_core: LinuxPPS API ver. 1 registered [ 0.705011] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.708063] PTP clock support registered [ 0.710129] EDAC MC: Ver: 3.0.0 [ 0.712119] PCI: Using ACPI for IRQ routing [ 0.713768] NetLabel: Initializing [ 0.715012] NetLabel: domain hash size = 128 [ 0.717009] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.719172] NetLabel: unlabeled traffic allowed by default [ 0.722155] vgaarb: loaded [ 0.723336] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.726015] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.732447] clocksource: Switched to clocksource kvm-clock [ 0.845472] VFS: Disk quotas dquot_6.6.0 [ 0.847191] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 0.849858] *** VALIDATE ramfs *** [ 0.851027] *** VALIDATE hugetlbfs *** [ 0.852806] pnp: PnP ACPI init [ 0.855709] pnp: PnP ACPI: found 6 devices [ 0.872172] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 0.875859] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 0.878021] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 0.880568] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 0.883361] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 0.886249] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 0.888839] NET: Registered protocol family 2 [ 0.890933] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 0.895150] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 0.898965] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 0.903631] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 0.906479] TCP: Hash tables configured (established 65536 bind 65536) [ 0.909446] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 0.911806] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.913827] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.916073] NET: Registered protocol family 1 [ 0.918218] RPC: Registered named UNIX socket transport module. [ 0.919648] RPC: Registered udp transport module. [ 0.921279] RPC: Registered tcp transport module. [ 0.922619] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 0.924435] NET: Registered protocol family 44 [ 0.926463] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 0.929052] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 0.931441] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 0.934376] PCI: CLS 0 bytes, default 64 [ 0.936291] Unpacking initramfs... [ 2.313560] debug: unmapping init [mem 0xffff9c873cc54000-0xffff9c873ffbffff] [ 2.317808] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 2.320518] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 2.323863] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 2.830268] Initialise system trusted keyrings [ 2.832242] Key type blacklist registered [ 2.834596] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 2.845358] zbud: loaded [ 2.848785] *** VALIDATE nfs *** [ 2.850198] *** VALIDATE nfs4 *** [ 2.852196] pstore: using deflate compression [ 2.856709] Platform Keyring initialized [ 2.973950] NET: Registered protocol family 38 [ 2.976126] Key type asymmetric registered [ 2.978043] Asymmetric key parser 'x509' registered [ 2.980124] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 2.984166] io scheduler mq-deadline registered [ 2.985702] io scheduler kyber registered [ 2.987122] io scheduler bfq registered [ 2.989079] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 2.991768] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 2.994301] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 2.996513] ACPI: Power Button [PWRF] [ 3.002475] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 3.010279] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 3.026048] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 3.033900] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 3.046081] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 3.073792] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 3.102503] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 3.106978] Non-volatile memory driver v1.3 [ 3.108937] Linux agpgart interface v0.103 [ 3.139045] virtio_blk virtio1: [vda] 134072 512-byte logical blocks (68.6 MB/65.5 MiB) [ 3.143696] vda: detected capacity change from 0 to 68644864 [ 3.161917] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 3.164985] vdb: detected capacity change from 0 to 1073741824 [ 3.181080] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.184733] vdc: detected capacity change from 0 to 2621440000 [ 3.201260] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.204700] vdd: detected capacity change from 0 to 2621440000 [ 3.219136] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.221524] vde: detected capacity change from 0 to 4294967296 [ 3.233474] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.236096] vdf: detected capacity change from 0 to 4294967296 [ 3.244890] libphy: Fixed MDIO Bus: probed [ 3.251059] usbcore: registered new interface driver usbserial_generic [ 3.253263] usbserial: USB Serial support registered for generic [ 3.255116] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 3.258750] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 3.260575] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 3.263310] mousedev: PS/2 mouse device common for all mice [ 3.266225] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 3.268649] rtc_cmos 00:05: RTC can wake from S4 [ 3.274340] rtc_cmos 00:05: registered as rtc0 [ 3.275588] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 3.277209] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 3.277261] intel_pstate: CPU model not supported [ 3.279744] hid: raw HID events driver (C) Jiri Kosina [ 3.288232] usbcore: registered new interface driver usbhid [ 3.288747] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 3.291476] usbhid: USB HID core driver [ 3.296135] drop_monitor: Initializing network drop monitor service [ 3.299233] Initializing XFRM netlink socket [ 3.301405] NET: Registered protocol family 10 [ 3.304592] Segment Routing with IPv6 [ 3.305856] NET: Registered protocol family 17 [ 3.307503] mpls_gso: MPLS GSO support [ 3.313289] RAS: Correctable Errors collector initialized. [ 3.315484] AVX version of gcm_enc/dec engaged. [ 3.316990] AES CTR mode by8 optimization enabled [ 3.391203] sched_clock: Marking stable (3391184041, 0)->(4346552792, -955368751) [ 3.394165] registered taskstats version 1 [ 3.395953] Loading compiled-in X.509 certificates [ 3.397758] zswap: loaded using pool lzo/zbud [ 3.421799] Key type big_key registered [ 3.435224] Key type encrypted registered [ 3.437067] ima: No TPM chip found, activating TPM-bypass! [ 3.439276] ima: Allocated hash algorithm: sha1 [ 3.440655] ima: No architecture policies found [ 3.442143] evm: Initialising EVM extended attributes: [ 3.443846] evm: security.selinux [ 3.445010] evm: security.ima [ 3.445744] evm: security.capability [ 3.446880] evm: HMAC attrs: 0x1 [ 3.449094] rtc_cmos 00:05: setting system clock to 2026-05-12 19:50:39 UTC (1778615439) [ 3.454549] debug: unmapping init [mem 0xffffffff98a03000-0xffffffff98bfffff] [ 3.457341] debug: unmapping init [mem 0xffffffff97782000-0xffffffff97a58fff] [ 3.466110] Write protecting the kernel read-only data: 28672k [ 3.470457] debug: unmapping init [mem 0xffffffff95e03000-0xffffffff95ffffff] [ 3.473667] debug: unmapping init [mem 0xffffffff96714000-0xffffffff967fffff] [ 3.505786] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 3.515251] systemd[1]: Detected virtualization kvm. [ 3.517435] systemd[1]: Detected architecture x86-64. [ 3.519678] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 3.550721] systemd[1]: No hostname configured. [ 3.552132] systemd[1]: Set hostname to . [ 3.554549] random: systemd: uninitialized urandom read (16 bytes read) [ 3.556413] systemd[1]: Initializing machine ID from random generator. [ 3.593786] random: ln: uninitialized urandom read (6 bytes read) [ 3.677621] random: systemd: uninitialized urandom read (16 bytes read) [ 3.681303] systemd[1]: Started Dispatch Password Requests to Console Directory Watch. [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ 3.689386] systemd[1]: Reached target Paths. [ OK ] Reached target Paths. [ 3.694254] systemd[1]: Reached target Slices. [ OK ] Reached target Slices. [ OK ] Reached target Timers. [ OK ] Listening on Journal Socket. Starting Setup Virtual Console... Starting Create list of required st…ce nodes for the current kernel... [ OK ] Started Memstrack Anylazing Service. [ OK ] Listening on Journal Socket (/dev/log). Starting Journal Service... [ OK ] Listening on udev Kernel Socket. [ OK ] Reached target Initrd Root Device. [ OK ] Reached target Swap. [ OK ] Reached target Local Encrypted Volumes. Starting Apply Kernel Variables... [ OK ] Listening on udev Control Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Local File Systems. Starting Create Volatile Files and Directories... [ OK ] Started Setup Virtual Console. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. [ OK ] Started Create Volatile Files and Directories. Starting Create Static Device Nodes in /dev... Starting dracut cmdline hook... [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started Journal Service. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 4.269715] device-mapper: uevent: version 1.0.3 [ 4.271826] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. Starting dracut initqueue hook... [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ 4.952242] random: fast init done [ 4.953810] virtio_net virtio0 ens2: renamed from eth0 [ 5.032132] scsi host0: ata_piix [ 5.053217] scsi host1: ata_piix [ 5.055216] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 5.057876] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 8.773739] dracut-initqueue[576]: RTNETLINK answers: File exists [ 9.952416] random: crng init done [ 9.954139] random: 7 urandom warning(s) missed due to ratelimiting Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. Mounting /sysroot... [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. [ 10.429501] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... [ OK ] Stopped target Timers. [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Initrd Root Device. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Sockets. [ OK ] Stopped target Slices. [ OK ] Stopped target Paths. [ OK ] Stopped target System Initialization. [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Local File Systems. [ OK ] Stopped target Swap. [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. Stopping udev Kernel Device Manager... [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Closed udev Kernel Socket. [ OK ] Closed udev Control Socket. Starting Cleanup udevd DB... [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 11.618623] printk: systemd: 26 output lines suppressed due to ratelimiting [ 11.927662] SELinux: Disabled at runtime. [ 11.995558] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 12.003031] systemd[1]: Detected virtualization kvm. [ 12.004912] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 12.557974] systemd[1]: initrd-switch-root.service: Succeeded. [ 12.561588] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 12.567520] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 12.571581] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 12.577396] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 12.584912] systemd[1]: Starting Journal Service... Starting Journal Service... [ 12.590152] systemd[1]: Created slice system-serial\x2dgetty.slice. [ OK ] Created slice system-serial\x2dgetty.slice. [ OK ] Listening on initctl Compatibility Named Pipe. [ OK ] Reached target rpc_pipefs.target. [ OK ] Stopped target Switch Root. [ OK ] Stopped target Initrd Root File System. [ OK ] Listening on udev Kernel Socket. [ OK ] Created slice User and Session Slice. [ OK ] Reached target Slices. Activating swap /dev/disk/by-label/SWAP... [ OK ] Listening on udev Control Socket. Starting udev Coldplug all Devices... [ OK ] Stopped target Initrd File Systems. Starting Create list of required st…ce nodes for the current kernel... [ 12.667495] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS [ OK ] Listening on Process Core Dump Socket. [ OK ] Created slice system-getty.slice. Starting Apply Kernel Variables... [ OK ] Started Forward Password Requests to Wall Directory Watch. [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. Mounting Kernel Debug File System... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Reached target Paths. [ OK ] Created slice system-sshd\x2dkeygen.slice. Mounting POSIX Message Queue File System... [ OK ] Listening on RPCbind Server Activation Socket. [ OK ] Reached target RPC Port Mapper. Mounting Huge Pages File System... Starting Remount Root and Kernel File Systems... [ OK ] Started Journal Service. [ OK ] Activated swap /dev/disk/by-label/SWAP. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. [ OK ] Mounted Kernel Debug File System. [ OK ] Mounted POSIX Message Queue File System. [ OK ] Mounted Huge Pages File System. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... [ OK ] Reached target Swap. Starting Flush Journal to Persistent Storage... [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Reached target Local File Systems (Pre). Mounting /home/green/git/lustre-release... Mounting /mnt... Starting udev Kernel Device Manager... [ OK ] Started udev Coldplug all Devices. [ OK ] Mounted /mnt. [ 13.064663] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Kernel Device Manager. [ 13.365846] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 13.385221] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 13.518375] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 13.537277] EDAC sbridge: Ver: 1.1.2 [ 15.446614] Key type dns_resolver registered [ 15.763149] NFS: Registering the id_resolver key type [ 15.765659] Key type id_resolver registered [ 15.767531] Key type id_legacy registered [ OK ] Started Configure read-only root support. [ OK ] Reached target Local File Systems. Starting Rebuild Dynamic Linker Cache... Starting Mark the need to relabel after reboot... Starting Load/Save Random Seed... Starting Create Volatile Files and Directories... [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Load/Save Random Seed. [ OK ] Started Create Volatile Files and Directories. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started dnf makecache --timer. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. [ OK ] Started irqbalance daemon. Starting Restore /run/initramfs on shutdown... [ OK ] Started D-Bus System Message Bus. [ OK ] Reached target sshd-keygen.target. [ OK ] Started daily update of the root trust anchor for DNSSEC. Starting Login Service... Starting Network Manager... [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Reached target Timers. [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Network Manager. [ OK ] Reached target Network. Starting OpenSSH server daemon... Starting Dynamic System Tuning Daemon... Starting GSSAPI Proxy Daemon... Starting Network Manager Wait Online... [ OK ] Started Login Service. [ OK ] Started OpenSSH server daemon. Starting Hostname Service... [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... [ OK ] Started Permit User Sessions. [ OK ] Started Command Scheduler. [ OK ] Started Getty on tty1. [ OK ] Started Serial Getty on ttyS0. [ OK ] Started Serial Getty on ttyS1. [ OK ] Reached target Login Prompts. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting System Logging Service... Starting Notify NFS peers of a restart... Starting Crash recovery kernel arming... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Starting Authorization Manager... [ OK ] Started Dynamic System Tuning Daemon. [ OK ] Reached target Multi-User System. [ OK ] Reached target Graphical Interface. Starting Update UTMP about System Runlevel Changes... [ OK ] Started Authorization Manager. [ OK ] Started Update UTMP about System Runlevel Changes. Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg403-server login: [ 42.874508] libcfs: loading out-of-tree module taints kernel. [ 42.892331] Key type ._llcrypt registered [ 42.894342] Key type .llcrypt registered [ 42.949914] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_hostid [ 61.337411] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing load_modules_local [ 63.863756] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 63.882934] alg: No test for adler32 (adler32-zlib) [ 65.773256] Lustre: Lustre: Build Version: 2.17.52_126_gcf32fc9 [ 67.256343] LNet: Added LNI 192.168.204.103@tcp [8/256/0/180] [ 69.103898] Key type lgssc registered [ 69.453330] hrtimer: interrupt took 8767254 ns [ 72.240579] Lustre: Echo OBD driver; http://www.lustre.org/ [ 92.826117] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 144.442381] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing load_modules_local [ 161.447806] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 161.498800] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 162.869441] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 162.918247] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 163.083671] Lustre: lustre-MDT0000: new disk, initializing [ 163.214092] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 163.261802] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 168.681992] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 186.325380] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 186.495317] Lustre: 6502:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 186.536963] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 186.547876] Lustre: Skipped 1 previous similar message [ 186.728794] Lustre: lustre-MDT0001: new disk, initializing [ 186.853279] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 186.921949] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 186.943379] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 193.099572] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 199.153309] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 210.333459] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 210.589959] Lustre: lustre-OST0000: new disk, initializing [ 210.596375] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 210.696585] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 218.172217] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 218.189325] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 218.331024] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 218.351512] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 237.365212] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 237.619482] Lustre: lustre-OST0001: new disk, initializing [ 237.626123] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 237.789367] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 245.304537] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 245.316252] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 245.456101] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 246.378134] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 261.425358] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 269.866744] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 277.358381] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing check_logdir /tmp/testlogs/ [ 283.603373] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing yml_node [ 287.822129] Lustre: DEBUG MARKER: Client: 2.17.52.126 [ 290.122615] Lustre: DEBUG MARKER: MDS: 2.17.52.126 [ 293.114362] Lustre: DEBUG MARKER: OSS: 2.17.52.126 [ 295.245486] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-dual ============----- Tue May 12 15:55:30 EDT 2026 [ 317.735758] Lustre: DEBUG MARKER: excepting tests: 14b 21b [ 319.454147] Lustre: DEBUG MARKER: skipping tests SLOW=no: 21b [ 321.012664] Lustre: DEBUG MARKER: === replay-dual: start setup 15:55:56 (1778615756) === [ 326.278767] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing check_config_client /mnt/lustre [ 351.460192] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 356.259654] Lustre: 13226:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 361.736192] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 368.652168] Lustre: DEBUG MARKER: === replay-dual: finish setup 15:56:43 (1778615803) === [ 371.990971] Lustre: DEBUG MARKER: == replay-dual test 0a: expired recovery with lost client ========================================================== 15:56:46 (1778615806) [ 384.261561] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 391.268322] Lustre: Failing over lustre-MDT0000 [ 391.660682] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 391.694820] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 391.723992] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 392.031889] Lustre: server umount lustre-MDT0000 complete [ 394.237452] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 394.272732] Lustre: Skipped 2 previous similar messages [ 398.612680] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 398.650526] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 8 previous similar messages [ 399.332192] LustreError: 6509:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 399.378169] LustreError: 6509:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 3 previous similar messages [ 403.711412] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 403.749191] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 1 previous similar message [ 408.850601] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 408.890064] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 5 previous similar messages [ 410.591650] Lustre: 3651:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778615830/real 1778615830] req@ffff9c87bf713800 x1865013529806720/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778615846 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 410.620063] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 413.890575] LustreError: 8412:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 413.921176] LustreError: 8412:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 11 previous similar messages [ 419.690366] LDISKFS-fs (dm-0): 10 truncates cleaned up [ 419.707219] LDISKFS-fs (dm-0): recovery complete [ 419.769195] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 421.265034] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 424.192687] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 426.506415] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 427.716403] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 529.500783] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 529.504533] Lustre: 14754:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 776cba71-9d02-4594-bb17-028c37cada47@192.168.204.3@tcp [ 529.516431] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 529.547146] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 529.555043] Lustre: 14754:0:(ldlm_lib.c:2933:target_recovery_thread()) too long recovery - read logs [ 529.571437] Lustre: Skipped 2 previous similar messages [ 529.591804] LustreError: dumping log to /tmp/lustre-log.1778615965.14754 [ 529.942218] Lustre: lustre-MDT0000: Recovery over after 1:45, of 3 clients 2 recovered and 1 was evicted. [ 530.006756] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:65) [ 530.010886] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:65) [ 553.797292] Lustre: DEBUG MARKER: == replay-dual test 0b: lost client during waiting for next transno ========================================================== 15:59:49 (1778615989) [ 564.804571] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 567.953408] Lustre: Failing over lustre-MDT0000 [ 568.100886] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (stopping) [ 568.432741] Lustre: server umount lustre-MDT0000 complete [ 569.829757] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 569.848517] LustreError: 7711:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 569.860983] Lustre: Skipped 1 previous similar message [ 569.887931] LustreError: 7711:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 15 previous similar messages [ 585.680157] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616005/real 1778616005] req@ffff9c86831ea680 x1865013529891456/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616021 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 585.724459] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 588.541513] LustreError: 15123:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 588.563470] LustreError: 15123:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 16 previous similar messages [ 593.228267] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 593.231307] LDISKFS-fs (dm-0): recovery complete [ 593.244471] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 595.423212] LustreError: 16443:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 595.438632] LustreError: 16443:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c86831e9880 x1865013529897344/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1778616031 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 595.482810] LustreError: 16443:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 595.939674] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c8789e59f80 x1865013529899008/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 596.394168] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 599.819860] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 601.391967] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 601.587246] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 614.621101] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:54 [ 619.797939] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:49 [ 624.910254] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:44 [ 627.182395] Lustre: lustre-MDT0001: haven't heard from client 776cba71-9d02-4594-bb17-028c37cada47 (at 192.168.204.3@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff9c8682cf7800, cur 1778616063 deadline 1778616062 last 1778615962 [ 630.020230] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:39 [ 635.134164] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:34 [ 645.397213] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:24 [ 645.413510] Lustre: Skipped 1 previous similar message [ 665.862260] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:03 [ 665.885967] Lustre: Skipped 3 previous similar messages [ 669.503102] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 669.510257] Lustre: 16477:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 28d3d4fe-9369-4428-8285-8b008976e07b@ [ 669.530472] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 701.700578] Lustre: lustre-MDT0001: haven't heard from client a121c1b4-aba3-4bb1-ae93-15d43d6e83b8 (at 192.168.204.3@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff9c8683153000, cur 1778616137 deadline 1778616135 last 1778616036 [ 701.701176] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 1:08 [ 701.758296] Lustre: Skipped 6 previous similar messages [ 768.260255] Lustre: lustre-MDT0000: Denying connection for new client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 0:02 [ 768.277937] Lustre: Skipped 12 previous similar messages [ 770.500149] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 770.510643] Lustre: 16477:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client a121c1b4-aba3-4bb1-ae93-15d43d6e83b8@192.168.204.3@tcp [ 770.520780] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 770.529775] Lustre: 16477:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 770.554639] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 770.556657] Lustre: 16477:0:(ldlm_lib.c:2933:target_recovery_thread()) too long recovery - read logs [ 770.561711] Lustre: Skipped 2 previous similar messages [ 770.576912] LustreError: dumping log to /tmp/lustre-log.1778616206.16477 [ 770.690142] Lustre: lustre-MDT0000: Recovery over after 2:51, of 3 clients 1 recovered and 2 were evicted. [ 770.720631] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:97) [ 770.728472] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:67 to 0x280000401:97) [ 783.020498] Lustre: DEBUG MARKER: == replay-dual test 1: |X| simple create ================= 16:03:38 (1778616218) [ 792.786947] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 795.399789] Lustre: Failing over lustre-MDT0000 [ 795.806076] Lustre: server umount lustre-MDT0000 complete [ 796.128333] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 796.142975] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 796.142987] Lustre: Skipped 2 previous similar messages [ 796.149161] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 796.231675] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 14 previous similar messages [ 812.511126] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616232/real 1778616232] req@ffff9c8685327b80 x1865013529985536/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616248 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 812.536061] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 821.744668] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 821.746908] LDISKFS-fs (dm-0): recovery complete [ 821.758171] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 823.201395] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 823.286759] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 825.094574] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 828.423923] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 828.651679] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 828.703582] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:129) [ 828.708880] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:129) [ 829.088825] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 838.997981] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 841.150674] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 851.938747] Lustre: DEBUG MARKER: == replay-dual test 2: |X| mkdir adir ==================== 16:04:47 (1778616287) [ 861.964674] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 864.361295] Lustre: Failing over lustre-MDT0000 [ 864.741836] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 864.755928] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 864.777022] Lustre: Skipped 3 previous similar messages [ 864.784041] LustreError: 6513:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 864.815917] LustreError: 6513:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 38 previous similar messages [ 866.682986] Lustre: server umount lustre-MDT0000 complete [ 885.728581] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616305/real 1778616305] req@ffff9c8790a5d180 x1865013530026368/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616321 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 885.755513] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 891.943495] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 891.949049] LDISKFS-fs (dm-0): recovery complete [ 891.964489] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 896.362854] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 896.412270] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 897.793938] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 901.112844] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 901.648365] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 901.667193] Lustre: Skipped 3 previous similar messages [ 901.776381] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 901.839246] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:161) [ 901.844323] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:161) [ 910.672172] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 913.030933] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 923.714755] Lustre: DEBUG MARKER: == replay-dual test 3: |X| mkdir adir, mkdir adir/bdir === 16:05:58 (1778616358) [ 934.090756] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 936.781302] Lustre: Failing over lustre-MDT0000 [ 937.174149] Lustre: server umount lustre-MDT0000 complete [ 937.455054] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 937.478176] Lustre: Skipped 6 previous similar messages [ 953.825839] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616373/real 1778616373] req@ffff9c8789e04380 x1865013530070016/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616389 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 953.861652] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 961.141944] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 961.145536] LDISKFS-fs (dm-0): recovery complete [ 961.162256] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 963.042312] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c8789e58380 x1865013530078208/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 963.424941] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 963.502091] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 965.390061] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 968.706604] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 968.717463] Lustre: Skipped 3 previous similar messages [ 968.898932] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 968.958786] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 968.989351] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:193) [ 968.991685] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:193) [ 976.732813] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 978.988724] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 989.759258] Lustre: DEBUG MARKER: == replay-dual test 4: |X| mkdir adir (-EEXIST), mkdir adir/bdir ========================================================== 16:07:05 (1778616425) [ 999.689341] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1002.225973] Lustre: Failing over lustre-MDT0000 [ 1002.521975] Lustre: server umount lustre-MDT0000 complete [ 1004.521447] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1004.524199] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1004.546328] Lustre: Skipped 3 previous similar messages [ 1004.589674] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 85 previous similar messages [ 1020.896897] Lustre: 3650:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616440/real 1778616440] req@ffff9c87c1aff800 x1865013530115328/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616456 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1020.918978] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1025.613914] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 1025.621976] LDISKFS-fs (dm-0): recovery complete [ 1025.650964] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1030.113327] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c87c1afce00 x1865013530123904/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1030.578848] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1031.946094] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1035.769589] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1035.773494] Lustre: Skipped 3 previous similar messages [ 1035.884197] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1035.943489] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:225) [ 1035.947106] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:225) [ 1036.234631] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1044.637413] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1046.782480] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1057.665189] Lustre: DEBUG MARKER: == replay-dual test 5: open, unlink |X| close ============ 16:08:12 (1778616492) [ 1066.431433] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1068.727395] Lustre: Failing over lustre-MDT0000 [ 1069.256943] Lustre: server umount lustre-MDT0000 complete [ 1071.592177] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1071.610813] Lustre: Skipped 3 previous similar messages [ 1087.967808] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616507/real 1778616507] req@ffff9c8789e15f80 x1865013530155776/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616523 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1088.028546] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1091.961319] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1091.964697] LDISKFS-fs (dm-0): recovery complete [ 1091.974311] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1098.512404] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1098.515577] Lustre: Skipped 1 previous similar message [ 1098.578107] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1100.373197] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1103.830447] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1103.867153] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1103.871169] Lustre: Skipped 3 previous similar messages [ 1103.978459] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 1104.056017] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:257) [ 1104.056717] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:257) [ 1112.397764] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1113.916869] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1123.134292] Lustre: DEBUG MARKER: == replay-dual test 6: open1, open2, unlink |X| close1 [fail mds1] close2 ========================================================== 16:09:18 (1778616558) [ 1132.164479] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1134.270177] Lustre: Failing over lustre-MDT0000 [ 1134.630841] Lustre: server umount lustre-MDT0000 complete [ 1156.067531] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616575/real 1778616575] req@ffff9c87b46af100 x1865013530193408/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616591 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1156.117267] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1161.999346] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1162.002068] LDISKFS-fs (dm-0): recovery complete [ 1162.026096] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1166.317151] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec3994d [ 1166.883420] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1167.110237] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1172.061608] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1172.132572] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:289) [ 1172.133407] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:289) [ 1172.292695] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1181.466681] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1183.248822] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1193.972877] Lustre: DEBUG MARKER: == replay-dual test 8: replay of resent request ========== 16:10:28 (1778616628) [ 1203.362635] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1204.432439] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1204.438352] LustreError: 8412:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87c1affb80 x1865013507944448/t38654705670(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:486/0 lens 512/448 e 0 to 0 dl 1778616651 ref 1 fl Interpret:/200/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1220.903764] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnecting [ 1220.944067] Lustre: 6508:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c8789452300 x1865013507944448/t38654705670(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:502/0 lens 512/2880 e 0 to 0 dl 1778616667 ref 1 fl Interpret:/202/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1224.759744] Lustre: Failing over lustre-MDT0000 [ 1225.066641] Lustre: server umount lustre-MDT0000 complete [ 1228.267696] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1228.289149] Lustre: Skipped 4 previous similar messages [ 1244.639120] Lustre: 3651:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616664/real 1778616664] req@ffff9c878418ca80 x1865013530240640/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616680 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1244.678746] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1248.144867] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1248.147046] LDISKFS-fs (dm-0): recovery complete [ 1248.162218] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1255.429200] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1256.723885] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1260.200865] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1260.554418] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1260.564226] Lustre: Skipped 8 previous similar messages [ 1260.639121] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1260.681714] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:321) [ 1260.682038] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:321) [ 1268.509819] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1270.278490] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1279.658814] Lustre: DEBUG MARKER: == replay-dual test 9: resending a replayed create ======= 16:11:55 (1778616715) [ 1288.469324] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1292.656258] Lustre: Failing over lustre-MDT0000 [ 1293.035525] Lustre: server umount lustre-MDT0000 complete [ 1296.359901] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1296.401360] LustreError: 6507:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 167 previous similar messages [ 1318.664507] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1318.666676] LDISKFS-fs (dm-0): recovery complete [ 1318.689364] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1322.049153] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec3a4d0 [ 1322.637880] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1327.709075] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1327.717253] LustreError: 31175:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b467a300 x1865013507963264/t42949672962(42949672962) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:605/0 lens 528/448 e 0 to 0 dl 1778616770 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1328.717297] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1339.690246] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1339.887701] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:353) [ 1339.892996] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:353) [ 1344.076963] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1346.549506] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1357.984956] Lustre: DEBUG MARKER: == replay-dual test 10: resending a replayed unlink ====== 16:13:13 (1778616793) [ 1366.844319] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1370.565689] Lustre: Failing over lustre-MDT0000 [ 1370.600490] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1370.608457] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1370.622083] Lustre: Skipped 5 previous similar messages [ 1370.627773] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 1370.822135] Lustre: server umount lustre-MDT0000 complete [ 1389.029466] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778616809/real 1778616809] req@ffff9c87bf712300 x1865013530320896/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778616825 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1389.072178] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1389.079301] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1389.093108] LustreError: Skipped 1 previous similar message [ 1395.464657] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1395.466556] LDISKFS-fs (dm-0): recovery complete [ 1395.475191] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1399.269175] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c879ebf7800 x1865013530329344/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1399.616802] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1399.622942] Lustre: Skipped 3 previous similar messages [ 1399.688883] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1400.079985] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1400.090786] Lustre: Skipped 1 previous similar message [ 1404.886273] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1404.950557] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1404.954745] LustreError: 33096:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b45c2680 x1865013507983232/t47244640260(47244640260) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:682/0 lens 528/448 e 0 to 0 dl 1778616847 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1416.462556] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnected, waiting for 3 clients in recovery for 1:29 [ 1416.574407] Lustre: lustre-MDT0000: Recovery over after 0:16, of 3 clients 3 recovered and 0 were evicted. [ 1416.587444] Lustre: Skipped 1 previous similar message [ 1416.644205] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:385) [ 1416.646589] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:385) [ 1420.923536] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1423.289976] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1437.150453] Lustre: DEBUG MARKER: == replay-dual test 11: both clients timeout during replay ========================================================== 16:14:31 (1778616871) [ 1447.531186] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1451.511562] Lustre: Failing over lustre-MDT0000 [ 1451.938295] Lustre: server umount lustre-MDT0000 complete [ 1452.514112] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1475.088221] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1475.094788] LDISKFS-fs (dm-0): recovery complete [ 1475.128437] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1482.724296] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c87a6709c00 x1865013530371712/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1488.318724] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1488.421556] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1488.430914] LustreError: 35021:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87844a4000 x1865013508006272/t51539607558(51539607558) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:10/0 lens 520/448 e 0 to 0 dl 1778616930 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1494.603994] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1499.930477] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1500.063935] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:417) [ 1500.069413] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:417) [ 1501.801206] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 5 sec [ 1510.237886] Lustre: DEBUG MARKER: == replay-dual test 12: open resend timeout ============== 16:15:45 (1778616945) [ 1519.179452] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1522.309904] Lustre: Failing over lustre-MDT0000 [ 1522.563605] Lustre: server umount lustre-MDT0000 complete [ 1546.211814] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1546.214180] LDISKFS-fs (dm-0): recovery complete [ 1546.221704] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1551.259370] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1551.277297] Lustre: Skipped 1 previous similar message [ 1556.469946] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1556.488040] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1556.488449] Lustre: Skipped 16 previous similar messages [ 1556.556747] Lustre: *** cfs_fail_loc=302, val=2147483648*** [ 1572.150051] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnected, waiting for 3 clients in recovery for 1:25 [ 1572.328358] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:449) [ 1572.334600] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:449) [ 1580.885663] Lustre: DEBUG MARKER: == replay-dual test 13: close resend timeout ============= 16:16:56 (1778617016) [ 1589.863670] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1592.854501] Lustre: Failing over lustre-MDT0000 [ 1593.331256] Lustre: server umount lustre-MDT0000 complete [ 1616.559251] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1616.561603] LDISKFS-fs (dm-0): recovery complete [ 1616.571333] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1623.075107] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c87a6709180 x1865013530448000/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1623.299834] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (not set up) [ 1623.315251] Lustre: Skipped 1 previous similar message [ 1627.545713] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1628.758119] Lustre: *** cfs_fail_loc=115, val=2147483648*** [ 1644.322532] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnected, waiting for 3 clients in recovery for 1:25 [ 1644.423921] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:481) [ 1644.424894] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:481) [ 1652.726296] Lustre: DEBUG MARKER: SKIP: replay-dual test_14b skipping ALWAYS excluded test 14b [ 1654.574915] Lustre: DEBUG MARKER: == replay-dual test 15a: timeout waiting for lost client during replay, 1 client completes ========================================================== 16:18:10 (1778617090) [ 1662.802618] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1666.360558] Lustre: Failing over lustre-MDT0000 [ 1666.843952] Lustre: server umount lustre-MDT0000 complete [ 1669.612357] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1669.635249] Lustre: Skipped 18 previous similar messages [ 1684.967979] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778617105/real 1778617105] req@ffff9c87a67caa00 x1865013530477568/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778617121 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1685.012055] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 1685.021543] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1685.040262] LustreError: Skipped 3 previous similar messages [ 1691.643994] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1691.649072] LDISKFS-fs (dm-0): recovery complete [ 1691.670156] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1695.202850] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c87b4164000 x1865013530486656/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1696.689623] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1696.695753] Lustre: Skipped 3 previous similar messages [ 1701.402558] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 1766.500205] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1766.503671] Lustre: 40541:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client e6e9a735-8b94-40f6-808d-3a0c72b3e971@ [ 1766.524389] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1767.140041] Lustre: lustre-MDT0000: Recovery over after 1:11, of 3 clients 2 recovered and 1 was evicted. [ 1767.155045] Lustre: Skipped 3 previous similar messages [ 1767.193654] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:513) [ 1767.201525] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:513) [ 1772.115041] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1773.813453] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1784.360150] Lustre: DEBUG MARKER: == replay-dual test 15c: remove multiple OST orphans ===== 16:20:19 (1778617219) [ 1793.431454] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1932.564584] Lustre: Failing over lustre-MDT0000 [ 1933.227081] Lustre: server umount lustre-MDT0000 complete [ 1936.147433] LustreError: 8412:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1936.171956] LustreError: 8412:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 224 previous similar messages [ 1936.358148] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1958.325551] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1958.330061] LDISKFS-fs (dm-0): recovery complete [ 1958.342872] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1961.954808] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c87afe6c700 x1865013530614016/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1962.489368] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1962.497481] Lustre: Skipped 4 previous similar messages [ 1962.588407] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1962.610618] Lustre: Skipped 2 previous similar messages [ 1967.925904] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2032.500323] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2032.522702] Lustre: 42456:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 1999c2f8-c82d-4f78-9ead-edee47d03e84@ [ 2032.542463] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2032.659965] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:1537) [ 2032.663070] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:1537) [ 2038.003933] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2039.737601] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2049.497232] Lustre: DEBUG MARKER: == replay-dual test 16: fail MDS during recovery (3571) == 16:24:44 (1778617484) [ 2060.006828] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2064.320537] Lustre: Failing over lustre-MDT0000 [ 2064.739398] Lustre: server umount lustre-MDT0000 complete [ 2090.543942] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2090.548755] LDISKFS-fs (dm-0): recovery complete [ 2090.573436] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2101.991085] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2102.282230] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2102.301892] Lustre: Skipped 15 previous similar messages [ 2128.682625] Lustre: Failing over lustre-MDT0000 [ 2128.771491] LustreError: 44773:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 2128.789611] Lustre: 44312:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2128.833137] Lustre: 44312:0:(ldlm_lib.c:1899:abort_req_replay_queue()) @@@ aborted: req@ffff9c878418d500 x1865013510488064/t0(73014444033) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:649/0 lens 528/0 e 2 to 0 dl 1778617569 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 2128.893420] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 2128.903503] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (stopping) [ 2128.959627] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 2128.980965] LustreError: 44312:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9c87b47b9500 x1865013530699904/t0(0) o700->lustre-MDT0001-osp-MDT0000@0@lo:30/10 lens 264/248 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'tgt_recover_0.0' uid:0 gid:0 projid:4294967295 [ 2129.004785] LustreError: 44312:0:(fid_request.c:213:seq_client_alloc_seq()) cli-cli-lustre-MDT0001-osp-MDT0000: Cannot allocate new meta-sequence: rc = -5 [ 2129.029144] LustreError: 44312:0:(fid_request.c:316:seq_client_alloc_fid()) cli-cli-lustre-MDT0001-osp-MDT0000: Can't allocate new sequence: rc = -5 [ 2129.713510] Lustre: server umount lustre-MDT0000 complete [ 2151.574277] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2159.627136] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c868bfeb800 x1865013530709376/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2165.610763] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2230.500657] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2230.512439] Lustre: 45231:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client cb40928f-8c53-49ce-94a9-cf6d6c7fba4f@ [ 2230.532641] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2231.396511] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1551 to 0x280000401:1569) [ 2231.397301] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1550 to 0x2c0000401:1569) [ 2236.248777] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2238.133613] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2248.789604] Lustre: DEBUG MARKER: == replay-dual test 17: fail OST during recovery (3571) == 16:28:04 (1778617684) [ 2259.512102] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 2262.436269] Lustre: Failing over lustre-OST0000 [ 2262.517763] Lustre: lustre-OST0000-osc-MDT0000: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2262.544757] Lustre: Skipped 11 previous similar messages [ 2262.579051] Lustre: lustre-OST0000: Not available for connect from 0@lo (stopping) [ 2262.587670] Lustre: Skipped 1 previous similar message [ 2262.681803] Lustre: server umount lustre-OST0000 complete [ 2290.820716] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 2290.823480] LDISKFS-fs (dm-2): recovery complete [ 2290.835559] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2292.392388] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 2292.399936] Lustre: Skipped 3 previous similar messages [ 2298.941460] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2325.448817] Lustre: Failing over lustre-OST0000 [ 2325.459225] LustreError: 47621:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 2325.470965] Lustre: 47074:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2325.480277] Lustre: 47074:0:(ldlm_lib.c:2389:target_recovery_overseer()) Skipped 2 previous similar messages [ 2325.494680] LustreError: 47074:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 2325.517238] Lustre: lustre-OST0000: Recovery over after 0:33, of 4 clients 0 recovered and 4 were evicted. [ 2325.529497] Lustre: Skipped 3 previous similar messages [ 2325.737870] Lustre: server umount lustre-OST0000 complete [ 2344.927439] Lustre: 3648:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778617728/real 1778617728] req@ffff9c8684ca4380 x1865013530779136/t0(0) o400->lustre-OST0000-osc-MDT0001@0@lo:28/4 lens 224/224 e 3 to 1 dl 1778617780 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 2344.966648] Lustre: 3648:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2347.429422] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2355.464701] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2419.500177] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 2419.514063] Lustre: 48061:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client 3ef78c66-e0ef-4f21-9177-5ad03044138c@ [ 2419.538814] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 2424.942688] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 2427.397220] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 2439.179745] Lustre: DEBUG MARKER: == replay-dual test 18: ldlm_handle_enqueue succeeds on evicted export (3822) ========================================================== 16:31:14 (1778617874) [ 2445.000426] LustreError: 14343:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b sleeping for 40000ms [ 2485.079147] LustreError: 14343:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b awake [ 2502.015726] Lustre: DEBUG MARKER: == replay-dual test 19: resend of open request =========== 16:32:17 (1778617937) [ 2510.897406] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2512.523669] Lustre: *** cfs_fail_loc=157, val=2147483648*** [ 2512.538422] LustreError: 14343:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b47b7b80 x1865013510616192/t0(0) o101->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:355/0 lens 576/688 e 0 to 0 dl 1778618030 ref 1 fl Interpret:/600/0 rc 0/0 job:'createmany.0' uid:0 gid:0 projid:0 [ 2599.169111] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnecting [ 2602.906613] Lustre: Failing over lustre-MDT0000 [ 2603.425485] Lustre: server umount lustre-MDT0000 complete [ 2604.318835] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2604.360136] LustreError: 6508:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 124 previous similar messages [ 2605.027752] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2621.413095] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2621.427597] LustreError: Skipped 3 previous similar messages [ 2628.936109] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2628.938114] LDISKFS-fs (dm-0): recovery complete [ 2628.960443] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2631.199151] LustreError: 50563:0:(import.c:337:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 2631.217326] LustreError: 50563:0:(import.c:361:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9c868bba8e00 x1865013530929408/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1778618067 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2631.259476] LustreError: 50563:0:(import.c:371:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 2631.671019] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c8684b9df80 x1865013530932736/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2632.274366] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2632.283669] Lustre: Skipped 4 previous similar messages [ 2632.331963] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2632.341597] Lustre: Skipped 4 previous similar messages [ 2637.370853] Lustre: 50595:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2637.464923] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2637.673964] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1601) [ 2637.674937] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1601) [ 2645.886074] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2647.680780] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2657.386173] Lustre: DEBUG MARKER: == replay-dual test 20: recovery time is not increasing == 16:34:52 (1778618092) [ 2666.598076] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2669.181340] Lustre: Failing over lustre-MDT0000 [ 2669.424655] Lustre: server umount lustre-MDT0000 complete [ 2693.261684] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2693.269775] LDISKFS-fs (dm-0): recovery complete [ 2693.278075] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2704.987602] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 2705.392780] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2705.400498] Lustre: Skipped 12 previous similar messages [ 2842.500193] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2842.505468] Lustre: 52421:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client d49cf180-a5af-41a6-9248-a3acbe5ffdfa@ [ 2842.514880] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2842.546499] Lustre: 52421:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2842.553420] Lustre: 52421:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 6 previous similar messages [ 2842.669700] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1603 to 0x2c0000401:1633) [ 2842.676657] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1633) [ 2847.392727] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2849.503331] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2860.797560] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2863.543808] Lustre: Failing over lustre-MDT0000 [ 2863.962598] Lustre: server umount lustre-MDT0000 complete [ 2864.109367] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2864.130384] Lustre: Skipped 11 previous similar messages [ 2887.067180] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2887.069466] LDISKFS-fs (dm-0): recovery complete [ 2887.085477] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2893.010947] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 2893.020315] Lustre: Skipped 3 previous similar messages [ 2895.489846] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3034.500501] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3034.503655] Lustre: 54090:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 1ca9701e-f4d2-4b32-86bf-284206584f3b@ [ 3034.531137] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3034.598335] Lustre: 54090:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3034.613361] Lustre: 54090:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3034.711769] Lustre: lustre-MDT0000: Recovery over after 2:21, of 3 clients 2 recovered and 1 was evicted. [ 3034.722760] Lustre: Skipped 3 previous similar messages [ 3034.769443] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1665) [ 3034.770756] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1665) [ 3039.858730] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3042.236909] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3053.153825] Lustre: DEBUG MARKER: == replay-dual test 21a: commit on sharing =============== 16:41:28 (1778618488) [ 3064.267301] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3067.210922] Lustre: Failing over lustre-MDT0000 [ 3067.722421] Lustre: server umount lustre-MDT0000 complete [ 3086.817225] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778618506/real 1778618506] req@ffff9c87b56a3800 x1865013531131392/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778618522 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3086.850766] Lustre: 3652:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 3092.590542] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3092.593880] LDISKFS-fs (dm-0): recovery complete [ 3092.605933] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3102.086962] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3237.504525] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3237.516096] Lustre: 56008:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client c75dfc8b-5516-4301-b550-02e0f9ec8423@ [ 3237.544401] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3237.614399] Lustre: 56008:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3237.628029] Lustre: 56008:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3237.699995] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1697) [ 3237.700525] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1697) [ 3248.440724] Lustre: DEBUG MARKER: SKIP: replay-dual test_21b skipping SLOW test 21b [ 3251.129456] Lustre: DEBUG MARKER: == replay-dual test 22a: c1 lfs mkdir -i 1 dir1, M1 drop reply [ 3252.991307] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3253.006839] LustreError: 15123:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c8790a5a300 x1865013510738688/t4294967346(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:339/0 lens 560/448 e 0 to 0 dl 1778618769 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3256.206403] Lustre: Failing over lustre-MDT0001 [ 3256.715803] Lustre: server umount lustre-MDT0001 complete [ 3258.676644] LustreError: 6509:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3258.704734] LustreError: 6509:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 143 previous similar messages [ 3260.391531] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3277.106642] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3277.687727] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3277.697215] Lustre: Skipped 3 previous similar messages [ 3277.763526] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3277.784220] Lustre: Skipped 3 previous similar messages [ 3283.075833] Lustre: 6507:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c87b4669f80 x1865013510738688/t4294967346(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:370/0 lens 560/2880 e 0 to 0 dl 1778618800 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3283.133259] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3292.597812] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3294.857485] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3306.510624] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3309.455255] Lustre: Failing over lustre-MDT0000 [ 3309.949637] Lustre: server umount lustre-MDT0000 complete [ 3313.639517] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3330.007272] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3330.019243] LustreError: Skipped 3 previous similar messages [ 3334.802307] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3334.805670] LDISKFS-fs (dm-0): recovery complete [ 3334.826274] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3340.271007] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c8789424700 x1865013531254272/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3340.324760] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 1 previous similar message [ 3340.543347] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (not set up) [ 3340.556055] Lustre: Skipped 1 previous similar message [ 3345.455892] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3345.907082] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 3345.916292] Lustre: Skipped 14 previous similar messages [ 3345.965640] Lustre: 58951:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3346.107406] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1729) [ 3346.108477] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1729) [ 3354.842340] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3356.590485] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3366.868657] Lustre: DEBUG MARKER: == replay-dual test 22b: c1 lfs mkdir -i 1 d1, M1 drop reply [ 3368.319022] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3368.323560] LustreError: 44562:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c868bbfea00 x1865013510781696/t8589934617(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:455/0 lens 560/448 e 0 to 0 dl 1778618885 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3371.061134] Lustre: Failing over lustre-MDT0000 [ 3371.581628] Lustre: server umount lustre-MDT0000 complete [ 3375.529470] LustreError: 6493:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778618811 with bad export cookie 17211453990685247089 [ 3375.530492] Lustre: Failing over lustre-MDT0001 [ 3375.536040] LustreError: 6493:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 3375.575031] LustreError: 59981:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9c87b557ca80 x1865013531287808/t0(0) o1000->lustre-MDT0000-osp-MDT0001@0@lo:24/4 lens 304/4320 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'umount.0' uid:0 gid:0 projid:4294967295 [ 3375.608784] LustreError: 59981:0:(osp_object.c:617:osp_attr_get()) lustre-MDT0000-osp-MDT0001: osp_attr_get update error [0x200000401:0x1:0x0]: rc = -5 [ 3376.043966] Lustre: server umount lustre-MDT0001 complete [ 3395.711697] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3395.804863] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3396.057135] LustreError: 60675:0:(llog.c:1646:llog_backup()) MGC192.168.204.103@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3396.072138] Lustre: 60675:0:(mgc_request_server.c:768:mgc_llog_local_copy()) MGC192.168.204.103@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3400.171783] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec60db7 [ 3405.493536] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3405.697847] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3406.903569] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:36 to 0x2c0000400:65) [ 3406.903864] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:36 to 0x280000400:65) [ 3406.958413] Lustre: 60697:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c87ba416300 x1865013510781696/t8589934617(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:493/0 lens 560/2880 e 0 to 0 dl 1778618923 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3413.734985] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1761) [ 3413.736672] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1761) [ 3418.177235] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3420.249047] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3422.249292] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3432.513173] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3434.644455] Lustre: Failing over lustre-MDT0000 [ 3434.927471] Lustre: server umount lustre-MDT0000 complete [ 3457.526479] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3457.528439] LDISKFS-fs (dm-0): recovery complete [ 3457.549581] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3468.589737] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3469.333569] Lustre: 62777:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3469.346333] Lustre: 62777:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3469.450501] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1793) [ 3469.451450] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1793) [ 3477.203407] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3479.024785] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3489.158382] Lustre: DEBUG MARKER: == replay-dual test 22c: c1 lfs mkdir -i 1 d1, M1 drop update [ 3490.759318] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3490.769759] LustreError: 8399:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87bb68aa00 x1865013531361024/t107374182411(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:507/0 lens 2520/4320 e 0 to 0 dl 1778618937 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3494.745530] Lustre: Failing over lustre-MDT0000 [ 3494.887915] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3494.899980] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3494.914797] Lustre: Skipped 25 previous similar messages [ 3494.931537] Lustre: Skipped 3 previous similar messages [ 3495.315504] Lustre: server umount lustre-MDT0000 complete [ 3515.724285] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3516.167363] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (not set up) [ 3516.178328] Lustre: Skipped 1 previous similar message [ 3521.282882] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3521.301556] Lustre: Skipped 6 previous similar messages [ 3521.587279] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3521.677975] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1825) [ 3521.677975] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1825) [ 3530.233536] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3532.164895] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3543.841263] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3546.871195] Lustre: Failing over lustre-MDT0000 [ 3546.906121] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (stopping) [ 3546.922970] Lustre: Skipped 1 previous similar message [ 3547.103968] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3547.452696] Lustre: server umount lustre-MDT0000 complete [ 3571.734838] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3571.736891] LDISKFS-fs (dm-0): recovery complete [ 3571.744928] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3578.873471] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec6242a [ 3583.818586] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3584.582866] Lustre: 65748:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3584.603491] Lustre: 65748:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3584.771736] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1857) [ 3584.774370] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1857) [ 3592.523220] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3594.496756] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3604.772101] Lustre: DEBUG MARKER: == replay-dual test 22d: c1 lfs mkdir -i 1 d1, M1 drop update [ 3609.914699] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3609.925228] LustreError: 64475:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87ba416300 x1865013531435520/t115964117002(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:626/0 lens 2520/4320 e 0 to 0 dl 1778619056 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3614.511062] Lustre: Failing over lustre-MDT0000 [ 3615.430501] Lustre: server umount lustre-MDT0000 complete [ 3619.687119] Lustre: Failing over lustre-MDT0001 [ 3619.693691] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778619055 with bad export cookie 17211453990685254698 [ 3619.713759] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 3619.726228] LustreError: 66874:0:(ldlm_resource.c:1180:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x79:0x0].0xf7117594 (ffff9c8781323800) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3619.782114] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.3@tcp (stopping) [ 3619.790578] Lustre: Skipped 4 previous similar messages [ 3626.495839] Lustre: server umount lustre-MDT0001 complete [ 3647.198830] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3647.524029] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3647.641396] LustreError: 67570:0:(llog.c:1646:llog_backup()) MGC192.168.204.103@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3647.657936] Lustre: 67570:0:(mgc_request_server.c:768:mgc_llog_local_copy()) MGC192.168.204.103@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3664.362419] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec62bcb [ 3670.376333] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3670.395377] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3671.119233] Lustre: lustre-MDT0000: Recovery over after 0:01, of 3 clients 3 recovered and 0 were evicted. [ 3671.131305] Lustre: Skipped 8 previous similar messages [ 3671.183423] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1889) [ 3671.191632] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1889) [ 3671.263777] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 3671.275491] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 3671.310768] Lustre: 67595:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c87a6b0df80 x1865013510869376/t12884901939(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:2/0 lens 560/2880 e 0 to 0 dl 1778619187 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3679.133480] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3681.195542] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3682.855504] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3693.808033] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3697.329669] Lustre: Failing over lustre-MDT0000 [ 3697.835686] Lustre: server umount lustre-MDT0000 complete [ 3717.599097] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778619137/real 1778619137] req@ffff9c87c1c03800 x1865013531479936/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778619153 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3717.636339] Lustre: 3649:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 16 previous similar messages [ 3722.430072] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3722.432218] LDISKFS-fs (dm-0): recovery complete [ 3722.462718] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3727.868308] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec63445 [ 3733.517050] Lustre: 69682:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3733.533093] Lustre: 69682:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3733.673836] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3733.715394] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1921) [ 3733.720049] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1921) [ 3743.148190] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3745.041046] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3755.577853] Lustre: DEBUG MARKER: == replay-dual test 23a: c1 rmdir d1, M1 drop reply and fail, client2 mkdir d1 ========================================================== 16:53:11 (1778619191) [ 3757.338857] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3757.346347] LustreError: 69009:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c868b824a80 x1865013510915712/t17179869210(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:88/0 lens 496/456 e 0 to 0 dl 1778619273 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3761.369828] Lustre: Failing over lustre-MDT0001 [ 3761.785380] Lustre: server umount lustre-MDT0001 complete [ 3764.193461] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3781.862356] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3787.496393] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3787.853950] Lustre: 67595:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c8790a5e680 x1865013510915712/t17179869210(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:118/0 lens 496/2888 e 0 to 0 dl 1778619303 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3787.855902] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:129) [ 3787.876336] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:129) [ 3796.166738] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3797.928970] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3808.478876] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3810.747461] Lustre: Failing over lustre-MDT0000 [ 3811.141080] Lustre: server umount lustre-MDT0000 complete [ 3813.363857] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3834.234991] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3834.239996] LDISKFS-fs (dm-0): recovery complete [ 3834.263640] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3845.746590] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3846.162389] Lustre: 72625:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3846.184745] Lustre: 72625:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 4 previous similar messages [ 3846.367046] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1953) [ 3846.372564] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1953) [ 3854.363156] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3856.565891] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3867.326507] Lustre: DEBUG MARKER: == replay-dual test 23b: c1 rmdir d1, M1 drop reply and fail M0/M1, c2 mkdir d1 ========================================================== 16:55:02 (1778619302) [ 3869.291501] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3869.301644] LustreError: 67597:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b0fc7100 x1865013510951424/t21474836483(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:200/0 lens 496/456 e 0 to 0 dl 1778619385 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3873.348389] Lustre: Failing over lustre-MDT0000 [ 3873.374027] LustreError: 6493:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778619309 with bad export cookie 17211453990685262013 [ 3873.381964] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3873.387924] Lustre: Skipped 6 previous similar messages [ 3873.633269] Lustre: server umount lustre-MDT0000 complete [ 3876.834628] LustreError: 64475:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3876.857300] LustreError: 64475:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 314 previous similar messages [ 3877.622425] LustreError: 9461:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778619313 with bad export cookie 17211453990685261908 [ 3877.624524] Lustre: Failing over lustre-MDT0001 [ 3877.630399] LustreError: 9461:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 3878.094611] Lustre: server umount lustre-MDT0001 complete [ 3898.279084] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3898.352241] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3898.580980] LustreError: 74399:0:(llog.c:1646:llog_backup()) MGC192.168.204.103@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3898.593659] Lustre: 74399:0:(mgc_request_server.c:768:mgc_llog_local_copy()) MGC192.168.204.103@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3903.456707] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c8681ed0700 x1865013531595904/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3903.476927] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 18 previous similar messages [ 3903.834194] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3903.843621] Lustre: Skipped 11 previous similar messages [ 3903.905606] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3903.915745] Lustre: Skipped 11 previous similar messages [ 3908.507840] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3909.148776] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3910.266735] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:161) [ 3910.266890] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:161) [ 3910.311804] Lustre: 74417:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c87906bce00 x1865013510951424/t21474836483(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:241/0 lens 496/2888 e 0 to 0 dl 1778619426 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3917.468309] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1985) [ 3917.469501] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1985) [ 3922.412439] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3924.632816] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3926.607948] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3937.776658] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3940.492528] Lustre: Failing over lustre-MDT0000 [ 3940.920870] Lustre: server umount lustre-MDT0000 complete [ 3943.399674] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3959.846719] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3959.852603] LustreError: Skipped 8 previous similar messages [ 3965.393744] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3965.398589] LDISKFS-fs (dm-0): recovery complete [ 3965.407620] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3969.004996] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec64fc0 [ 3969.016214] Lustre: MGC192.168.204.103@tcp: Connection restored to 0@lo (at 0@lo) [ 3969.019755] Lustre: Skipped 48 previous similar messages [ 3974.356923] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 3974.692085] Lustre: 76502:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3974.709083] Lustre: 76502:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 8 previous similar messages [ 3974.957730] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2017) [ 3974.963024] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2017) [ 3983.107828] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3984.879657] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3994.552940] Lustre: DEBUG MARKER: == replay-dual test 23c: c1 rmdir d1, M0 drop update reply and fail M0, c2 mkdir d1 ========================================================== 16:57:10 (1778619430) [ 3996.479960] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3996.491741] LustreError: 8399:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87906bf800 x1865013531667200/t137438953491(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:258/0 lens 1984/4320 e 0 to 0 dl 1778619443 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 4000.523080] Lustre: Failing over lustre-MDT0000 [ 4000.559545] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (stopping) [ 4000.569679] Lustre: Skipped 1 previous similar message [ 4000.738805] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 4002.858179] Lustre: server umount lustre-MDT0000 complete [ 4021.884495] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4037.715665] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4037.738866] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2049) [ 4037.740279] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2049) [ 4046.050372] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4048.037716] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4058.095483] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4060.301702] Lustre: Failing over lustre-MDT0000 [ 4060.614948] Lustre: server umount lustre-MDT0000 complete [ 4063.215946] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 4083.293158] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4083.296431] LDISKFS-fs (dm-0): recovery complete [ 4083.307822] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4094.733970] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4095.811886] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2081) [ 4095.812562] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2081) [ 4104.243717] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4106.297503] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4116.951635] Lustre: DEBUG MARKER: == replay-dual test 23d: c1 rmdir d1, M0 drop update reply and fail M0/M1, c2 mkdir d1 ========================================================== 16:59:12 (1778619552) [ 4122.782641] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 4122.795749] LustreError: 64475:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b689f450 x1865013531744128/t146028888081(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:384/0 lens 1984/4320 e 0 to 0 dl 1778619569 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 4127.483661] Lustre: Failing over lustre-MDT0000 [ 4128.074088] Lustre: server umount lustre-MDT0000 complete [ 4131.308360] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4131.341795] Lustre: Skipped 41 previous similar messages [ 4132.892748] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778619568 with bad export cookie 17211453990685269174 [ 4132.921669] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 4132.925823] Lustre: Failing over lustre-MDT0001 [ 4132.945134] LustreError: 80586:0:(ldlm_resource.c:1180:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x81:0x0].0x0 (ffff9c8798e13800) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 4132.999629] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.3@tcp (stopping) [ 4133.006264] Lustre: Skipped 1 previous similar message [ 4139.453333] Lustre: server umount lustre-MDT0001 complete [ 4160.244793] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4160.253431] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4177.381818] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec66434 [ 4178.167224] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_connect to node 0@lo failed: rc = -114 [ 4182.157609] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4182.389879] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4183.531816] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 4183.553580] Lustre: Skipped 11 previous similar messages [ 4186.975636] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2113) [ 4186.977808] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2113) [ 4201.428135] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:193) [ 4201.438427] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:193) [ 4201.440834] Lustre: 82048:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c8683217480 x1865013511032832/t25769803783(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:532/0 lens 496/2888 e 0 to 0 dl 1778619717 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 4206.041549] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4208.041600] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4209.883697] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4220.352528] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4222.762881] Lustre: Failing over lustre-MDT0000 [ 4222.952407] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4222.962727] Lustre: Skipped 6 previous similar messages [ 4223.037690] Lustre: server umount lustre-MDT0000 complete [ 4249.800568] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4249.803080] LDISKFS-fs (dm-0): recovery complete [ 4249.815654] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4251.650105] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec66c30 [ 4257.049924] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4257.354682] Lustre: 83416:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4257.381907] Lustre: 83416:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 17 previous similar messages [ 4257.652883] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2115 to 0x280000401:2145) [ 4257.657700] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2115 to 0x2c0000401:2145) [ 4266.160734] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4268.455497] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4281.608184] Lustre: DEBUG MARKER: == replay-dual test 24: reconstruct on non-existing object ========================================================== 17:01:56 (1778619716) [ 4283.303163] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 4283.310350] LustreError: 82048:0:(ldlm_lib.c:3328:target_send_reply_msg()) @@@ dropping reply req@ffff9c87b043e680 x1865013511074816/t154618822673(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:614/0 lens 488/456 e 0 to 0 dl 1778619799 ref 1 fl Interpret:/200/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4368.691026] Lustre: lustre-MDT0000: Client a9b271a4-70b1-42e0-be39-2d76a5da0732 (at 192.168.204.3@tcp) reconnecting [ 4368.710047] Lustre: 82050:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9c87b0499500 x1865013511074816/t154618822673(0) o36->a9b271a4-70b1-42e0-be39-2d76a5da0732@192.168.204.3@tcp:699/0 lens 488/3152 e 0 to 0 dl 1778619884 ref 1 fl Interpret:/202/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4376.406248] Lustre: DEBUG MARKER: == replay-dual test 25: replay|resend ==================== 17:03:31 (1778619811) [ 4378.422782] Lustre: *** cfs_fail_loc=304, val=0*** [ 4381.409659] Lustre: Failing over lustre-OST0000 [ 4381.747611] Lustre: server umount lustre-OST0000 complete [ 4404.181076] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4406.563633] Lustre: lustre-OST0000: Recovery over after 0:01, of 4 clients 4 recovered and 0 were evicted. [ 4406.586093] Lustre: Skipped 12 previous similar messages [ 4411.198236] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4420.543453] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4422.727584] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4434.552158] Lustre: DEBUG MARKER: == replay-dual test 26: dbench and tar with mds failover ========================================================== 17:04:30 (1778619870) [ 4446.574352] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4450.435310] Lustre: DEBUG MARKER: test_26 fail mds1 1 times [ 4452.569319] Lustre: Failing over lustre-MDT0000 [ 4452.709987] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.3@tcp (stopping) [ 4453.029699] Lustre: server umount lustre-MDT0000 complete [ 4471.263462] Lustre: 3650:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778619891/real 1778619891] req@ffff9c87b043d180 x1865013531945856/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778619907 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4471.302884] Lustre: 3650:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 16 previous similar messages [ 4475.699323] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 4475.702631] LDISKFS-fs (dm-0): recovery complete [ 4475.716191] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4481.292900] LustreError: 81311:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4481.315504] LustreError: 81311:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 235 previous similar messages [ 4487.191603] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4489.340059] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2170 to 0x2c0000401:2209) [ 4489.348289] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2171 to 0x280000401:2209) [ 4496.028484] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4498.395649] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4513.805469] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4517.971265] Lustre: DEBUG MARKER: test_26 fail mds2 2 times [ 4520.630784] Lustre: Failing over lustre-MDT0001 [ 4520.658910] LustreError: 10261:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 192.168.204.3@tcp arrived at 1778619956 with bad export cookie 17211453990685271470 [ 4520.752577] LustreError: 81313:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9c87b46ad180 x1865013532082048/t0(0) o101->lustre-MDT0000-osp-MDT0001@0@lo:24/4 lens 328/344 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'mdt00_002.0' uid:0 gid:0 projid:4294967295 [ 4521.075693] Lustre: server umount lustre-MDT0001 complete [ 4522.990026] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 4523.008312] LustreError: Skipped 1 previous similar message [ 4545.355239] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4545.358104] LDISKFS-fs (dm-1): recovery complete [ 4545.367728] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4545.907101] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 4545.913582] Lustre: Skipped 9 previous similar messages [ 4545.961159] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 4545.970572] Lustre: Skipped 9 previous similar messages [ 4551.209800] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4553.984986] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:267 to 0x280000400:289) [ 4553.988584] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:267 to 0x2c0000400:289) [ 4560.712652] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4562.823918] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4577.273822] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4581.442343] Lustre: DEBUG MARKER: test_26 fail mds1 3 times [ 4584.087728] Lustre: Failing over lustre-MDT0000 [ 4584.493894] Lustre: server umount lustre-MDT0000 complete [ 4603.360089] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4603.381504] LustreError: Skipped 5 previous similar messages [ 4608.576791] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4608.581902] LDISKFS-fs (dm-0): recovery complete [ 4608.588289] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4613.604974] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9c868f777480 x1865013532284672/t0(0) o250->MGC192.168.204.103@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4613.633376] LustreError: 3648:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 12 previous similar messages [ 4618.294504] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4619.261213] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 4619.275733] Lustre: Skipped 33 previous similar messages [ 4620.918852] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2288 to 0x2c0000401:2305) [ 4620.920919] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2288 to 0x280000401:2305) [ 4627.525225] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4630.481745] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4644.001915] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4648.348902] Lustre: DEBUG MARKER: test_26 fail mds2 4 times [ 4650.949091] Lustre: Failing over lustre-MDT0001 [ 4657.517367] Lustre: server umount lustre-MDT0001 complete [ 4681.127759] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4681.131849] LDISKFS-fs (dm-1): recovery complete [ 4681.150794] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4686.102992] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4688.322059] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:338 to 0x280000400:353) [ 4688.325898] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:338 to 0x2c0000400:353) [ 4695.264154] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4697.445940] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4771.170844] Lustre: DEBUG MARKER: == replay-dual test 28: lock replay should be ordered: waiting after granted ========================================================== 17:10:06 (1778620206) [ 4789.767502] Lustre: Failing over lustre-OST0000 [ 4789.939207] Lustre: server umount lustre-OST0000 complete [ 4790.240210] LustreError: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 4790.253409] LustreError: Skipped 1 previous similar message [ 4790.268278] Lustre: lustre-OST0000-osc-MDT0000: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4790.281521] Lustre: Skipped 22 previous similar messages [ 4809.357039] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4810.496260] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 4810.503303] Lustre: Skipped 6 previous similar messages [ 4811.476967] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4816.539610] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4825.570795] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4827.546282] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4839.393966] Lustre: DEBUG MARKER: == replay-dual test 29: replay vs update with the same xid ========================================================== 17:11:14 (1778620274) [ 4840.893566] Lustre: DEBUG MARKER: SKIP: replay-dual test_29 needs >= 2 clients [ 4843.054140] Lustre: DEBUG MARKER: == replay-dual test 30: layout lock replay is not blocked on IO ========================================================== 17:11:18 (1778620278) [ 4846.567970] Lustre: Failing over lustre-MDT0000 [ 4847.755653] LustreError: 9461:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_convert from 192.168.204.3@tcp arrived at 1778620283 with bad export cookie 17211453990685360930 [ 4849.018670] Lustre: server umount lustre-MDT0000 complete [ 4869.554113] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4876.284530] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xeedb5e36dec996cb [ 4881.952664] Lustre: 95340:0:(ldlm_lib.c:2070:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4881.969136] Lustre: 95340:0:(ldlm_lib.c:2070:extend_recovery_timer()) Skipped 634 previous similar messages [ 4882.059221] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2382 to 0x280000401:2401) [ 4882.060151] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2381 to 0x2c0000401:2401) [ 4882.183133] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4890.193191] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4892.008954] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4901.420516] Lustre: DEBUG MARKER: == replay-dual test 31: deadlock on file_remove_privs and occupied mod rpc slots ========================================================== 17:12:16 (1778620336) [ 4905.291436] Lustre: Failing over lustre-OST0000 [ 4905.429991] Lustre: server umount lustre-OST0000 complete [ 4925.277193] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4933.042691] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4942.590480] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4944.792090] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in IDLE [ 4957.613202] Lustre: DEBUG MARKER: == replay-dual test 32: gap in update llog shouldn't break recovery ========================================================== 17:13:12 (1778620392) [ 4958.932921] Lustre: *** cfs_fail_loc=131d, val=10*** [ 4959.454084] Lustre: *** cfs_fail_loc=131d, val=4*** [ 4959.457492] Lustre: Skipped 5 previous similar messages [ 4960.478455] Lustre: *** cfs_fail_loc=131d, val=4294967286*** [ 4960.482484] Lustre: Skipped 13 previous similar messages [ 4964.120694] Lustre: Failing over lustre-MDT0001 [ 4964.704817] Lustre: server umount lustre-MDT0001 complete [ 4970.200849] Lustre: Failing over lustre-MDT0000 [ 4970.804624] Lustre: server umount lustre-MDT0000 complete [ 4979.260581] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4979.830060] Lustre: *** cfs_fail_loc=131d, val=4294967266*** [ 4979.832449] Lustre: Skipped 19 previous similar messages [ 4985.386845] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 4994.100214] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4994.281646] Lustre: *** cfs_fail_loc=131d, val=4294967262*** [ 4994.284814] Lustre: Skipped 3 previous similar messages [ 4999.259068] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 5000.335245] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:381 to 0x2c0000400:417) [ 5000.337204] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:381 to 0x280000400:417) [ 5000.354199] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2381 to 0x2c0000401:2433) [ 5000.355254] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2497) [ 5012.827806] Lustre: DEBUG MARKER: == replay-dual test 33: Check for OBD_INCOMPAT_MULTI_RPCS in last_rcvd after abort_recovery ========================================================== 17:14:08 (1778620448) [ 5020.566729] Lustre: Failing over lustre-MDT0001 [ 5020.658673] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 5020.665399] Lustre: Skipped 14 previous similar messages [ 5020.907696] Lustre: server umount lustre-MDT0001 complete [ 5043.808116] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5050.388824] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 5058.185862] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount REPLAY_WAIT mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5060.587990] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in REPLAY_WAIT state after 0 sec [ 5061.434150] Lustre: lustre-MDT0001: Aborting client recovery [ 5061.437389] LustreError: 100794:0:(ldlm_lib.c:2986:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 5061.442118] Lustre: 100273:0:(ldlm_lib.c:2389:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 5061.447331] Lustre: 100273:0:(ldlm_lib.c:2389:target_recovery_overseer()) Skipped 2 previous similar messages [ 5061.452742] Lustre: 100273:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 52619953-a0b7-42ef-9e66-f313570bb9aa@ [ 5061.459835] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 5061.484930] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 5061.494890] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 5061.532157] Lustre: lustre-MDT0001: Recovery over after 0:15, of 3 clients 2 recovered and 1 was evicted. [ 5061.544301] Lustre: Skipped 9 previous similar messages [ 5061.581181] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:381 to 0x2c0000400:449) [ 5061.587376] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:381 to 0x280000400:449) [ 5065.716661] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5067.770625] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5070.937790] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 5077.966137] Lustre: Failing over lustre-MDT0001 [ 5078.393711] Lustre: server umount lustre-MDT0001 complete [ 5082.909278] LustreError: 98363:0:(ldlm_lib.c:1180:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.204.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 5082.944764] LustreError: 98363:0:(ldlm_lib.c:1180:target_handle_connect()) Skipped 193 previous similar messages [ 5086.936139] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5091.755181] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug -1 all [ 5092.914703] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:381 to 0x2c0000400:481) [ 5092.914995] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:381 to 0x280000400:481) [ 5098.562820] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5101.051290] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5104.817365] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 5116.160771] Lustre: DEBUG MARKER: == replay-dual test complete, duration 4819 sec ========== 17:15:51 (1778620551) [ 5118.709929] Lustre: DEBUG MARKER: === replay-dual: start cleanup 17:15:53 (1778620553) === [ 5133.960813] Lustre: DEBUG MARKER: === replay-dual: finish cleanup 17:16:09 (1778620569) === [ 5136.111967] Lustre: Failing over lustre-MDT0000 [ 5136.451672] Lustre: server umount lustre-MDT0000 complete [ 5155.297866] Lustre: 3650:0:(client.c:2479:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1778620574/real 1778620574] req@ffff9c87b447aa00 x1865013533362432/t0(0) o400->MGC192.168.204.103@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1778620590 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 5155.332108] Lustre: 3650:0:(client.c:2479:ptlrpc_expire_one_request()) Skipped 6 previous similar messages [ 5164.947960] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5166.045396] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 5166.050399] Lustre: Skipped 9 previous similar messages [ 5166.122363] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 5166.141394] Lustre: Skipped 9 previous similar messages [ 5171.298085] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 5306.501423] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 5306.514658] Lustre: 103719:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 52619953-a0b7-42ef-9e66-f313570bb9aa@ [ 5306.522275] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 5306.540251] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 5306.548850] Lustre: Skipped 30 previous similar messages [ 5306.576172] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2381 to 0x2c0000401:2465) [ 5306.579344] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2529) [ 5311.743686] Lustre: DEBUG MARKER: oleg403-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 5313.537629] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5325.687249] Lustre: server umount lustre-MDT0000 complete [ 5333.648702] LustreError: 6493:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1778620769 with bad export cookie 17211453990685496373 [ 5333.649301] LustreError: MGC192.168.204.103@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 5333.661718] LustreError: 6493:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 5333.676688] LustreError: Skipped 3 previous similar messages [ 5334.063794] Lustre: server umount lustre-MDT0001 complete [ 5352.265927] Lustre: server umount lustre-OST0000 complete [ 5372.995361] Lustre: server umount lustre-OST0001 complete [ 5391.228569] Lustre: DEBUG MARKER: oleg403-server.virtnet: executing unload_modules_local [ 5394.516398] Key type lgssc unregistered [ 5394.928036] LNet: 106527:0:(lib-ptl.c:967:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 5394.955684] LNetError: 106527:0:(acceptor.c:252:lnet_acceptor_remove_socket()) Interface ens2 not found [ 5394.992797] LNet: Removed LNI 192.168.204.103@tcp [ 5396.193867] Key type .llcrypt unregistered [ 5396.199623] Key type ._llcrypt unregistered