[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-10.fc44 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 670519807 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2399.998 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE247C 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE2318 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 0022D8 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE238C 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE241C 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2454 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe2318-0xbffe238b] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe2317] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe238c-0xbffe241b] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe241c-0xbffe2453] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2454-0xbffe247b] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524584K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001013] APIC: Switch to symmetric I/O mode setup [ 0.002000] x2apic enabled [ 0.002000] Switched APIC routing to physical x2apic. [ 0.002000] kvm-guest: setup PV IPIs [ 0.002000] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.002000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x229835b7123, max_idle_ns: 440795242976 ns [ 0.002000] Calibrating delay loop (skipped) preset value.. 4799.99 BogoMIPS (lpj=2399998) [ 0.002000] pid_max: default: 32768 minimum: 301 [ 0.002000] LSM: Security Framework initializing [ 0.002000] Yama: becoming mindful. [ 0.002032] SELinux: Initializing. [ 0.003074] *** VALIDATE selinux *** [ 0.011821] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.017361] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.018160] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.020080] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.021111] *** VALIDATE tmpfs *** [ 0.023356] *** VALIDATE proc *** [ 0.024227] *** VALIDATE cgroup *** [ 0.025009] *** VALIDATE cgroup2 *** [ 0.027019] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.028178] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.029009] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.030026] Spectre V2 : User space: Vulnerable [ 0.031008] Speculative Store Bypass: Vulnerable [ 0.034309] debug: unmapping init [mem 0xffffffffbaa59000-0xffffffffbaa60fff] [ 0.037158] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.038686] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.039025] ... version: 2 [ 0.040009] ... bit width: 48 [ 0.041012] ... generic registers: 4 [ 0.042012] ... value mask: 0000ffffffffffff [ 0.043014] ... max period: 00007fffffffffff [ 0.044014] ... fixed-purpose events: 3 [ 0.045013] ... event mask: 000000070000000f [ 0.046274] rcu: Hierarchical SRCU implementation. [ 0.048392] smp: Bringing up secondary CPUs ... [ 0.049548] x86: Booting SMP configuration: [ 0.050025] .... node #0, CPUs: #1 #2 #3 [ 0.056220] smp: Brought up 1 node, 4 CPUs [ 0.058013] smpboot: Max logical packages: 1 [ 0.059015] smpboot: Total of 4 processors activated (19199.98 BogoMIPS) [ 0.105323] node 0 deferred pages initialised in 44ms [ 0.109284] devtmpfs: initialized [ 0.110200] x86/mm: Memory block size: 128MB [ 0.113237] gcov: version magic: 0x41383552 [ 0.115277] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.130147] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.131251] pinctrl core: initialized pinctrl subsystem [ 0.132201] [ 0.132790] ************************************************************* [ 0.135018] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.138012] ** ** [ 0.140010] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.142010] ** ** [ 0.144011] ** This means that this kernel is built to expose internal ** [ 0.147010] ** IOMMU data structures, which may compromise security on ** [ 0.149009] ** your system. ** [ 0.152011] ** ** [ 0.154007] ** If you see this message and you are not debugging the ** [ 0.156009] ** kernel, report this immediately to your vendor! ** [ 0.157006] ** ** [ 0.159010] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.161009] ************************************************************* [ 0.165078] NET: Registered protocol family 16 [ 0.167048] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.169059] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.172054] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.175097] cpuidle: using governor menu [ 0.176470] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.179453] PCI: Using configuration type 1 for base access [ 0.181151] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.191060] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.193047] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.196397] cryptd: max_cpu_qlen set to 1000 [ 0.199095] ACPI: Added _OSI(Module Device) [ 0.201012] ACPI: Added _OSI(Processor Device) [ 0.202014] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.204013] ACPI: Added _OSI(Processor Aggregator Device) [ 0.208783] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.215401] ACPI: Interpreter enabled [ 0.217103] ACPI: PM: (supports S0 S3 S4 S5) [ 0.219013] ACPI: Using IOAPIC for interrupt routing [ 0.222261] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.229760] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.242316] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.245052] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.248034] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.252083] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.259579] acpiphp: Slot [2] registered [ 0.261129] acpiphp: Slot [5] registered [ 0.262095] acpiphp: Slot [6] registered [ 0.264090] acpiphp: Slot [7] registered [ 0.265115] acpiphp: Slot [8] registered [ 0.267128] acpiphp: Slot [9] registered [ 0.269212] acpiphp: Slot [10] registered [ 0.271198] acpiphp: Slot [3] registered [ 0.273146] acpiphp: Slot [4] registered [ 0.275147] acpiphp: Slot [11] registered [ 0.277122] acpiphp: Slot [12] registered [ 0.279097] acpiphp: Slot [13] registered [ 0.280128] acpiphp: Slot [14] registered [ 0.282185] acpiphp: Slot [15] registered [ 0.284187] acpiphp: Slot [16] registered [ 0.286149] acpiphp: Slot [17] registered [ 0.287086] acpiphp: Slot [18] registered [ 0.289191] acpiphp: Slot [19] registered [ 0.291162] acpiphp: Slot [20] registered [ 0.292094] acpiphp: Slot [21] registered [ 0.293287] acpiphp: Slot [22] registered [ 0.294076] acpiphp: Slot [23] registered [ 0.295071] acpiphp: Slot [24] registered [ 0.297110] acpiphp: Slot [25] registered [ 0.298264] acpiphp: Slot [26] registered [ 0.299073] acpiphp: Slot [27] registered [ 0.300080] acpiphp: Slot [28] registered [ 0.301075] acpiphp: Slot [29] registered [ 0.302095] acpiphp: Slot [30] registered [ 0.303101] acpiphp: Slot [31] registered [ 0.305054] PCI host bridge to bus 0000:00 [ 0.306028] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.309027] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.312027] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.315026] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.318036] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.322014] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.323172] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.325964] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.329227] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.338652] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.344051] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.346015] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.348014] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.350013] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.351618] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.354655] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.357035] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.359715] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.366020] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.380015] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.388023] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.396211] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.412016] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.429016] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.472016] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.496176] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.506015] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.520015] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.553013] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.568030] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.578017] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.602017] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.643018] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.659829] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.671016] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.683017] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.721024] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.739027] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.749017] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.766023] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.814019] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.832987] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.843016] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.854015] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.880015] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.898938] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.901391] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.904451] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.907443] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.909280] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.914217] iommu: Default domain type: Passthrough [ 0.916461] SCSI subsystem initialized [ 0.918208] ACPI: bus type USB registered [ 0.920180] usbcore: registered new interface driver usbfs [ 0.922087] usbcore: registered new interface driver hub [ 0.924107] usbcore: registered new device driver usb [ 0.927181] pps_core: LinuxPPS API ver. 1 registered [ 0.929010] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.932068] PTP clock support registered [ 0.934125] EDAC MC: Ver: 3.0.0 [ 0.936105] PCI: Using ACPI for IRQ routing [ 0.937640] NetLabel: Initializing [ 0.939015] NetLabel: domain hash size = 128 [ 0.941008] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.943076] NetLabel: unlabeled traffic allowed by default [ 0.945118] vgaarb: loaded [ 0.947255] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.949013] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.957297] clocksource: Switched to clocksource kvm-clock [ 1.081697] VFS: Disk quotas dquot_6.6.0 [ 1.083628] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 1.086575] *** VALIDATE ramfs *** [ 1.088046] *** VALIDATE hugetlbfs *** [ 1.089859] pnp: PnP ACPI init [ 1.092610] pnp: PnP ACPI: found 6 devices [ 1.121432] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 1.125143] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 1.127806] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 1.130338] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 1.133053] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 1.135884] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 1.139142] NET: Registered protocol family 2 [ 1.141952] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 1.147220] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 1.151486] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 1.159212] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 1.162892] TCP: Hash tables configured (established 65536 bind 65536) [ 1.166223] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 1.170299] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.173730] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.177146] NET: Registered protocol family 1 [ 1.180941] RPC: Registered named UNIX socket transport module. [ 1.183498] RPC: Registered udp transport module. [ 1.185301] RPC: Registered tcp transport module. [ 1.187207] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 1.189938] NET: Registered protocol family 44 [ 1.191845] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 1.194115] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 1.196487] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 1.199051] PCI: CLS 0 bytes, default 64 [ 1.200790] Unpacking initramfs... [ 2.780743] debug: unmapping init [mem 0xffff9ce3fcc54000-0xffff9ce3fffbffff] [ 2.786028] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 2.789183] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 2.792762] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x229835b7123, max_idle_ns: 440795242976 ns [ 3.333573] Initialise system trusted keyrings [ 3.336687] Key type blacklist registered [ 3.339140] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 3.348882] zbud: loaded [ 3.352048] *** VALIDATE nfs *** [ 3.353208] *** VALIDATE nfs4 *** [ 3.355110] pstore: using deflate compression [ 3.359202] Platform Keyring initialized [ 3.466971] NET: Registered protocol family 38 [ 3.469393] Key type asymmetric registered [ 3.471160] Asymmetric key parser 'x509' registered [ 3.473324] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 3.476632] io scheduler mq-deadline registered [ 3.477866] io scheduler kyber registered [ 3.479751] io scheduler bfq registered [ 3.481399] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 3.484291] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 3.486729] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 3.490504] ACPI: Power Button [PWRF] [ 3.496623] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 3.502266] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 3.529193] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 3.544263] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 3.580435] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 3.612501] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 3.641756] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 3.648870] Non-volatile memory driver v1.3 [ 3.652933] Linux agpgart interface v0.103 [ 3.697095] virtio_blk virtio1: [vda] 150080 512-byte logical blocks (76.8 MB/73.3 MiB) [ 3.702555] vda: detected capacity change from 0 to 76840960 [ 3.733518] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 3.736795] vdb: detected capacity change from 0 to 1073741824 [ 3.761658] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.764969] vdc: detected capacity change from 0 to 2621440000 [ 3.786713] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.789925] vdd: detected capacity change from 0 to 2621440000 [ 3.810961] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.814147] vde: detected capacity change from 0 to 4294967296 [ 3.835203] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.839112] vdf: detected capacity change from 0 to 4294967296 [ 3.851119] libphy: Fixed MDIO Bus: probed [ 3.856784] usbcore: registered new interface driver usbserial_generic [ 3.859996] usbserial: USB Serial support registered for generic [ 3.863176] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 3.869103] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 3.870965] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 3.873962] mousedev: PS/2 mouse device common for all mice [ 3.877396] rtc_cmos 00:05: RTC can wake from S4 [ 3.880629] rtc_cmos 00:05: registered as rtc0 [ 3.882422] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 3.887943] intel_pstate: CPU model not supported [ 3.890584] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 3.897595] hid: raw HID events driver (C) Jiri Kosina [ 3.900141] usbcore: registered new interface driver usbhid [ 3.902459] usbhid: USB HID core driver [ 3.904699] drop_monitor: Initializing network drop monitor service [ 3.907702] Initializing XFRM netlink socket [ 3.910714] NET: Registered protocol family 10 [ 3.912925] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 3.921596] Segment Routing with IPv6 [ 3.923109] NET: Registered protocol family 17 [ 3.925428] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 3.925941] mpls_gso: MPLS GSO support [ 3.935634] RAS: Correctable Errors collector initialized. [ 3.938224] AVX version of gcm_enc/dec engaged. [ 3.939955] AES CTR mode by8 optimization enabled [ 4.069696] sched_clock: Marking stable (4069674005, 0)->(5084986405, -1015312400) [ 4.072889] registered taskstats version 1 [ 4.074756] Loading compiled-in X.509 certificates [ 4.076717] zswap: loaded using pool lzo/zbud [ 4.104394] Key type big_key registered [ 4.118406] Key type encrypted registered [ 4.120708] ima: No TPM chip found, activating TPM-bypass! [ 4.123530] ima: Allocated hash algorithm: sha1 [ 4.125665] ima: No architecture policies found [ 4.128084] evm: Initialising EVM extended attributes: [ 4.129609] evm: security.selinux [ 4.131118] evm: security.ima [ 4.131975] evm: security.capability [ 4.133305] evm: HMAC attrs: 0x1 [ 4.136549] rtc_cmos 00:05: setting system clock to 2026-09-07 03:43:24 UTC (1788752604) [ 4.144925] debug: unmapping init [mem 0xffffffffbba03000-0xffffffffbbbfffff] [ 4.148770] debug: unmapping init [mem 0xffffffffba782000-0xffffffffbaa58fff] [ 4.153168] Write protecting the kernel read-only data: 28672k [ 4.156961] debug: unmapping init [mem 0xffffffffb8e03000-0xffffffffb8ffffff] [ 4.160336] debug: unmapping init [mem 0xffffffffb9714000-0xffffffffb97fffff] [ 4.199195] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 4.209865] systemd[1]: Detected virtualization kvm. [ 4.212411] systemd[1]: Detected architecture x86-64. [ 4.214408] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 4.244212] systemd[1]: No hostname configured. [ 4.247157] systemd[1]: Set hostname to . [ 4.250938] random: systemd: uninitialized urandom read (16 bytes read) [ 4.254228] systemd[1]: Initializing machine ID from random generator. [ 4.317566] random: ln: uninitialized urandom read (6 bytes read) [ 4.427981] random: systemd: uninitialized urandom read (16 bytes read) [ 4.431181] systemd[1]: Reached target Swap. [ OK ] Reached target Swap. [ 4.436365] systemd[1]: Reached target Timers. [ OK ] Reached target Timers. [ 4.443941] systemd[1]: Listening on udev Control Socket. [ OK ] Listening on udev Control Socket. [ OK ] Reached target Slices. [ OK ] Reached target Initrd Root Device. [ OK ] Reached target Local File Systems. [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Reached target Paths. [ OK ] Listening on Journal Socket. Starting Create Volatile Files and Directories... Starting Create list of required st…ce nodes for the current kernel... Starting Setup Virtual Console... [ OK ] Started Memstrack Anylazing Service. [ OK ] Listening on udev Kernel Socket. Starting Apply Kernel Variables... [ OK ] Listening on Journal Socket (/dev/log). [ OK ] Reached target Sockets. Starting Journal Service... [ OK ] Started Create Volatile Files and Directories. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Setup Virtual Console. [ OK ] Started Apply Kernel Variables. Starting dracut cmdline hook... Starting Create Static Device Nodes in /dev... [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started Journal Service. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 5.189962] device-mapper: uevent: version 1.0.3 [ 5.192431] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. [ 6.026851] random: fast init done [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. Starting dracut initqueue hook... [ 6.140139] virtio_net virtio0 ens2: renamed from eth0 [ 6.214698] scsi host0: ata_piix [ 6.230654] scsi host1: ata_piix [ 6.233032] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 6.235831] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 10.875426] random: crng init done [ 10.876909] random: 7 urandom warning(s) missed due to ratelimiting [ 11.036155] dracut-initqueue[603]: RTNETLINK answers: File exists Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Mounting /sysroot... [ 12.402624] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped target Timers. [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Initrd Root Device. [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Paths. [ OK ] Stopped target Slices. [ OK ] Stopped target Sockets. [ OK ] Stopped target System Initialization. [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped target Local File Systems. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. Stopping udev Kernel Device Manager... [ OK ] Stopped target Swap. [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Closed udev Control Socket. [ OK ] Closed udev Kernel Socket. Starting Cleanup udevd DB... [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 13.917725] printk: systemd: 26 output lines suppressed due to ratelimiting [ 14.235310] SELinux: Disabled at runtime. [ 14.300941] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 14.311089] systemd[1]: Detected virtualization kvm. [ 14.313721] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 15.127395] systemd[1]: initrd-switch-root.service: Succeeded. [ 15.134995] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 15.139852] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 15.143224] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 15.146084] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 15.153194] systemd[1]: Starting Journal Service... Starting Journal Service... [ 15.161753] systemd[1]: Activating swap /dev/disk/by-label/SWAP... Activating swap /dev/disk/by-label/SWAP... [ OK ] Created slice system-sshd\x2dkeygen.slice. [ OK ] Listening on udev Control Socket. [ OK ] Listening on udev Kernel Socket. Starting udev Coldplug all Devic[ 15.192291] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS es... [ OK ] Stopped target Switch Root. Mounting Huge Pages File System... Starting Remount Root and Kernel File Systems... [ OK ] Stopped target Initrd File Systems. Mounting Kernel Debug File System... [ OK ] Created slice system-serial\x2dgetty.slice. [ OK ] Listening on initctl Compatibility Named Pipe. [ OK ] Listening on RPCbind Server Activation Socket. [ OK ] Reached target RPC Port Mapper. [ OK ] Reached target rpc_pipefs.target. [ OK ] Started Forward Password Requests to Wall Directory Watch. [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Local Encrypted Volumes. Mounting POSIX Message Queue File System... [ OK ] Created slice User and Session Slice. [ OK ] Reached target Slices. [ OK ] Created slice system-getty.slice. Starting Create list of required st…ce nodes for the current kernel... [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. Starting Apply Kernel Variables... [ OK ] Stopped target Initrd Root File System. [ OK ] Listening on Process Core Dump Socket. [ OK ] Started Journal Service. [ OK ] Activated swap /dev/disk/by-label/SWAP. [ OK ] Mounted Huge Pages File System. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. [ OK ] Mounted Kernel Debug File System. [ OK ] Mounted POSIX Message Queue File System. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... [ OK ] Reached target Swap. Starting Flush Journal to Persistent Storage... [ OK ] Started udev Coldplug all Devices. [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started Create Static Device Nodes in /dev. Starting udev Kernel Device Manager... [ OK ] Reached target Local File Systems (Pre). Mounting /mnt... Mounting /home/green/git/lustre-release... [ OK ] Mounted /mnt. [ 16.189387] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Kernel Device Manager. [ 17.094726] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 17.378131] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 17.490259] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 17.502140] EDAC sbridge: Ver: 1.1.2 [ 19.589686] Key type dns_resolver registered [ 19.956735] NFS: Registering the id_resolver key type [ 19.958990] Key type id_resolver registered [ 19.961043] Key type id_legacy registered [ OK ] Started Configure read-only root support. [ OK ] Reached target Local File Systems. Starting Mark the need to relabel after reboot... Starting Rebuild Dynamic Linker Cache... Starting Create Volatile Files and Directories... Starting Load/Save Random Seed... [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Load/Save Random Seed. [ OK ] Started Create Volatile Files and Directories. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. [ OK ] Started irqbalance daemon. Starting Login Service... Starting Restore /run/initramfs on shutdown... [ OK ] Started dnf makecache --timer. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ OK ] Started D-Bus System Message Bus. Starting Network Manager... [ OK ] Started daily update of the root trust anchor for DNSSEC. [ OK ] Reached target Timers. [ OK ] Reached target sshd-keygen.target. [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Login Service. [ OK ] Started Network Manager. Starting Network Manager Wait Online... [ OK ] Reached target Network. Starting GSSAPI Proxy Daemon... Starting OpenSSH server daemon... Starting Dynamic System Tuning Daemon... [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... Starting Hostname Service... [ OK ] Started OpenSSH server daemon. [ OK ] Started Permit User Sessions. [ OK ] Started Serial Getty on ttyS1. [ OK ] Started Serial Getty on ttyS0. [ OK ] Started Command Scheduler. [ OK ] Started Getty on tty1. [ OK ] Reached target Login Prompts. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting Crash recovery kernel arming... Starting Notify NFS peers of a restart... Starting System Logging Service... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Starting Authorization Manager... [ OK ] Started Dynamic System Tuning Daemon. [ OK ] Reached target Multi-User System. [ OK ] Reached target Graphical Interface. Starting Update UTMP about System Runlevel Changes... [ OK ] Started Update UTMP about System Runlevel Changes. [ OK ] Started Authorization Manager. Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg325-server login: [ 47.707523] libcfs: loading out-of-tree module taints kernel. [ 47.726173] Key type ._llcrypt registered [ 47.727880] Key type .llcrypt registered [ 47.787467] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_hostid [ 56.705486] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing load_modules_local [ 57.417205] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 57.425932] alg: No test for adler32 (adler32-zlib) [ 59.039841] Lustre: Lustre: Build Version: 2.17.58_43_g6a42d2c [ 60.317567] LNet: Added LNI 192.168.203.125@tcp [8/256/0/180] [ 62.160258] Key type lgssc registered [ 64.492328] Lustre: Echo OBD driver; http://www.lustre.org/ [ 71.351080] hrtimer: interrupt took 7456765 ns [ 82.776643] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 119.996645] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing load_modules_local [ 132.790739] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 132.868347] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 134.193106] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 134.217353] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 134.290480] Lustre: lustre-MDT0000: new disk, initializing [ 134.362841] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 134.376901] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 139.525934] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 152.927367] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 153.035977] Lustre: 6504:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 153.106829] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 153.112427] Lustre: Skipped 1 previous similar message [ 153.209897] Lustre: lustre-MDT0001: new disk, initializing [ 153.301568] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 153.338423] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 153.344051] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 157.903144] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 162.648924] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 171.928322] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 172.170942] Lustre: lustre-OST0000: new disk, initializing [ 172.175980] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 172.183668] Lustre: 8443:0:(osd_compat.c:1353:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 172.258403] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 177.715212] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 177.729572] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 177.738411] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 177.841134] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 190.947469] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 191.071839] Lustre: lustre-OST0001: new disk, initializing [ 191.076781] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 191.081696] Lustre: 9513:0:(osd_compat.c:1353:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 191.139546] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 197.158888] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 197.171083] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 197.240310] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 197.746086] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 208.217536] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 213.203499] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 218.681767] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing check_logdir /tmp/testlogs/ [ 223.091319] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing yml_node [ 226.952310] Lustre: DEBUG MARKER: Client: 2.17.58.43 [ 229.367589] Lustre: DEBUG MARKER: MDS: 2.17.58.43 [ 231.976565] Lustre: DEBUG MARKER: OSS: 2.17.58.43 [ 233.575079] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-dual ============----- Sun Sep 6 23:47:13 EDT 2026 [ 249.658045] Lustre: DEBUG MARKER: excepting tests: 14b 21b [ 251.257694] Lustre: DEBUG MARKER: skipping tests SLOW=no: 21b [ 252.657592] Lustre: DEBUG MARKER: === replay-dual: start setup 23:47:32 (1788752852) === [ 259.051708] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing check_config_client /mnt/lustre [ 276.443656] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 279.730218] Lustre: 13372:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 284.528323] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 290.299941] Lustre: DEBUG MARKER: === replay-dual: finish setup 23:48:09 (1788752889) === [ 292.568879] Lustre: DEBUG MARKER: == replay-dual test 0a: expired recovery with lost client ========================================================== 23:48:11 (1788752891) [ 300.371727] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 304.657307] Lustre: Failing over lustre-MDT0000 [ 307.023419] Lustre: server umount lustre-MDT0000 complete [ 307.170180] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 307.179572] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 309.738647] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 309.753106] Lustre: Skipped 2 previous similar messages [ 314.691713] LustreError: 6512:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 314.716264] LustreError: 6512:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 8 previous similar messages [ 319.812218] LustreError: 6513:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 319.833066] LustreError: 6513:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 6 previous similar messages [ 324.940514] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 324.961043] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 5 previous similar messages [ 326.114659] Lustre: 3640:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788752910/real 1788752910] req@ffff9ce45ffdd180 x1875643109324416/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788752926 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 326.149115] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 329.059951] LDISKFS-fs (dm-0): 10 truncates cleaned up [ 329.064979] LDISKFS-fs (dm-0): recovery complete [ 329.080984] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 330.042978] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 330.058157] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 11 previous similar messages [ 335.177792] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 335.201991] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 7 previous similar messages [ 336.362866] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce445288000 x1875643109333504/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 336.700326] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 336.912216] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 340.561616] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 342.010722] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 443.500285] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 443.506198] Lustre: 14913:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client d645e103-2c89-4d79-8fb5-0a0ea1b2ef7d@192.168.203.25@tcp [ 443.531224] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 443.544859] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 443.546886] Lustre: 14913:0:(ldlm_lib.c:2989:target_recovery_thread()) too long recovery - read logs [ 443.552913] Lustre: Skipped 2 previous similar messages [ 443.562722] LustreError: dumping log to /tmp/lustre-log.1788753043.14913 [ 443.694483] Lustre: lustre-MDT0000: Recovery over after 1:47, of 3 clients 2 recovered and 1 was evicted. [ 443.734938] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:65) [ 443.736181] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:65) [ 463.871635] Lustre: DEBUG MARKER: == replay-dual test 0b: lost client during waiting for next transno ========================================================== 23:51:03 (1788753063) [ 470.599410] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 472.556224] Lustre: Failing over lustre-MDT0000 [ 472.749983] Lustre: server umount lustre-MDT0000 complete [ 474.593970] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 474.600799] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 474.610947] LustreError: 6516:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 474.624359] LustreError: 6516:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 4 previous similar messages [ 491.488367] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753075/real 1788753075] req@ffff9ce344eed500 x1875643109400192/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753091 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 491.510944] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 491.522899] LustreError: 6517:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 491.547190] LustreError: 6517:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 22 previous similar messages [ 493.552131] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 493.554508] LDISKFS-fs (dm-0): recovery complete [ 493.559798] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 500.706810] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce445288000 x1875643109408512/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 501.147798] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 503.107117] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 504.979491] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 506.357147] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 506.762993] Lustre: lustre-MDT0000: Client 99d9a105-66d9-4885-8009-b31c81e2a8d3 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:06 [ 517.366343] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:56 [ 522.553747] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:50 [ 527.673674] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:45 [ 532.796870] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:40 [ 537.914552] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:35 [ 542.194657] Lustre: lustre-MDT0001: haven't heard from client d645e103-2c89-4d79-8fb5-0a0ea1b2ef7d (at 192.168.203.25@tcp) in 102 seconds. I think it's dead, and I am evicting it. exp ffff9ce3432a5800, cur 1788753142 deadline 1788753140 last 1788753040 [ 548.154719] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:25 [ 548.186370] Lustre: Skipped 1 previous similar message [ 568.632435] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:04 [ 568.651226] Lustre: Skipped 3 previous similar messages [ 573.502343] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 573.513815] Lustre: 16652:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client d6f2fb16-ad43-4eb2-baed-b06618eb297d@ [ 573.528177] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 604.142047] Lustre: lustre-MDT0001: haven't heard from client 99d9a105-66d9-4885-8009-b31c81e2a8d3 (at 192.168.203.25@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff9ce343015800, cur 1788753204 deadline 1788753203 last 1788753103 [ 604.471443] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 1:10 [ 604.493486] Lustre: Skipped 6 previous similar messages [ 671.039669] Lustre: lustre-MDT0000: Denying connection for new client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 0:03 [ 671.058833] Lustre: Skipped 12 previous similar messages [ 674.500845] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 674.512635] Lustre: 16652:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 99d9a105-66d9-4885-8009-b31c81e2a8d3@192.168.203.25@tcp [ 674.523967] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 674.530290] Lustre: 16652:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 674.541609] Lustre: 16652:0:(ldlm_lib.c:2989:target_recovery_thread()) too long recovery - read logs [ 674.543724] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 674.553650] LustreError: dumping log to /tmp/lustre-log.1788753274.16652 [ 674.574948] Lustre: Skipped 2 previous similar messages [ 674.706113] Lustre: lustre-MDT0000: Recovery over after 2:51, of 3 clients 1 recovered and 2 were evicted. [ 674.742489] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:97) [ 674.746981] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:97) [ 683.855927] Lustre: DEBUG MARKER: == replay-dual test 1: |X| simple create ================= 23:54:43 (1788753283) [ 691.051763] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 692.927412] Lustre: Failing over lustre-MDT0000 [ 693.205942] Lustre: server umount lustre-MDT0000 complete [ 695.265621] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 695.274298] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 695.283399] Lustre: Skipped 3 previous similar messages [ 695.291511] LustreError: 7751:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 695.304668] LustreError: 7751:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 12 previous similar messages [ 711.648146] Lustre: 3637:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753296/real 1788753296] req@ffff9ce3420b6300 x1875643109498240/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753312 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 711.688634] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 715.468876] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 715.472173] LDISKFS-fs (dm-0): recovery complete [ 715.484801] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 721.928379] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b80b337 [ 721.937351] Lustre: MGC192.168.203.125@tcp: Connection restored to 0@lo (at 0@lo) [ 722.228950] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 722.276864] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 723.275543] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 726.283930] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 727.718083] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 727.768359] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:129) [ 727.771827] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:129) [ 735.326105] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 736.751824] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 745.218985] Lustre: DEBUG MARKER: == replay-dual test 2: |X| mkdir adir ==================== 23:55:44 (1788753344) [ 752.699270] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 754.430771] Lustre: Failing over lustre-MDT0000 [ 754.662493] Lustre: server umount lustre-MDT0000 complete [ 758.246614] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 758.269144] Lustre: Skipped 6 previous similar messages [ 763.362682] LustreError: 6516:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 763.385561] LustreError: 6516:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 44 previous similar messages [ 774.631331] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753358/real 1788753358] req@ffff9ce44f74e300 x1875643109533312/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753374 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 774.658204] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 776.749452] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 776.751962] LDISKFS-fs (dm-0): recovery complete [ 776.759235] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 785.261844] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 785.352935] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 786.315530] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 790.078561] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 790.515950] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 790.521208] Lustre: Skipped 4 previous similar messages [ 790.615893] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 790.662983] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:161) [ 790.664096] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:161) [ 799.914724] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 801.366249] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 810.239448] Lustre: DEBUG MARKER: == replay-dual test 3: |X| mkdir adir, mkdir adir/bdir === 23:56:49 (1788753409) [ 817.805128] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 820.060662] Lustre: Failing over lustre-MDT0000 [ 820.400439] Lustre: server umount lustre-MDT0000 complete [ 821.222805] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 821.242795] Lustre: Skipped 2 previous similar messages [ 837.611779] Lustre: 3637:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753421/real 1788753421] req@ffff9ce4607a9f80 x1875643109576064/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753437 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 837.642611] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 842.747783] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 842.752096] LDISKFS-fs (dm-0): recovery complete [ 842.777559] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 848.204121] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 848.454072] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 852.393749] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 853.483704] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 853.491766] Lustre: Skipped 3 previous similar messages [ 853.678385] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 853.748962] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:193) [ 853.751309] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:193) [ 863.053387] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 864.786176] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 874.003595] Lustre: DEBUG MARKER: == replay-dual test 4: |X| mkdir adir (-EEXIST), mkdir adir/bdir ========================================================== 23:57:53 (1788753473) [ 881.702755] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 883.546270] Lustre: Failing over lustre-MDT0000 [ 883.877313] Lustre: server umount lustre-MDT0000 complete [ 884.195945] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 884.206694] Lustre: Skipped 4 previous similar messages [ 894.265747] LustreError: 19007:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 894.298590] LustreError: 19007:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 87 previous similar messages [ 900.579976] Lustre: 3640:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753484/real 1788753484] req@ffff9ce345ad4000 x1875643109620352/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753500 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 900.607306] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 904.913927] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 904.920901] LDISKFS-fs (dm-0): recovery complete [ 904.944079] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 909.793885] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce3420b4e00 x1875643109628800/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 910.201750] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 910.209874] Lustre: Skipped 1 previous similar message [ 910.253955] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 911.257117] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 914.844948] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 915.440625] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 915.444192] Lustre: Skipped 3 previous similar messages [ 915.576251] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 915.607103] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:225) [ 915.607397] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:225) [ 923.980642] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 925.412674] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 933.557141] Lustre: DEBUG MARKER: == replay-dual test 5: open, unlink |X| close ============ 23:58:53 (1788753533) [ 941.141091] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 943.278474] Lustre: Failing over lustre-MDT0000 [ 943.523424] Lustre: server umount lustre-MDT0000 complete [ 946.149548] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 946.169874] Lustre: Skipped 1 previous similar message [ 962.528981] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753546/real 1788753546] req@ffff9ce45fbb5880 x1875643109659008/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753562 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 962.555834] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 964.881877] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 964.884231] LDISKFS-fs (dm-0): recovery complete [ 964.890301] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 971.755539] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b80cdc4 [ 972.227844] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 973.117460] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 975.854392] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 977.586022] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 977.655538] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:257) [ 977.655816] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:257) [ 986.746854] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 988.434313] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 997.962889] Lustre: DEBUG MARKER: == replay-dual test 6: open1, open2, unlink |X| close1 [fail mds1] close2 ========================================================== 23:59:57 (1788753597) [ 1004.864696] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1006.503979] Lustre: Failing over lustre-MDT0000 [ 1006.860358] Lustre: server umount lustre-MDT0000 complete [ 1024.484157] Lustre: 3637:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753608/real 1788753608] req@ffff9ce47178c000 x1875643109695104/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753624 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1024.504047] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1026.554319] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1026.556825] LDISKFS-fs (dm-0): recovery complete [ 1026.562962] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1033.699853] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce3420b4380 x1875643109703296/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1034.107592] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1034.558145] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1038.371257] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1039.353178] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1039.368221] Lustre: Skipped 8 previous similar messages [ 1039.441796] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1039.491195] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:289) [ 1039.491430] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:289) [ 1046.980242] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1048.160464] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1055.250167] Lustre: DEBUG MARKER: == replay-dual test 8: replay of resent request ========== 00:00:54 (1788753654) [ 1062.250049] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1063.158108] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1063.162333] LustreError: 10352:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce4793b3100 x1875643099000832/t38654705670(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:124/0 lens 512/448 e 0 to 0 dl 1788753674 ref 1 fl Interpret:/200/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1079.631174] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnecting [ 1079.666731] Lustre: 6512:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce460749880 x1875643099000832/t38654705670(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:141/0 lens 512/2880 e 0 to 0 dl 1788753691 ref 1 fl Interpret:/202/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1082.923475] Lustre: Failing over lustre-MDT0000 [ 1083.207372] Lustre: server umount lustre-MDT0000 complete [ 1085.412701] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1085.432235] Lustre: Skipped 9 previous similar messages [ 1101.792142] Lustre: 3640:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753685/real 1788753685] req@ffff9ce3431be680 x1875643109738368/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753701 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1101.821168] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1106.643968] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1106.645927] LDISKFS-fs (dm-0): recovery complete [ 1106.662193] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1112.041946] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce345ad5880 x1875643109747584/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1112.632398] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1112.647947] Lustre: Skipped 2 previous similar messages [ 1112.730355] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1114.544358] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1116.976967] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1117.746687] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 1117.780309] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:321) [ 1117.781633] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:321) [ 1125.251484] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1126.613579] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1134.627786] Lustre: DEBUG MARKER: == replay-dual test 9: resending a replayed create ======= 00:02:13 (1788753733) [ 1142.062648] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1144.797916] Lustre: Failing over lustre-MDT0000 [ 1145.128716] Lustre: server umount lustre-MDT0000 complete [ 1148.392598] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1150.266556] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1150.293269] LustreError: 6511:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 153 previous similar messages [ 1165.207786] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1165.210616] LDISKFS-fs (dm-0): recovery complete [ 1165.218442] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1173.993676] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b80df60 [ 1174.002824] Lustre: MGC192.168.203.125@tcp: Connection restored to 0@lo (at 0@lo) [ 1174.017713] Lustre: Skipped 7 previous similar messages [ 1174.309285] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1177.867502] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1179.668204] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1179.676290] LustreError: 32160:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce47178f800 x1875643099020032/t42949672962(42949672962) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:235/0 lens 528/448 e 0 to 0 dl 1788753785 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1190.741261] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:29 [ 1190.876686] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:353) [ 1190.877147] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:353) [ 1197.408968] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1199.086430] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1208.337297] Lustre: DEBUG MARKER: == replay-dual test 10: resending a replayed unlink ====== 00:03:27 (1788753807) [ 1215.566926] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1218.257652] Lustre: Failing over lustre-MDT0000 [ 1218.495626] Lustre: server umount lustre-MDT0000 complete [ 1220.592891] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1220.618995] Lustre: Skipped 6 previous similar messages [ 1235.937601] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788753820/real 1788753820] req@ffff9ce34bc93b80 x1875643109813120/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788753836 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1235.984119] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1235.998260] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1236.007717] LustreError: Skipped 1 previous similar message [ 1241.341442] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1241.345673] LDISKFS-fs (dm-0): recovery complete [ 1241.354714] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1246.176812] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce44f675180 x1875643109821696/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1246.486369] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1246.938952] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1246.947592] Lustre: Skipped 1 previous similar message [ 1251.011433] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1251.880802] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1251.890982] LustreError: 34216:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce44faf4380 x1875643099040128/t47244640260(47244640260) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:308/0 lens 528/448 e 0 to 0 dl 1788753858 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1262.427611] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:30 [ 1262.512427] Lustre: lustre-MDT0000: Recovery over after 0:16, of 3 clients 3 recovered and 0 were evicted. [ 1262.522071] Lustre: Skipped 1 previous similar message [ 1262.571066] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:385) [ 1262.573431] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:385) [ 1268.005542] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1269.339840] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1278.507405] Lustre: DEBUG MARKER: == replay-dual test 11: both clients timeout during replay ========================================================== 00:04:38 (1788753878) [ 1285.255482] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1288.263373] Lustre: Failing over lustre-MDT0000 [ 1288.594054] Lustre: server umount lustre-MDT0000 complete [ 1310.748667] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1310.750724] LDISKFS-fs (dm-0): recovery complete [ 1310.758617] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1319.420411] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b80ebfb [ 1324.348966] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1325.099891] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1325.106680] LustreError: 36270:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce345b27800 x1875643099058944/t51539607554(51539607554) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:381/0 lens 528/448 e 0 to 0 dl 1788753931 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1331.474967] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1336.649251] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1336.798703] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:417) [ 1336.799901] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:417) [ 1338.256950] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 5 sec [ 1344.841899] Lustre: DEBUG MARKER: == replay-dual test 12: open resend timeout ============== 00:05:44 (1788753944) [ 1351.252502] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1353.813252] Lustre: Failing over lustre-MDT0000 [ 1354.122703] Lustre: server umount lustre-MDT0000 complete [ 1374.759799] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1374.766913] LDISKFS-fs (dm-0): recovery complete [ 1374.776717] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1382.671319] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1382.675599] Lustre: Skipped 3 previous similar messages [ 1382.720307] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1382.732103] Lustre: Skipped 1 previous similar message [ 1387.031373] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1388.117153] Lustre: *** cfs_fail_loc=302, val=2147483648*** [ 1404.764733] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:23 [ 1404.859442] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:449) [ 1404.861426] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:449) [ 1410.706482] Lustre: DEBUG MARKER: == replay-dual test 13: close resend timeout ============= 00:06:50 (1788754010) [ 1417.669462] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1420.038492] Lustre: Failing over lustre-MDT0000 [ 1420.258410] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1420.301302] Lustre: server umount lustre-MDT0000 complete [ 1441.214181] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1441.216656] LDISKFS-fs (dm-0): recovery complete [ 1441.222336] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1454.619740] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1456.113272] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 1456.117956] Lustre: Skipped 17 previous similar messages [ 1456.162400] Lustre: *** cfs_fail_loc=115, val=2147483648*** [ 1472.335113] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnected, waiting for 3 clients in recovery for 1:24 [ 1472.447759] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:481) [ 1472.451247] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:481) [ 1477.804581] Lustre: DEBUG MARKER: SKIP: replay-dual test_14b skipping ALWAYS excluded test 14b [ 1479.039982] Lustre: DEBUG MARKER: == replay-dual test 15a: timeout waiting for lost client during replay, 1 client completes ========================================================== 00:07:58 (1788754078) [ 1485.487622] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1487.879117] Lustre: Failing over lustre-MDT0000 [ 1488.114070] Lustre: server umount lustre-MDT0000 complete [ 1491.947322] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1491.973272] Lustre: Skipped 14 previous similar messages [ 1508.321082] Lustre: 3638:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788754092/real 1788754092] req@ffff9ce45fc47b80 x1875643109957504/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788754108 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1508.343691] Lustre: 3638:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 1508.351844] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1508.364128] LustreError: Skipped 3 previous similar messages [ 1508.407676] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1508.419340] LDISKFS-fs (dm-0): recovery complete [ 1508.435129] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1520.084244] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1520.098192] Lustre: Skipped 3 previous similar messages [ 1522.895094] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1590.500208] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1590.506699] Lustre: 41937:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 80cb104f-02cc-4441-94ef-dc3b8a268294@ [ 1590.519294] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1591.076443] Lustre: lustre-MDT0000: Recovery over after 1:11, of 3 clients 2 recovered and 1 was evicted. [ 1591.083331] Lustre: Skipped 3 previous similar messages [ 1591.108107] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:513) [ 1591.110211] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:513) [ 1596.293381] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1597.717088] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1608.466223] Lustre: DEBUG MARKER: == replay-dual test 15c: remove multiple OST orphans ===== 00:10:07 (1788754207) [ 1615.646191] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1724.882881] Lustre: Failing over lustre-MDT0000 [ 1725.366767] Lustre: server umount lustre-MDT0000 complete [ 1726.962269] LustreError: 10352:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1726.983494] LustreError: 10352:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 234 previous similar messages [ 1746.347545] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1746.354859] LDISKFS-fs (dm-0): recovery complete [ 1746.365929] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1753.591162] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b82c7a0 [ 1753.603327] LustreError: 43928:0:(ldlm_resource.c:1207:ldlm_resource_complain()) MGC192.168.203.125@tcp: namespace resource [0x65727473756c:0x5:0x0].0x0 (ffff9ce4430aea00) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 1754.036114] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1754.052111] Lustre: Skipped 2 previous similar messages [ 1758.989917] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1825.503324] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1825.515124] Lustre: 43953:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client f503669e-1750-4d13-993e-e461bec9c5fd@ [ 1825.529940] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1825.604933] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:1537) [ 1825.606015] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:1537) [ 1830.439266] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1831.825613] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1840.432847] Lustre: DEBUG MARKER: == replay-dual test 16: fail MDS during recovery (3571) == 00:13:59 (1788754439) [ 1848.360566] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1851.468451] Lustre: Failing over lustre-MDT0000 [ 1851.730284] Lustre: server umount lustre-MDT0000 complete [ 1856.481205] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1873.041133] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1873.046803] LDISKFS-fs (dm-0): recovery complete [ 1873.057043] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1886.372066] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 1910.206023] Lustre: Failing over lustre-MDT0000 [ 1910.224770] LustreError: 46398:0:(ldlm_lib.c:3042:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 1910.231762] Lustre: 45924:0:(ldlm_lib.c:2442:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 1910.240892] Lustre: 45924:0:(ldlm_lib.c:1952:abort_req_replay_queue()) @@@ aborted: req@ffff9ce45ff3b100 x1875643101515648/t0(73014444033) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:215/0 lens 528/0 e 2 to 0 dl 1788754520 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1910.279532] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 1910.302801] Lustre: lustre-MDT0000: Not available for connect from 192.168.203.25@tcp (stopping) [ 1910.312727] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 1910.335817] LustreError: 45924:0:(client.c:1394:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9ce345ad4700 x1875643110164992/t0(0) o700->lustre-MDT0001-osp-MDT0000@0@lo:30/10 lens 264/248 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'tgt_recover_0.0' uid:0 gid:0 projid:4294967295 [ 1910.356797] LustreError: 45924:0:(fid_request.c:217:seq_client_alloc_seq()) cli-cli-lustre-MDT0001-osp-MDT0000: Cannot allocate new meta-sequence: rc = -5 [ 1910.367270] LustreError: 45924:0:(fid_request.c:321:seq_client_alloc_fid()) cli-cli-lustre-MDT0001-osp-MDT0000: Can't allocate new sequence: rc = -5 [ 1910.612786] Lustre: server umount lustre-MDT0000 complete [ 1929.810864] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1930.404081] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1930.409750] Lustre: Skipped 4 previous similar messages [ 1934.884251] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2001.500429] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2001.511498] Lustre: 46856:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 033823cc-ce9a-472a-b9ba-014f13e89205@ [ 2001.532116] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2002.041274] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2002.055160] Lustre: Skipped 18 previous similar messages [ 2002.073781] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1550 to 0x2c0000401:1569) [ 2002.079636] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1551 to 0x280000401:1569) [ 2007.709552] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2009.206824] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2017.902703] Lustre: DEBUG MARKER: == replay-dual test 17: fail OST during recovery (3571) == 00:16:57 (1788754617) [ 2026.390372] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 2028.360532] Lustre: Failing over lustre-OST0000 [ 2028.563664] Lustre: server umount lustre-OST0000 complete [ 2032.099297] LustreError: lustre-OST0000-osc-MDT0001: operation ost_statfs to node 0@lo failed: rc = -107 [ 2032.110560] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2032.138356] Lustre: Skipped 13 previous similar messages [ 2050.263774] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 2050.267263] LDISKFS-fs (dm-2): recovery complete [ 2050.277278] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2051.625159] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 2051.638230] Lustre: Skipped 3 previous similar messages [ 2055.900430] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2080.540699] Lustre: Failing over lustre-OST0000 [ 2080.553726] LustreError: 49367:0:(ldlm_lib.c:3042:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 2080.586080] Lustre: 48803:0:(ldlm_lib.c:2442:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2080.594293] Lustre: 48803:0:(ldlm_lib.c:2442:target_recovery_overseer()) Skipped 2 previous similar messages [ 2080.601585] LustreError: 48803:0:(ofd_obd.c:1325:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 2080.823575] Lustre: server umount lustre-OST0000 complete [ 2092.000854] Lustre: 3636:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788754652/real 1788754652] req@ffff9ce44fbb5c00 x1875643110236032/t0(0) o400->lustre-OST0000-osc-MDT0000@0@lo:28/4 lens 224/224 e 2 to 1 dl 1788754692 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 2092.048905] Lustre: 3636:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 2099.339454] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2107.827754] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2170.501440] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 2170.504934] Lustre: 49804:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client 73032a7a-4b2a-4471-9149-f10dfaa51726@ [ 2170.517281] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 2170.548329] Lustre: lustre-OST0000: Recovery over after 1:10, of 4 clients 3 recovered and 1 was evicted. [ 2170.560183] Lustre: Skipped 4 previous similar messages [ 2176.308407] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid 1475 0 [ 2177.971547] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 2187.590275] Lustre: DEBUG MARKER: == replay-dual test 18: ldlm_handle_enqueue succeeds on evicted export (3822) ========================================================== 00:19:46 (1788754786) [ 2191.438306] LustreError: 19007:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b sleeping for 40000ms [ 2231.449727] LustreError: 19007:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b awake [ 2243.421477] Lustre: DEBUG MARKER: == replay-dual test 19: resend of open request =========== 00:20:42 (1788754842) [ 2250.322038] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2251.627594] Lustre: *** cfs_fail_loc=157, val=2147483648*** [ 2251.640549] LustreError: 6513:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce45f9b1180 x1875643101637248/t0(0) o101->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:627/0 lens 576/688 e 0 to 0 dl 1788754932 ref 1 fl Interpret:/600/0 rc 0/0 job:'createmany.0' uid:0 gid:0 projid:0 [ 2337.597230] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnecting [ 2340.926828] Lustre: Failing over lustre-MDT0000 [ 2341.245213] Lustre: server umount lustre-MDT0000 complete [ 2342.390230] LustreError: 7751:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2342.423760] LustreError: 7751:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 102 previous similar messages [ 2357.731771] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2357.756063] LustreError: Skipped 3 previous similar messages [ 2364.301679] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2364.311710] LDISKFS-fs (dm-0): recovery complete [ 2364.328121] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2367.984371] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b832682 [ 2368.434182] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2368.453222] Lustre: Skipped 4 previous similar messages [ 2372.683989] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2373.644798] Lustre: 52462:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2373.820018] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1601) [ 2373.823638] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1601) [ 2382.426918] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2384.226746] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2391.331905] Lustre: DEBUG MARKER: == replay-dual test 20: recovery time is not increasing == 00:23:10 (1788754990) [ 2398.715459] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2401.165766] Lustre: Failing over lustre-MDT0000 [ 2401.405619] Lustre: server umount lustre-MDT0000 complete [ 2423.903934] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2423.906580] LDISKFS-fs (dm-0): recovery complete [ 2423.916364] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2430.951371] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce47178e680 x1875643110412800/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2436.083205] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2571.500245] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2571.506683] Lustre: 54405:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client e4c1a7be-52a2-4df2-b4cf-eb486d7d17da@ [ 2571.531808] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2571.577801] Lustre: 54405:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2571.585925] Lustre: 54405:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 6 previous similar messages [ 2571.700719] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1633) [ 2571.700719] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1603 to 0x2c0000401:1633) [ 2576.710290] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2577.926195] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2586.739638] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2588.451769] Lustre: Failing over lustre-MDT0000 [ 2588.646600] Lustre: server umount lustre-MDT0000 complete [ 2608.236313] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2608.238442] LDISKFS-fs (dm-0): recovery complete [ 2608.250457] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2615.781175] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce44fc56d80 x1875643110492800/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2616.151474] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2616.166826] Lustre: Skipped 4 previous similar messages [ 2620.669284] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2621.428729] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2621.434182] Lustre: Skipped 11 previous similar messages [ 2757.500106] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2757.505355] Lustre: 56185:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 7398d48d-5c68-499b-81fd-2d271f09c1e7@ [ 2757.516382] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2757.536172] Lustre: 56185:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2757.544262] Lustre: 56185:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 2757.652778] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1665) [ 2757.654957] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1665) [ 2763.748176] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2765.060875] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2776.086989] Lustre: DEBUG MARKER: == replay-dual test 21a: commit on sharing =============== 00:29:35 (1788755375) [ 2784.474850] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2786.692260] Lustre: Failing over lustre-MDT0000 [ 2786.947494] Lustre: server umount lustre-MDT0000 complete [ 2788.334256] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2788.348839] LustreError: Skipped 1 previous similar message [ 2788.359802] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2788.381125] Lustre: Skipped 13 previous similar messages [ 2805.668398] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788755390/real 1788755390] req@ffff9ce44f629c00 x1875643110570112/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788755406 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2805.697445] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2808.697989] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2808.702036] LDISKFS-fs (dm-0): recovery complete [ 2808.718172] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2815.975935] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b83382c [ 2815.995807] LustreError: 58199:0:(ldlm_resource.c:1207:ldlm_resource_complain()) MGC192.168.203.125@tcp: namespace resource [0x65727473756c:0x5:0x0].0x0 (ffff9ce465d1c000) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 2817.296889] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 2817.312832] Lustre: Skipped 4 previous similar messages [ 2820.173402] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 2957.500547] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2957.506487] Lustre: 58226:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 84714096-e0ac-4473-8bcc-b1436ada9c91@ [ 2957.513364] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2957.538170] Lustre: 58226:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2957.549571] Lustre: 58226:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 2957.560289] Lustre: lustre-MDT0000: Recovery over after 2:20, of 3 clients 2 recovered and 1 was evicted. [ 2957.564607] Lustre: Skipped 3 previous similar messages [ 2957.587073] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1697) [ 2957.587522] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1697) [ 2965.232347] Lustre: DEBUG MARKER: SKIP: replay-dual test_21b skipping SLOW test 21b [ 2966.640383] Lustre: DEBUG MARKER: == replay-dual test 22a: c1 lfs mkdir -i 1 dir1, M1 drop reply [ 2967.937823] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 2967.943828] LustreError: 6512:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce472288a80 x1875643101757056/t4294967346(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:588/0 lens 560/448 e 0 to 0 dl 1788755648 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 2970.351636] Lustre: Failing over lustre-MDT0001 [ 2970.665424] Lustre: server umount lustre-MDT0001 complete [ 2973.531336] LustreError: 19007:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2973.552229] LustreError: 19007:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 135 previous similar messages [ 2975.211955] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 2988.085115] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2988.564771] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 2988.572203] Lustre: Skipped 3 previous similar messages [ 2993.705171] Lustre: 14889:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce452a55f80 x1875643101757056/t4294967346(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:614/0 lens 560/2880 e 0 to 0 dl 1788755674 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 2993.962078] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3004.273982] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3006.464723] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3016.578307] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3019.377781] Lustre: Failing over lustre-MDT0000 [ 3019.611051] Lustre: lustre-MDT0000: Not available for connect from 192.168.203.25@tcp (stopping) [ 3019.619058] Lustre: Skipped 1 previous similar message [ 3019.762970] Lustre: server umount lustre-MDT0000 complete [ 3024.357469] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3040.721199] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3040.736971] LustreError: Skipped 3 previous similar messages [ 3042.250644] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3042.252814] LDISKFS-fs (dm-0): recovery complete [ 3042.261156] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3055.254984] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3056.663846] Lustre: 61290:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3056.785678] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1729) [ 3056.786405] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1729) [ 3064.372606] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3065.915899] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3073.577452] Lustre: DEBUG MARKER: == replay-dual test 22b: c1 lfs mkdir -i 1 d1, M1 drop reply [ 3074.739794] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3074.745564] LustreError: 10352:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce44f758380 x1875643101797376/t8589934617(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:695/0 lens 560/448 e 0 to 0 dl 1788755755 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3077.461909] Lustre: Failing over lustre-MDT0000 [ 3077.601952] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3077.615743] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3077.878606] Lustre: server umount lustre-MDT0000 complete [ 3081.635080] LustreError: 6497:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788755681 with bad export cookie 7181418748430205178 [ 3081.640159] Lustre: Failing over lustre-MDT0001 [ 3081.642376] LustreError: 6497:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 3081.669851] LustreError: 62432:0:(client.c:1394:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9ce471789c00 x1875643110717696/t0(0) o1000->lustre-MDT0000-osp-MDT0001@0@lo:24/4 lens 304/4320 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'umount.0' uid:0 gid:0 projid:4294967295 [ 3081.693874] LustreError: 62432:0:(osp_object.c:618:osp_attr_get()) lustre-MDT0000-osp-MDT0001: osp_attr_get update error [0x200000401:0x1:0x0]: rc = -5 [ 3082.086046] Lustre: server umount lustre-MDT0001 complete [ 3101.346656] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3101.405111] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3101.732616] LustreError: 63143:0:(llog.c:1655:llog_backup()) MGC192.168.203.125@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3101.740084] Lustre: 63143:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.203.125@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3112.281152] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3112.476382] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3114.066154] Lustre: 63171:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce47a860380 x1875643101797376/t8589934617(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:734/0 lens 560/2880 e 0 to 0 dl 1788755794 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3114.069840] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:36 to 0x280000400:65) [ 3114.073389] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:36 to 0x2c0000400:65) [ 3118.001883] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1761) [ 3118.002505] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1761) [ 3124.626449] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3127.861374] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3130.328845] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3142.487320] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3144.729875] Lustre: Failing over lustre-MDT0000 [ 3145.030926] Lustre: server umount lustre-MDT0000 complete [ 3148.775685] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3168.191944] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3168.195743] LDISKFS-fs (dm-0): recovery complete [ 3168.218166] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3174.886472] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce3441f4380 x1875643110763904/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3180.094450] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3180.578157] Lustre: 65362:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3180.595741] Lustre: 65362:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 3180.705694] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1793) [ 3180.706990] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1793) [ 3190.033628] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3191.300167] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3200.505434] Lustre: DEBUG MARKER: == replay-dual test 22c: c1 lfs mkdir -i 1 d1, M1 drop update [ 3201.640938] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3201.644501] LustreError: 8436:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce460749500 x1875643110793984/t107374182411(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:753/0 lens 2520/4320 e 0 to 0 dl 1788755813 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3205.629985] Lustre: Failing over lustre-MDT0000 [ 3206.231685] Lustre: server umount lustre-MDT0000 complete [ 3224.916840] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3231.713021] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce47a8bbb80 x1875643110803584/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3232.008664] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 3232.020074] Lustre: Skipped 6 previous similar messages [ 3236.862146] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3237.371603] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 3237.378650] Lustre: Skipped 25 previous similar messages [ 3237.479148] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1825) [ 3237.480501] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1825) [ 3246.736064] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3248.457624] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3259.818225] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3262.381582] Lustre: Failing over lustre-MDT0000 [ 3262.706262] Lustre: server umount lustre-MDT0000 complete [ 3262.951245] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3286.660916] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3286.665429] LDISKFS-fs (dm-0): recovery complete [ 3286.678411] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3299.274848] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3299.352422] Lustre: 68552:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3299.366710] Lustre: 68552:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 3299.521581] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1857) [ 3299.524469] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1857) [ 3309.318392] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3310.929713] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3320.376325] Lustre: DEBUG MARKER: == replay-dual test 22d: c1 lfs mkdir -i 1 d1, M1 drop update [ 3325.097629] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3325.101628] LustreError: 8437:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce45fc43100 x1875643110868864/t115964117002(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:121/0 lens 2520/4320 e 0 to 0 dl 1788755936 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3328.388756] Lustre: Failing over lustre-MDT0000 [ 3328.855212] Lustre: server umount lustre-MDT0000 complete [ 3332.671192] LustreError: 10332:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788755933 with bad export cookie 7181418748430212892 [ 3332.673220] Lustre: Failing over lustre-MDT0001 [ 3332.685272] LustreError: 10332:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 3332.706916] LustreError: 69789:0:(ldlm_resource.c:1207:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x79:0x0].0xf7117594 (ffff9ce45f668200) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3332.769987] Lustre: lustre-MDT0001: Not available for connect from 192.168.203.25@tcp (stopping) [ 3334.970521] Lustre: lustre-MDT0001: Not available for connect from 192.168.203.25@tcp (stopping) [ 3339.053476] Lustre: server umount lustre-MDT0001 complete [ 3359.942262] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3360.055533] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3378.080709] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce47a8baa00 x1875643110876032/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3382.450569] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3382.751356] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3384.684173] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1889) [ 3384.684202] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1889) [ 3399.762531] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 3399.762568] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 3399.786483] Lustre: 70516:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce3420b5c00 x1875643101888000/t12884901939(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:265/0 lens 560/2880 e 0 to 0 dl 1788756080 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3405.782687] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3407.309936] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3408.895922] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3418.932398] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3421.693827] Lustre: Failing over lustre-MDT0000 [ 3422.103408] Lustre: server umount lustre-MDT0000 complete [ 3423.209571] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3423.231570] Lustre: Skipped 36 previous similar messages [ 3438.993138] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788756023/real 1788756023] req@ffff9ce45fc47b80 x1875643110912896/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788756039 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3439.051953] Lustre: 3639:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 21 previous similar messages [ 3444.907741] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3444.911847] LDISKFS-fs (dm-0): recovery complete [ 3444.928463] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3450.336114] LustreError: 72711:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 3450.339175] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce45f9c7b80 x1875643110922496/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3450.396180] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x63a9864d8b836eff [ 3451.200839] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3451.210181] Lustre: Skipped 9 previous similar messages [ 3455.730567] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3456.055152] Lustre: 72747:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3456.074658] Lustre: 72747:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 3456.186868] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1921) [ 3456.191119] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1921) [ 3465.780242] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3467.439414] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3476.949820] Lustre: DEBUG MARKER: == replay-dual test 23a: c1 rmdir d1, M1 drop reply and fail, client2 mkdir d1 ========================================================== 00:41:16 (1788756076) [ 3478.724114] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3478.733747] LustreError: 70515:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce34cd5d180 x1875643101934848/t17179869210(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:344/0 lens 496/456 e 0 to 0 dl 1788756159 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3482.903677] Lustre: Failing over lustre-MDT0001 [ 3483.323671] Lustre: server umount lustre-MDT0001 complete [ 3486.707465] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3486.719235] LustreError: Skipped 1 previous similar message [ 3502.913639] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3508.805597] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:129) [ 3508.809279] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:129) [ 3508.830413] Lustre: 72205:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce344902680 x1875643101934848/t17179869210(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:374/0 lens 496/2888 e 0 to 0 dl 1788756189 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3509.212496] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3520.306548] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3521.941591] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3532.083816] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3534.632460] Lustre: Failing over lustre-MDT0000 [ 3535.161108] Lustre: server umount lustre-MDT0000 complete [ 3559.292907] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3559.298084] LDISKFS-fs (dm-0): recovery complete [ 3559.306810] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3565.540507] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce467c6ad80 x1875643110996224/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3565.828511] Lustre: lustre-MDT0000: Not available for connect from 192.168.203.25@tcp (not set up) [ 3565.840834] Lustre: Skipped 6 previous similar messages [ 3571.224326] Lustre: 75915:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3571.239248] Lustre: 75915:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 4 previous similar messages [ 3571.346252] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3571.381870] Lustre: lustre-MDT0000: Recovery over after 0:01, of 3 clients 3 recovered and 0 were evicted. [ 3571.394351] Lustre: Skipped 11 previous similar messages [ 3571.445257] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1953) [ 3571.446203] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1953) [ 3582.248175] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3584.296806] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3596.016387] Lustre: DEBUG MARKER: == replay-dual test 23b: c1 rmdir d1, M1 drop reply and fail M0/M1, c2 mkdir d1 ========================================================== 00:43:14 (1788756194) [ 3597.547070] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3597.557426] LustreError: 70516:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce3420b6d80 x1875643101973248/t21474836483(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:462/0 lens 496/456 e 0 to 0 dl 1788756277 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3601.737744] Lustre: Failing over lustre-MDT0000 [ 3601.777127] Lustre: lustre-MDT0000: Not available for connect from 192.168.203.25@tcp (stopping) [ 3601.811806] LustreError: 6498:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788756202 with bad export cookie 7181418748430220200 [ 3602.165146] Lustre: server umount lustre-MDT0000 complete [ 3605.835742] LustreError: 6496:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788756206 with bad export cookie 7181418748430220102 [ 3605.836975] Lustre: Failing over lustre-MDT0001 [ 3605.843734] LustreError: 6496:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 3606.192877] Lustre: server umount lustre-MDT0001 complete [ 3625.471067] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3625.645640] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3625.834732] LustreError: 77795:0:(llog.c:1655:llog_backup()) MGC192.168.203.125@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3625.845287] Lustre: 77795:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.203.125@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3626.247562] LustreError: 77817:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3626.268568] LustreError: 77817:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 315 previous similar messages [ 3632.082995] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3632.093670] Lustre: Skipped 11 previous similar messages [ 3637.213127] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3637.831716] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3638.398283] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:161) [ 3638.417971] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:161) [ 3638.461164] Lustre: 77818:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce44f759f80 x1875643101973248/t21474836483(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:503/0 lens 496/2888 e 0 to 0 dl 1788756318 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3641.347348] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1985) [ 3641.359261] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1985) [ 3648.206616] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3650.171770] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3651.727953] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3660.956369] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3663.017074] Lustre: Failing over lustre-MDT0000 [ 3663.423039] Lustre: server umount lustre-MDT0000 complete [ 3682.722768] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3682.741165] LustreError: Skipped 8 previous similar messages [ 3687.289655] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3687.292046] LDISKFS-fs (dm-0): recovery complete [ 3687.300059] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3698.108841] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3698.955834] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2017) [ 3698.956578] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2017) [ 3707.949952] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3709.539354] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3717.810684] Lustre: DEBUG MARKER: == replay-dual test 23c: c1 rmdir d1, M0 drop update reply and fail M0, c2 mkdir d1 ========================================================== 00:45:17 (1788756317) [ 3719.122973] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3719.127291] LustreError: 63886:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce445ea8000 x1875643111098624/t137438953491(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:515/0 lens 1744/4320 e 0 to 0 dl 1788756330 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3722.581262] Lustre: Failing over lustre-MDT0000 [ 3722.929344] Lustre: server umount lustre-MDT0000 complete [ 3740.279646] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3745.403861] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3745.875355] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2049) [ 3745.876620] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2049) [ 3754.798849] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3756.181650] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3766.000263] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3768.696604] Lustre: Failing over lustre-MDT0000 [ 3769.186813] Lustre: server umount lustre-MDT0000 complete [ 3771.377648] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3771.414737] LustreError: Skipped 2 previous similar messages [ 3792.050043] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3792.052264] LDISKFS-fs (dm-0): recovery complete [ 3792.058675] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3804.217207] Lustre: 83226:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3804.226962] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3804.231940] Lustre: 83226:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 17 previous similar messages [ 3804.487379] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2081) [ 3804.490590] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2081) [ 3815.088642] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3816.746534] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3826.354451] Lustre: DEBUG MARKER: == replay-dual test 23d: c1 rmdir d1, M0 drop update reply and fail M0/M1, c2 mkdir d1 ========================================================== 00:47:05 (1788756425) [ 3830.590573] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3830.597071] LustreError: 8436:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce465521c00 x1875643111172096/t146028888081(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:626/0 lens 1984/4320 e 0 to 0 dl 1788756441 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3834.192189] Lustre: Failing over lustre-MDT0000 [ 3834.494458] Lustre: server umount lustre-MDT0000 complete [ 3838.347902] LustreError: 6497:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788756438 with bad export cookie 7181418748430227326 [ 3838.354361] Lustre: Failing over lustre-MDT0001 [ 3838.361161] LustreError: 6497:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 3838.378758] LustreError: 84465:0:(ldlm_resource.c:1207:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x81:0x0].0x0 (ffff9ce4554f5e00) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3838.416669] Lustre: lustre-MDT0001: Not available for connect from 192.168.203.25@tcp (stopping) [ 3838.427350] Lustre: Skipped 4 previous similar messages [ 3844.506497] Lustre: server umount lustre-MDT0001 complete [ 3864.273958] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3864.437583] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3864.718873] LustreError: 85158:0:(llog.c:1655:llog_backup()) MGC192.168.203.125@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3864.725542] Lustre: 85158:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.203.125@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3883.496855] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3883.506579] Lustre: Skipped 11 previous similar messages [ 3884.035560] Lustre: lustre-MDT0001-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 3884.039796] Lustre: Skipped 43 previous similar messages [ 3888.357448] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2113) [ 3888.381052] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2113) [ 3888.558924] Lustre: 85181:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce47a861180 x1875643102052736/t25769803783(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:753/0 lens 496/2888 e 0 to 0 dl 1788756568 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3888.559623] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:193) [ 3888.564102] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:193) [ 3888.726858] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3889.865063] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3902.364973] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3904.290883] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3905.742354] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3917.286875] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3919.512638] Lustre: Failing over lustre-MDT0000 [ 3919.939497] Lustre: server umount lustre-MDT0000 complete [ 3943.924932] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3943.929773] LDISKFS-fs (dm-0): recovery complete [ 3943.939153] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3951.106991] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 3951.766757] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2115 to 0x2c0000401:2145) [ 3951.767957] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2115 to 0x280000401:2145) [ 3960.887617] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3962.649845] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3971.343223] Lustre: DEBUG MARKER: == replay-dual test 24: reconstruct on non-existing object ========================================================== 00:49:30 (1788756570) [ 3972.948434] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3972.958989] LustreError: 85179:0:(ldlm_lib.c:3386:target_send_reply_msg()) @@@ dropping reply req@ffff9ce44f40b480 x1875643102094336/t154618822673(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:83/0 lens 488/456 e 0 to 0 dl 1788756653 ref 1 fl Interpret:/200/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4058.081130] Lustre: lustre-MDT0000: Client 602be7b2-4b89-43fd-b603-a1c8727cdde1 (at 192.168.203.25@tcp) reconnecting [ 4058.098568] Lustre: 85255:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9ce445b0ad80 x1875643102094336/t154618822673(0) o36->602be7b2-4b89-43fd-b603-a1c8727cdde1@192.168.203.25@tcp:168/0 lens 488/3152 e 0 to 0 dl 1788756738 ref 1 fl Interpret:/202/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4066.142039] Lustre: DEBUG MARKER: == replay-dual test 25: replay|resend ==================== 00:51:05 (1788756665) [ 4068.735509] Lustre: *** cfs_fail_loc=304, val=0*** [ 4071.601193] Lustre: Failing over lustre-OST0000 [ 4071.799190] Lustre: server umount lustre-OST0000 complete [ 4072.947045] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4072.961493] Lustre: Skipped 35 previous similar messages [ 4090.030955] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4091.814403] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 4091.826600] Lustre: Skipped 10 previous similar messages [ 4097.040448] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4107.067606] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid 1475 0 [ 4108.617584] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4117.804408] Lustre: DEBUG MARKER: == replay-dual test 26: dbench and tar with mds failover ========================================================== 00:51:57 (1788756717) [ 4129.996119] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4133.938614] Lustre: DEBUG MARKER: test_26 fail mds1 1 times [ 4136.384051] Lustre: Failing over lustre-MDT0000 [ 4136.436062] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4136.445390] Lustre: Skipped 6 previous similar messages [ 4138.759212] Lustre: server umount lustre-MDT0000 complete [ 4156.897503] Lustre: 3640:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788756741/real 1788756741] req@ffff9ce34d541c00 x1875643111372672/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788756757 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4156.939627] Lustre: 3640:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 17 previous similar messages [ 4161.361517] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 4161.363786] LDISKFS-fs (dm-0): recovery complete [ 4161.391817] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4167.140031] LustreError: 3636:0:(client.c:1404:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9ce3431ddc00 x1875643111379584/t0(0) o250->MGC192.168.203.125@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4172.794883] Lustre: 91322:0:(ldlm_lib.c:2123:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4172.815278] Lustre: 91322:0:(ldlm_lib.c:2123:extend_recovery_timer()) Skipped 17 previous similar messages [ 4173.836303] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4174.717556] Lustre: lustre-MDT0000: Recovery over after 0:06, of 3 clients 3 recovered and 0 were evicted. [ 4174.728732] Lustre: Skipped 9 previous similar messages [ 4174.817578] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2172 to 0x2c0000401:2209) [ 4174.818676] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2173 to 0x280000401:2209) [ 4184.527062] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 4186.530986] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4198.825642] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4202.797834] Lustre: DEBUG MARKER: test_26 fail mds2 2 times [ 4205.041604] Lustre: Failing over lustre-MDT0001 [ 4205.093313] Lustre: lustre-MDT0001: Not available for connect from 192.168.203.25@tcp (stopping) [ 4205.102560] Lustre: Skipped 3 previous similar messages [ 4205.575647] Lustre: server umount lustre-MDT0001 complete [ 4228.169841] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4228.171317] LDISKFS-fs (dm-1): recovery complete [ 4228.177427] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4233.622949] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4237.314334] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:287 to 0x280000400:321) [ 4237.317118] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:288 to 0x2c0000400:321) [ 4247.166244] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4248.986784] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4262.341156] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4266.429789] Lustre: DEBUG MARKER: test_26 fail mds1 3 times [ 4268.776421] Lustre: Failing over lustre-MDT0000 [ 4275.720351] Lustre: server umount lustre-MDT0000 complete [ 4276.555050] LustreError: 85179:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.203.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4276.578686] LustreError: 85179:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 275 previous similar messages [ 4296.678459] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4296.689743] LustreError: Skipped 5 previous similar messages [ 4299.810591] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4299.812744] LDISKFS-fs (dm-0): recovery complete [ 4299.824337] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4307.378947] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 4307.391664] Lustre: Skipped 10 previous similar messages [ 4311.810917] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4313.804132] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2271 to 0x280000401:2305) [ 4313.804264] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2271 to 0x2c0000401:2305) [ 4322.780239] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 4325.013399] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4385.829183] Lustre: DEBUG MARKER: == replay-dual test 28: lock replay should be ordered: waiting after granted ========================================================== 00:56:24 (1788756984) [ 4406.019991] Lustre: Failing over lustre-OST0000 [ 4406.183247] Lustre: server umount lustre-OST0000 complete [ 4406.755674] LustreError: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 4406.783261] LustreError: Skipped 5 previous similar messages [ 4426.780102] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4428.757514] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4435.535623] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4448.859601] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid 1475 0 [ 4450.915696] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4465.351103] Lustre: DEBUG MARKER: == replay-dual test 29: replay vs update with the same xid ========================================================== 00:57:43 (1788757063) [ 4467.734106] Lustre: DEBUG MARKER: SKIP: replay-dual test_29 needs >= 2 clients [ 4469.935921] Lustre: DEBUG MARKER: == replay-dual test 30: layout lock replay is not blocked on IO ========================================================== 00:57:49 (1788757069) [ 4473.527352] Lustre: Failing over lustre-MDT0000 [ 4474.237632] Lustre: server umount lustre-MDT0000 complete [ 4493.808711] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4494.226996] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 4494.231750] Lustre: Skipped 7 previous similar messages [ 4499.220558] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4499.461363] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 4499.477019] Lustre: Skipped 24 previous similar messages [ 4499.543580] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2326 to 0x2c0000401:2369) [ 4499.543902] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2328 to 0x280000401:2369) [ 4508.202441] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 4509.856729] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4518.550113] Lustre: DEBUG MARKER: == replay-dual test 31: deadlock on file_remove_privs and occupied mod rpc slots ========================================================== 00:58:37 (1788757117) [ 4522.510256] Lustre: Failing over lustre-OST0000 [ 4522.920805] Lustre: server umount lustre-OST0000 complete [ 4541.889610] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4549.607844] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4559.447095] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid 1475 0 [ 4560.964180] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in IDLE [ 4569.983984] Lustre: DEBUG MARKER: == replay-dual test 32: gap in update llog shouldn't break recovery ========================================================== 00:59:29 (1788757169) [ 4571.148047] Lustre: *** cfs_fail_loc=131d, val=10*** [ 4571.693099] Lustre: *** cfs_fail_loc=131d, val=4294967294*** [ 4571.700718] Lustre: Skipped 11 previous similar messages [ 4572.709657] Lustre: *** cfs_fail_loc=131d, val=4294967276*** [ 4572.711936] Lustre: Skipped 17 previous similar messages [ 4575.047662] Lustre: Failing over lustre-MDT0001 [ 4575.410426] Lustre: server umount lustre-MDT0001 complete [ 4579.247911] Lustre: Failing over lustre-MDT0000 [ 4579.696368] Lustre: server umount lustre-MDT0000 complete [ 4588.246082] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4588.798303] Lustre: *** cfs_fail_loc=131d, val=4294967266*** [ 4588.802759] Lustre: Skipped 9 previous similar messages [ 4593.992755] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4601.993021] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4602.195588] Lustre: *** cfs_fail_loc=131d, val=4294967262*** [ 4602.199084] Lustre: Skipped 3 previous similar messages [ 4607.576360] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4607.623224] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:379 to 0x2c0000400:417) [ 4607.624655] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:379 to 0x280000400:417) [ 4607.711609] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2409 to 0x280000401:2497) [ 4607.712502] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2326 to 0x2c0000401:2401) [ 4623.880779] Lustre: DEBUG MARKER: == replay-dual test 33: Check for OBD_INCOMPAT_MULTI_RPCS in last_rcvd after abort_recovery ========================================================== 01:00:22 (1788757222) [ 4631.622976] Lustre: Failing over lustre-MDT0001 [ 4631.960632] Lustre: server umount lustre-MDT0001 complete [ 4652.935911] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4657.966598] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4666.348633] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount REPLAY_WAIT mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4667.745757] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in REPLAY_WAIT state after 0 sec [ 4668.753930] Lustre: lustre-MDT0001: Aborting client recovery [ 4668.757651] LustreError: 104161:0:(ldlm_lib.c:3042:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 4668.763891] Lustre: 103516:0:(ldlm_lib.c:2442:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 4668.770720] Lustre: 103516:0:(ldlm_lib.c:2442:target_recovery_overseer()) Skipped 2 previous similar messages [ 4668.780945] Lustre: 103516:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 551fc322-0c21-4eb2-85de-3a635a685350@ [ 4668.790797] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 4668.806602] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 4668.818446] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 4668.905051] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:379 to 0x2c0000400:449) [ 4668.905692] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:379 to 0x280000400:449) [ 4674.388785] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4675.789309] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4679.613459] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4686.673775] Lustre: Failing over lustre-MDT0001 [ 4687.001631] Lustre: server umount lustre-MDT0001 complete [ 4689.383718] Lustre: lustre-MDT0001-lwp-OST0001: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4689.390092] Lustre: Skipped 30 previous similar messages [ 4697.520617] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4700.270896] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 4700.280653] Lustre: Skipped 9 previous similar messages [ 4703.753964] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug -1 all [ 4703.795260] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:379 to 0x2c0000400:481) [ 4703.797037] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:379 to 0x280000400:481) [ 4714.038346] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4715.816187] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4720.324857] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4728.772730] Lustre: DEBUG MARKER: == replay-dual test complete, duration 4494 sec ========== 01:02:08 (1788757328) [ 4730.123338] Lustre: DEBUG MARKER: === replay-dual: start cleanup 01:02:09 (1788757329) === [ 4745.951771] Lustre: DEBUG MARKER: === replay-dual: finish cleanup 01:02:24 (1788757344) === [ 4748.329717] Lustre: Failing over lustre-MDT0000 [ 4749.210365] Lustre: server umount lustre-MDT0000 complete [ 4766.177949] Lustre: 3638:0:(client.c:2504:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1788757350/real 1788757350] req@ffff9ce450396300 x1875643112526208/t0(0) o400->MGC192.168.203.125@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1788757366 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4766.209803] Lustre: 3638:0:(client.c:2504:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 4775.765053] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4781.739828] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4918.500123] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 4918.504266] Lustre: 107329:0:(genops.c:1600:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 551fc322-0c21-4eb2-85de-3a635a685350@ [ 4918.526281] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 4918.548468] Lustre: lustre-MDT0000: Recovery over after 2:20, of 3 clients 2 recovered and 1 was evicted. [ 4918.565656] Lustre: Skipped 9 previous similar messages [ 4918.603822] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2326 to 0x2c0000401:2433) [ 4918.609171] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2409 to 0x280000401:2529) [ 4924.997342] Lustre: DEBUG MARKER: oleg325-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 4926.365769] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4934.114788] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4934.118864] Lustre: Skipped 12 previous similar messages [ 4938.384764] Lustre: server umount lustre-MDT0000 complete [ 4940.775864] LustreError: 91944:0:(ldlm_lib.c:1202:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4940.801312] LustreError: 91944:0:(ldlm_lib.c:1202:target_handle_connect()) Skipped 175 previous similar messages [ 4947.123984] LustreError: 7453:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1788757547 with bad export cookie 7181418748430414072 [ 4947.137103] LustreError: MGC192.168.203.125@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4947.141665] LustreError: 7453:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 4947.169895] LustreError: Skipped 3 previous similar messages [ 4947.535475] Lustre: server umount lustre-MDT0001 complete [ 4965.708977] Lustre: server umount lustre-OST0000 complete [ 4984.749815] Lustre: server umount lustre-OST0001 complete [ 5002.589576] Lustre: DEBUG MARKER: oleg325-server.virtnet: executing unload_modules_local [ 5005.673831] Key type lgssc unregistered [ 5006.020506] LNet: 110270:0:(lib-ptl.c:964:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 5006.031078] LNetError: 110270:0:(acceptor.c:252:lnet_acceptor_remove_socket()) Interface ens2 not found [ 5006.055395] LNet: Removed LNI 192.168.203.125@tcp [ 5007.130193] Key type .llcrypt unregistered [ 5007.133741] Key type ._llcrypt unregistered