[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-10.fc44 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 490618498 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2399.988 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE247C 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE2318 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 0022D8 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE238C 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE241C 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2454 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe2318-0xbffe238b] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe2317] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe238c-0xbffe241b] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe241c-0xbffe2453] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2454-0xbffe247b] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524584K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001011] APIC: Switch to symmetric I/O mode setup [ 0.003228] x2apic enabled [ 0.004005] Switched APIC routing to physical x2apic. [ 0.005014] kvm-guest: setup PV IPIs [ 0.008000] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.008000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x22982c12b6d, max_idle_ns: 440795281273 ns [ 0.008023] Calibrating delay loop (skipped) preset value.. 4799.97 BogoMIPS (lpj=2399988) [ 0.009016] pid_max: default: 32768 minimum: 301 [ 0.010132] LSM: Security Framework initializing [ 0.012058] Yama: becoming mindful. [ 0.012902] SELinux: Initializing. [ 0.013129] *** VALIDATE selinux *** [ 0.021791] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.026344] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.027171] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.028110] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.029134] *** VALIDATE tmpfs *** [ 0.031201] *** VALIDATE proc *** [ 0.032257] *** VALIDATE cgroup *** [ 0.033013] *** VALIDATE cgroup2 *** [ 0.034332] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.035180] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.036016] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.037036] Spectre V2 : User space: Vulnerable [ 0.038010] Speculative Store Bypass: Vulnerable [ 0.041726] debug: unmapping init [mem 0xffffffff98659000-0xffffffff98660fff] [ 0.043301] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.044788] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.045039] ... version: 2 [ 0.046013] ... bit width: 48 [ 0.047013] ... generic registers: 4 [ 0.048015] ... value mask: 0000ffffffffffff [ 0.049016] ... max period: 00007fffffffffff [ 0.050017] ... fixed-purpose events: 3 [ 0.051016] ... event mask: 000000070000000f [ 0.052421] rcu: Hierarchical SRCU implementation. [ 0.054701] smp: Bringing up secondary CPUs ... [ 0.055728] x86: Booting SMP configuration: [ 0.056029] .... node #0, CPUs: #1 #2 #3 [ 0.059367] smp: Brought up 1 node, 4 CPUs [ 0.061017] smpboot: Max logical packages: 1 [ 0.062021] smpboot: Total of 4 processors activated (19199.90 BogoMIPS) [ 0.116245] node 0 deferred pages initialised in 52ms [ 0.118260] devtmpfs: initialized [ 0.119219] x86/mm: Memory block size: 128MB [ 0.121586] gcov: version magic: 0x41383552 [ 0.122683] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.126081] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.129333] pinctrl core: initialized pinctrl subsystem [ 0.131201] [ 0.131891] ************************************************************* [ 0.134011] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.136014] ** ** [ 0.139011] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.141008] ** ** [ 0.142007] ** This means that this kernel is built to expose internal ** [ 0.143000] ** IOMMU data structures, which may compromise security on ** [ 0.145010] ** your system. ** [ 0.147010] ** ** [ 0.148009] ** If you see this message and you are not debugging the ** [ 0.150013] ** kernel, report this immediately to your vendor! ** [ 0.153011] ** ** [ 0.155010] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.158012] ************************************************************* [ 0.160717] NET: Registered protocol family 16 [ 0.162445] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.165055] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.167056] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.173066] cpuidle: using governor menu [ 0.175420] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.176455] PCI: Using configuration type 1 for base access [ 0.177091] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.185139] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.187076] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.192061] cryptd: max_cpu_qlen set to 1000 [ 0.196207] ACPI: Added _OSI(Module Device) [ 0.197000] ACPI: Added _OSI(Processor Device) [ 0.198015] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.200012] ACPI: Added _OSI(Processor Aggregator Device) [ 0.203398] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.209647] ACPI: Interpreter enabled [ 0.211052] ACPI: PM: (supports S0 S3 S4 S5) [ 0.212011] ACPI: Using IOAPIC for interrupt routing [ 0.213090] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.217100] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.226225] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.228042] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.232023] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.234127] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.241516] acpiphp: Slot [2] registered [ 0.243132] acpiphp: Slot [5] registered [ 0.244101] acpiphp: Slot [6] registered [ 0.245154] acpiphp: Slot [7] registered [ 0.247122] acpiphp: Slot [8] registered [ 0.248130] acpiphp: Slot [9] registered [ 0.249130] acpiphp: Slot [10] registered [ 0.250112] acpiphp: Slot [3] registered [ 0.252104] acpiphp: Slot [4] registered [ 0.253113] acpiphp: Slot [11] registered [ 0.255104] acpiphp: Slot [12] registered [ 0.256126] acpiphp: Slot [13] registered [ 0.258116] acpiphp: Slot [14] registered [ 0.259108] acpiphp: Slot [15] registered [ 0.261102] acpiphp: Slot [16] registered [ 0.262124] acpiphp: Slot [17] registered [ 0.264119] acpiphp: Slot [18] registered [ 0.265101] acpiphp: Slot [19] registered [ 0.267107] acpiphp: Slot [20] registered [ 0.269109] acpiphp: Slot [21] registered [ 0.270163] acpiphp: Slot [22] registered [ 0.271000] acpiphp: Slot [23] registered [ 0.271000] acpiphp: Slot [24] registered [ 0.273117] acpiphp: Slot [25] registered [ 0.274125] acpiphp: Slot [26] registered [ 0.276103] acpiphp: Slot [27] registered [ 0.277100] acpiphp: Slot [28] registered [ 0.279115] acpiphp: Slot [29] registered [ 0.281108] acpiphp: Slot [30] registered [ 0.282105] acpiphp: Slot [31] registered [ 0.284066] PCI host bridge to bus 0000:00 [ 0.286022] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.288024] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.291031] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.294030] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.296022] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.299023] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.301199] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.303986] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.307394] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.318014] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.323598] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.326021] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.328020] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.331020] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.336591] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.339876] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.343045] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.347098] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.352015] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.365016] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.371015] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.378395] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.387020] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.398016] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.419020] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.438783] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.451016] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.459018] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.474020] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.483594] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.494024] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.502024] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.529023] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.538711] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.546018] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.553016] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.575018] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.584277] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.592025] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.599028] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.626029] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.641515] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.650031] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.660022] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.684022] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.696873] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.699447] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.702451] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.707454] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.710249] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.716040] iommu: Default domain type: Passthrough [ 0.718555] SCSI subsystem initialized [ 0.719167] ACPI: bus type USB registered [ 0.721133] usbcore: registered new interface driver usbfs [ 0.723133] usbcore: registered new interface driver hub [ 0.727163] usbcore: registered new device driver usb [ 0.729641] pps_core: LinuxPPS API ver. 1 registered [ 0.731014] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.735092] PTP clock support registered [ 0.737104] EDAC MC: Ver: 3.0.0 [ 0.739135] PCI: Using ACPI for IRQ routing [ 0.741735] NetLabel: Initializing [ 0.743024] NetLabel: domain hash size = 128 [ 0.744013] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.746109] NetLabel: unlabeled traffic allowed by default [ 0.748194] vgaarb: loaded [ 0.750332] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.751020] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.757000] clocksource: Switched to clocksource kvm-clock [ 0.869836] VFS: Disk quotas dquot_6.6.0 [ 0.871489] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 0.874805] *** VALIDATE ramfs *** [ 0.876095] *** VALIDATE hugetlbfs *** [ 0.877537] pnp: PnP ACPI init [ 0.879719] pnp: PnP ACPI: found 6 devices [ 0.896571] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 0.900205] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 0.902704] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 0.904889] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 0.907655] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 0.910306] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 0.913631] NET: Registered protocol family 2 [ 0.916149] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 0.920987] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 0.924755] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 0.930209] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 0.933881] TCP: Hash tables configured (established 65536 bind 65536) [ 0.936723] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 0.939761] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.942768] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.945936] NET: Registered protocol family 1 [ 0.952301] RPC: Registered named UNIX socket transport module. [ 0.954879] RPC: Registered udp transport module. [ 0.956317] RPC: Registered tcp transport module. [ 0.957853] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 0.959958] NET: Registered protocol family 44 [ 0.961465] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 0.963325] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 0.965050] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 0.967156] PCI: CLS 0 bytes, default 64 [ 0.969268] Unpacking initramfs... [ 2.367069] debug: unmapping init [mem 0xffff93953cc54000-0xffff93953ffbffff] [ 2.373306] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 2.375748] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 2.379162] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x22982c12b6d, max_idle_ns: 440795281273 ns [ 2.887214] Initialise system trusted keyrings [ 2.889129] Key type blacklist registered [ 2.891526] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 2.900637] zbud: loaded [ 2.903633] *** VALIDATE nfs *** [ 2.905103] *** VALIDATE nfs4 *** [ 2.906935] pstore: using deflate compression [ 2.913336] Platform Keyring initialized [ 3.021115] NET: Registered protocol family 38 [ 3.023452] Key type asymmetric registered [ 3.025155] Asymmetric key parser 'x509' registered [ 3.027320] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 3.030697] io scheduler mq-deadline registered [ 3.032592] io scheduler kyber registered [ 3.034420] io scheduler bfq registered [ 3.036385] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 3.040084] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 3.043252] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 3.046529] ACPI: Power Button [PWRF] [ 3.052236] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 3.059723] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 3.071461] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 3.078045] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 3.096471] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 3.127419] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 3.157275] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 3.161210] Non-volatile memory driver v1.3 [ 3.162840] Linux agpgart interface v0.103 [ 3.197072] virtio_blk virtio1: [vda] 136600 512-byte logical blocks (69.9 MB/66.7 MiB) [ 3.200408] vda: detected capacity change from 0 to 69939200 [ 3.219375] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 3.231095] vdb: detected capacity change from 0 to 1073741824 [ 3.254946] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.262775] vdc: detected capacity change from 0 to 2621440000 [ 3.303773] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.308248] vdd: detected capacity change from 0 to 2621440000 [ 3.343191] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.355502] vde: detected capacity change from 0 to 4294967296 [ 3.397732] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.402439] vdf: detected capacity change from 0 to 4294967296 [ 3.426781] libphy: Fixed MDIO Bus: probed [ 3.479151] usbcore: registered new interface driver usbserial_generic [ 3.481610] usbserial: USB Serial support registered for generic [ 3.484078] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 3.489221] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 3.491140] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 3.493472] mousedev: PS/2 mouse device common for all mice [ 3.497129] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 3.498321] rtc_cmos 00:05: RTC can wake from S4 [ 3.508182] rtc_cmos 00:05: registered as rtc0 [ 3.512929] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 3.518281] intel_pstate: CPU model not supported [ 3.522710] hid: raw HID events driver (C) Jiri Kosina [ 3.524640] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 3.528760] usbcore: registered new interface driver usbhid [ 3.528768] usbhid: USB HID core driver [ 3.529889] drop_monitor: Initializing network drop monitor service [ 3.531130] Initializing XFRM netlink socket [ 3.552835] NET: Registered protocol family 10 [ 3.556203] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 3.558694] Segment Routing with IPv6 [ 3.568592] NET: Registered protocol family 17 [ 3.571305] mpls_gso: MPLS GSO support [ 3.579174] RAS: Correctable Errors collector initialized. [ 3.585269] AVX version of gcm_enc/dec engaged. [ 3.592627] AES CTR mode by8 optimization enabled [ 3.763557] sched_clock: Marking stable (3763524720, 0)->(4681411056, -917886336) [ 3.770457] registered taskstats version 1 [ 3.774158] Loading compiled-in X.509 certificates [ 3.777573] zswap: loaded using pool lzo/zbud [ 3.846615] Key type big_key registered [ 3.904331] Key type encrypted registered [ 3.906894] ima: No TPM chip found, activating TPM-bypass! [ 3.915084] ima: Allocated hash algorithm: sha1 [ 3.921684] ima: No architecture policies found [ 3.925613] evm: Initialising EVM extended attributes: [ 3.928301] evm: security.selinux [ 3.929874] evm: security.ima [ 3.931912] evm: security.capability [ 3.933353] evm: HMAC attrs: 0x1 [ 3.938980] rtc_cmos 00:05: setting system clock to 2026-06-12 03:25:15 UTC (1781234715) [ 3.968414] debug: unmapping init [mem 0xffffffff99603000-0xffffffff997fffff] [ 3.985445] debug: unmapping init [mem 0xffffffff98382000-0xffffffff98658fff] [ 3.993186] Write protecting the kernel read-only data: 28672k [ 4.007780] debug: unmapping init [mem 0xffffffff96a03000-0xffffffff96bfffff] [ 4.017621] debug: unmapping init [mem 0xffffffff97314000-0xffffffff973fffff] [ 4.160461] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 4.185729] systemd[1]: Detected virtualization kvm. [ 4.193267] systemd[1]: Detected architecture x86-64. [ 4.200209] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 4.240867] systemd[1]: No hostname configured. [ 4.243442] systemd[1]: Set hostname to . [ 4.249140] random: systemd: uninitialized urandom read (16 bytes read) [ 4.253035] systemd[1]: Initializing machine ID from random generator. [ 4.454476] random: ln: uninitialized urandom read (6 bytes read) [ 4.629000] random: systemd: uninitialized urandom read (16 bytes read) [ 4.636769] systemd[1]: Started Dispatch Password Requests to Console Directory Watch. [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ 4.655643] systemd[1]: Listening on Journal Socket (/dev/log). [ OK ] Listening on Journal Socket (/dev/log). [ 4.674259] systemd[1]: Listening on udev Control Socket. [ OK ] Listening on udev Control Socket. [ OK ] Listening on Journal Socket. Starting Journal Service... [ OK ] Started Memstrack Anylazing Service. Starting Setup Virtual Console... [ OK ] Reached target Timers. [ OK ] Reached target Local Encrypted Volumes. Starting Apply Kernel Variables... [ OK ] Reached target Slices. [ OK ] Reached target Paths. [ OK ] Reached target Initrd Root Device. [ OK ] Listening on udev Kernel Socket. [ OK ] Reached target Sockets. Starting Create list of required st…ce nodes for the current kernel... [ OK ] Reached target Local File Systems. Starting Create Volatile Files and Directories... [ OK ] Reached target Swap. [ OK ] Started Setup Virtual Console. [ OK ] Started Apply Kernel Variables. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Create Volatile Files and Directories. Starting Create Static Device Nodes in /dev... Starting dracut cmdline hook... [ OK ] Started Journal Service. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 6.393847] device-mapper: uevent: version 1.0.3 [ 6.397915] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. Starting dracut initqueue hook... [ 8.149032] random: fast init done [ 8.161534] virtio_net virtio0 ens2: renamed from eth0 [ 8.213128] scsi host0: ata_piix [ 8.282498] scsi host1: ata_piix [ 8.284325] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 8.290157] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 13.449530] random: crng init done [ 13.455666] random: 7 urandom warning(s) missed due to ratelimiting [ 14.252128] dracut-initqueue[591]: RTNETLINK answers: File exists Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Mounting /sysroot... [ 15.709630] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Timers. [ OK ] Stopped target Initrd Root Device. Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Slices. [ OK ] Stopped target Paths. [ OK ] Stopped target Sockets. [ OK ] Stopped target System Initialization. [ OK ] Stopped target Swap. [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. Stopping udev Kernel Device Manager... [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Local File Systems. [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Closed udev Control Socket. [ OK ] Closed udev Kernel Socket. Starting Cleanup udevd DB... [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 18.779113] printk: systemd: 25 output lines suppressed due to ratelimiting [ 19.547190] SELinux: Disabled at runtime. [ 19.671339] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 19.695834] systemd[1]: Detected virtualization kvm. [ 19.698893] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 21.388818] systemd[1]: initrd-switch-root.service: Succeeded. [ 21.400724] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 21.422195] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 21.441167] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 21.447547] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 21.476039] systemd[1]: Starting Journal Service... Starting Journal Service... [ 21.492372] systemd[1]: Listening on udev Kernel Socket. [ OK ] Listening on udev Kernel Socket. [ OK ] Created slice system-getty.slice. [ OK ] Created slice system-sshd\x2dkeygen.slice. [ OK ] Reached target rpc_pipefs.target. Activating swap /dev/disk/by-label/SWAP... [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. [ OK ] Created slice system-serial\x2dgetty.slice. Mounting Huge Pages File System... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ 21.663850] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS [ OK ] Listening on Process Core Dump Socket. [ OK ] Listening on udev Control Socket. Starting udev Coldplug all Devices... [ OK ] Listening on RPCbind Server Activation Socket. [ OK ] Reached target RPC Port Mapper. Starting Remount Root and Kernel File Systems... Starting Apply Kernel Variables... Starting Create list of required st…ce nodes for the current kernel... [ OK ] Listening on initctl Compatibility Named Pipe. [ OK ] Created slice User and Session Slice. [ OK ] Reached target Slices. Mounting Kernel Debug File System... Mounting POSIX Message Queue File System... [ OK ] Stopped target Switch Root. [ OK ] Stopped target Initrd File Systems. [ OK ] Stopped target Initrd Root File System. [ OK ] Started Forward Password Requests to Wall Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Started Journal Service. [ OK ] Activated swap /dev/disk/by-label/SWAP. [ OK ] Mounted Huge Pages File System. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. [ OK ] Started Apply Kernel Variables. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Mounted Kernel Debug File System. [ OK ] Mounted POSIX Message Queue File System. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... [ OK ] Reached target Swap. Starting Flush Journal to Persistent Storage... [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started udev Coldplug all Devices. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Reached target Local File Systems (Pre). Mounting /home/green/git/lustre-release... Mounting /mnt... Starting udev Kernel Device Manager... [ OK ] Mounted /mnt. [ 22.817273] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Kernel Device Manager. [ 23.860467] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 23.886190] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 24.161485] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 24.218098] EDAC sbridge: Ver: 1.1.2 [* ] A start job is running for Configur…-only root support (7s / no limit) [** ] A start job is running for Configur…-only root support (7s / no limit) [*** ] A start job is running for Configur…-only root support (8s / no limit)[ 30.245158] Key type dns_resolver registered [ *** ] A start job is running for Configur…-only root support (9s / no limit) [ *** ] A start job is running for Configur…-only root support (9s / no limit) [ ***] A start job is running for Configur…only root support (10s / no limit)[ 31.392257] NFS: Registering the id_resolver key type [ 31.395080] Key type id_resolver registered [ 31.397361] Key type id_legacy registered [ **] A start job is running for Configur…only root support (10s / no limit) [ OK ] Started Configure read-only root support. [ OK ] Reached target Local File Systems. Starting Mark the need to relabel after reboot... Starting Rebuild Dynamic Linker Cache... Starting Load/Save Random Seed... Starting Create Volatile Files and Directories... [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Load/Save Random Seed. [ OK ] Started Create Volatile Files and Directories. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started dnf makecache --timer. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Started daily update of the root trust anchor for DNSSEC. [ OK ] Reached target Timers. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. [ OK ] Started irqbalance daemon. Starting Restore /run/initramfs on shutdown... [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ OK ] Reached target sshd-keygen.target. Starting Login Service... [ OK ] Started D-Bus System Message Bus. Starting Network Manager... [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Network Manager. [ OK ] Reached target Network. Starting OpenSSH server daemon... Starting GSSAPI Proxy Daemon... Starting Dynamic System Tuning Daemon... Starting Network Manager Wait Online... [ OK ] Started OpenSSH server daemon. [ OK ] Started Login Service. Starting Hostname Service... [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... [ OK ] Started Permit User Sessions. [ OK ] Started Serial Getty on ttyS0. [ OK ] Started Serial Getty on ttyS1. [ OK ] Started Getty on tty1. [ OK ] Reached target Login Prompts. [ OK ] Started Command Scheduler. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting System Logging Service... Starting Crash recovery kernel arming... Starting Notify NFS peers of a restart... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg438-server login: [ 55.809009] hrtimer: interrupt took 2112552 ns [ 95.468051] libcfs: loading out-of-tree module taints kernel. [ 95.595947] Key type ._llcrypt registered [ 95.605094] Key type .llcrypt registered [ 95.756920] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_hostid [ 114.866183] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing load_modules_local [ 116.579539] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 116.594295] alg: No test for adler32 (adler32-zlib) [ 118.107909] Lustre: Lustre: Build Version: 2.17.53_82_geaee21b [ 119.171809] LNet: Added LNI 192.168.204.138@tcp [8/256/0/180] [ 121.031342] Key type lgssc registered [ 123.225816] Lustre: Echo OBD driver; http://www.lustre.org/ [ 139.934237] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 180.208464] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing load_modules_local [ 194.530524] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 194.577634] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 195.860677] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 195.905380] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 195.998530] Lustre: lustre-MDT0000: new disk, initializing [ 196.085541] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 196.108051] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 200.128661] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 213.163324] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 213.281471] Lustre: 6517:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 213.309991] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 213.316142] Lustre: Skipped 1 previous similar message [ 213.464194] Lustre: lustre-MDT0001: new disk, initializing [ 213.572451] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 213.630063] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 213.645255] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 217.657080] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 222.291957] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 233.211904] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 233.502286] Lustre: lustre-OST0000: new disk, initializing [ 233.506210] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 233.511495] Lustre: 8420:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 233.590363] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 239.137384] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 239.154616] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 239.260762] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 240.316354] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 256.236183] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 256.385535] Lustre: lustre-OST0001: new disk, initializing [ 256.400138] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 256.417448] Lustre: 9476:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 256.514214] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 262.230657] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 262.247848] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 262.337735] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 263.739754] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 277.357386] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 288.932887] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 295.294848] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing check_logdir /tmp/testlogs/ [ 300.745303] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing yml_node [ 306.765402] Lustre: DEBUG MARKER: Client: 2.17.53.82 [ 309.186666] Lustre: DEBUG MARKER: MDS: 2.17.53.82 [ 311.721644] Lustre: DEBUG MARKER: OSS: 2.17.53.82 [ 313.416491] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-dual ============----- Thu Jun 11 23:30:22 EDT 2026 [ 330.821971] Lustre: DEBUG MARKER: excepting tests: 14b 21b [ 332.255634] Lustre: DEBUG MARKER: skipping tests SLOW=no: 21b [ 334.092799] Lustre: DEBUG MARKER: === replay-dual: start setup 23:30:43 (1781235043) === [ 338.878517] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing check_config_client /mnt/lustre [ 360.855732] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 365.430478] Lustre: 13238:0:(mgs_llog.c:1437:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 369.689551] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 376.740615] Lustre: DEBUG MARKER: === replay-dual: finish setup 23:31:25 (1781235085) === [ 378.473454] Lustre: DEBUG MARKER: == replay-dual test 0a: expired recovery with lost client ========================================================== 23:31:27 (1781235087) [ 386.852281] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 390.622544] Lustre: Failing over lustre-MDT0000 [ 391.064856] Lustre: server umount lustre-MDT0000 complete [ 392.677867] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 392.690486] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 395.233877] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 395.256476] Lustre: Skipped 1 previous similar message [ 396.313347] LustreError: 10989:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.38@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 396.353384] LustreError: 10989:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 9 previous similar messages [ 400.353959] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 400.376166] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 3 previous similar messages [ 401.427240] LustreError: 6523:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.38@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 405.478268] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 405.501937] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 4 previous similar messages [ 410.463135] Lustre: 3669:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235106/real 1781235106] req@ffff9394847503c0 x1867760093229184/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235122 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 410.492925] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 410.525176] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 410.550375] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 5 previous similar messages [ 413.748421] LDISKFS-fs (dm-0): 10 truncates cleaned up [ 413.750659] LDISKFS-fs (dm-0): recovery complete [ 413.759783] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 420.843081] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 420.872271] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 11 previous similar messages [ 421.149311] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 422.331183] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 425.572767] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 426.494649] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 528.500275] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 528.503456] Lustre: 14763:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 15f5e7c7-c0fa-4613-b44a-e695ca628f0b@192.168.204.38@tcp [ 528.518458] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 528.542441] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 528.543653] Lustre: 14763:0:(ldlm_lib.c:2932:target_recovery_thread()) too long recovery - read logs [ 528.548029] Lustre: Skipped 2 previous similar messages [ 528.563765] LustreError: dumping log to /tmp/lustre-log.1781235240.14763 [ 528.659778] Lustre: lustre-MDT0000: Recovery over after 1:46, of 3 clients 2 recovered and 1 was evicted. [ 528.701054] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:65) [ 528.714664] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:65) [ 551.910379] Lustre: DEBUG MARKER: == replay-dual test 0b: lost client during waiting for next transno ========================================================== 23:34:20 (1781235260) [ 560.224347] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 562.725473] Lustre: Failing over lustre-MDT0000 [ 563.069707] Lustre: server umount lustre-MDT0000 complete [ 564.718702] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 564.726966] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 564.730985] Lustre: Skipped 4 previous similar messages [ 564.735920] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 564.744770] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 3 previous similar messages [ 581.087170] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235276/real 1781235276] req@ffff9394854b0000 x1867760093308928/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235292 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 581.115329] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 585.996098] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 585.998363] LDISKFS-fs (dm-0): recovery complete [ 586.006364] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 592.351145] LustreError: 16454:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 592.362262] LustreError: 16454:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9394854b1a40 x1867760093314944/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1781235302 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 592.398771] LustreError: 16454:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 592.807800] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 592.926349] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 597.291826] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 598.004458] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 610.514952] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:51 [ 615.956516] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:46 [ 621.075835] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:41 [ 626.199224] Lustre: lustre-MDT0001: haven't heard from client 15f5e7c7-c0fa-4613-b44a-e695ca628f0b (at 192.168.204.38@tcp) in 101 seconds. I think it's dead, and I am evicting it. exp ffff939483179000, cur 1781235337 deadline 1781235336 last 1781235236 [ 626.208457] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:36 [ 631.333097] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:31 [ 641.554978] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:20 [ 641.570384] Lustre: Skipped 1 previous similar message [ 662.042479] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:00 [ 662.072955] Lustre: Skipped 3 previous similar messages [ 662.500298] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 662.509955] Lustre: 16489:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 055dfeb1-8117-49cd-9108-a54927885f0e@ [ 662.542068] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 697.877754] Lustre: lustre-MDT0000: Denying connection for new client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 1:05 [ 697.908121] Lustre: Skipped 6 previous similar messages [ 700.395387] Lustre: lustre-MDT0001: haven't heard from client c782bdc5-9758-4017-ae83-200ec846473a (at 192.168.204.38@tcp) in 102 seconds. I think it's dead, and I am evicting it. exp ffff9394826a6800, cur 1781235411 deadline 1781235409 last 1781235309 [ 763.500921] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 763.511081] Lustre: 16489:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client c782bdc5-9758-4017-ae83-200ec846473a@192.168.204.38@tcp [ 763.524056] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 763.532602] Lustre: 16489:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 763.559094] Lustre: 16489:0:(ldlm_lib.c:2932:target_recovery_thread()) too long recovery - read logs [ 763.559547] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 763.575659] LustreError: dumping log to /tmp/lustre-log.1781235475.16489 [ 763.577061] Lustre: Skipped 2 previous similar messages [ 763.758712] Lustre: lustre-MDT0000: Recovery over after 2:51, of 3 clients 1 recovered and 2 were evicted. [ 763.805230] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:97) [ 763.808837] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:97) [ 773.872847] Lustre: DEBUG MARKER: == replay-dual test 1: |X| simple create ================= 23:38:03 (1781235483) [ 783.225657] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 785.288879] Lustre: Failing over lustre-MDT0000 [ 785.655331] Lustre: server umount lustre-MDT0000 complete [ 785.979517] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.38@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 785.991652] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 34 previous similar messages [ 787.431506] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 787.442180] Lustre: Skipped 1 previous similar message [ 802.784187] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235498/real 1781235498] req@ffff9395b3c2de00 x1867760093408256/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235514 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 802.833951] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 809.685979] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 809.690129] LDISKFS-fs (dm-0): recovery complete [ 809.702265] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 813.035780] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395916f52c0 x1867760093416704/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 813.415060] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 813.478210] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 814.660691] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 817.639412] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 818.688031] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 818.895371] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 818.957611] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:129) [ 818.957941] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:129) [ 825.948936] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 827.692732] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 835.931341] Lustre: DEBUG MARKER: == replay-dual test 2: |X| mkdir adir ==================== 23:39:05 (1781235545) [ 843.997601] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 846.363060] Lustre: Failing over lustre-MDT0000 [ 846.793610] Lustre: server umount lustre-MDT0000 complete [ 849.380515] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 849.394264] Lustre: Skipped 3 previous similar messages [ 850.514069] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.38@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 850.545126] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 41 previous similar messages [ 865.759240] Lustre: 3667:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235560/real 1781235560] req@ffff9395b59f2d00 x1867760093442816/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235576 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 865.795625] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 872.838609] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 872.840935] LDISKFS-fs (dm-0): recovery complete [ 872.858194] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 876.000611] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395b58c5680 x1867760093451136/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 876.537057] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 876.605367] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 877.090300] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 881.698314] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 881.709500] Lustre: Skipped 3 previous similar messages [ 881.819108] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 881.891984] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:161) [ 881.896541] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:161) [ 882.283384] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 891.589401] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 893.487645] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 902.497695] Lustre: DEBUG MARKER: == replay-dual test 3: |X| mkdir adir, mkdir adir/bdir === 23:40:11 (1781235611) [ 911.325049] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 913.722377] Lustre: Failing over lustre-MDT0000 [ 913.949486] Lustre: server umount lustre-MDT0000 complete [ 915.426973] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 915.440719] Lustre: Skipped 5 previous similar messages [ 930.789569] Lustre: 3669:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235626/real 1781235626] req@ffff939599ae3c00 x1867760093487360/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235642 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 930.831450] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 937.226499] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 937.230735] LDISKFS-fs (dm-0): recovery complete [ 937.241957] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 941.473938] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 941.543749] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 943.643577] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 946.524132] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 946.681041] Lustre: lustre-MDT0000-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 946.695136] Lustre: Skipped 3 previous similar messages [ 946.809134] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 946.889848] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:193) [ 946.890675] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:193) [ 955.090906] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 956.968696] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 966.651962] Lustre: DEBUG MARKER: == replay-dual test 4: |X| mkdir adir (-EEXIST), mkdir adir/bdir ========================================================== 23:41:15 (1781235675) [ 974.726590] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 977.165894] Lustre: Failing over lustre-MDT0000 [ 977.378558] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 977.393297] Lustre: Skipped 1 previous similar message [ 977.460685] Lustre: server umount lustre-MDT0000 complete [ 979.521350] LustreError: 15136:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.38@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 979.550254] LustreError: 15136:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 77 previous similar messages [ 998.882937] Lustre: 3668:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235694/real 1781235694] req@ffff93959d3612c0 x1867760093530240/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235710 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 998.916971] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1000.151183] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 1000.155380] LDISKFS-fs (dm-0): recovery complete [ 1000.164949] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1010.544910] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1011.239631] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1015.066397] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1015.831336] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1015.856213] Lustre: Skipped 3 previous similar messages [ 1015.929449] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1015.955227] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:225) [ 1015.955258] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:225) [ 1022.600548] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1024.751545] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1035.259730] Lustre: DEBUG MARKER: == replay-dual test 5: open, unlink |X| close ============ 23:42:24 (1781235744) [ 1044.694906] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1046.911451] Lustre: Failing over lustre-MDT0000 [ 1047.008822] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1047.023155] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1047.033250] Lustre: Skipped 2 previous similar messages [ 1047.037958] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 1047.267330] Lustre: server umount lustre-MDT0000 complete [ 1068.000211] Lustre: 3669:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235763/real 1781235763] req@ffff9395a73a4f00 x1867760093570560/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235779 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1068.023128] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1070.885083] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1070.888427] LDISKFS-fs (dm-0): recovery complete [ 1070.902212] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1078.253638] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c8277d6 [ 1078.582217] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1078.588938] Lustre: Skipped 1 previous similar message [ 1078.647769] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1079.351304] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1082.971615] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1083.899166] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1083.903235] Lustre: Skipped 4 previous similar messages [ 1084.080707] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1084.152137] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:257) [ 1084.156955] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:257) [ 1091.845734] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1093.514586] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1102.954764] Lustre: DEBUG MARKER: == replay-dual test 6: open1, open2, unlink |X| close1 [fail mds1] close2 ========================================================== 23:43:31 (1781235811) [ 1111.759661] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1114.096409] Lustre: Failing over lustre-MDT0000 [ 1114.397768] Lustre: server umount lustre-MDT0000 complete [ 1114.601209] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1114.621977] Lustre: Skipped 5 previous similar messages [ 1136.098162] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235831/real 1781235831] req@ffff9395c1cf8b40 x1867760093607424/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235847 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1136.124977] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1137.093638] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1137.095955] LDISKFS-fs (dm-0): recovery complete [ 1137.102725] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1146.881205] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1147.416811] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1151.423420] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1152.079202] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1152.126789] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:289) [ 1152.126930] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:289) [ 1160.000918] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1161.768767] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1170.999293] Lustre: DEBUG MARKER: == replay-dual test 8: replay of resent request ========== 23:44:39 (1781235879) [ 1178.364868] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1179.437339] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1179.445573] LustreError: 15136:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395b5b3f480 x1867760072412288/t38654705670(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:641/0 lens 512/448 e 0 to 0 dl 1781235901 ref 1 fl Interpret:/200/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1196.099871] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnecting [ 1196.138823] Lustre: 10989:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff939483017840 x1867760072412288/t38654705670(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:658/0 lens 512/2880 e 0 to 0 dl 1781235918 ref 1 fl Interpret:/202/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1199.242079] Lustre: Failing over lustre-MDT0000 [ 1199.518143] Lustre: server umount lustre-MDT0000 complete [ 1219.554833] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781235914/real 1781235914] req@ffff939599ae3840 x1867760093652224/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781235930 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1219.580825] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1222.537177] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1222.539014] LDISKFS-fs (dm-0): recovery complete [ 1222.549373] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1229.793018] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395b3c712c0 x1867760093660800/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1230.282920] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1231.110242] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1235.457352] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1235.465748] Lustre: Skipped 7 previous similar messages [ 1235.601073] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1235.678212] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:321) [ 1235.680995] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:321) [ 1236.468542] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1245.844229] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1247.603673] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1255.654754] Lustre: DEBUG MARKER: == replay-dual test 9: resending a replayed create ======= 23:46:04 (1781235964) [ 1262.469514] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1264.735790] Lustre: Failing over lustre-MDT0000 [ 1264.977907] Lustre: server umount lustre-MDT0000 complete [ 1266.150518] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1266.157995] LustreError: 15136:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1266.169792] Lustre: Skipped 8 previous similar messages [ 1266.180478] LustreError: 15136:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 172 previous similar messages [ 1287.473218] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1287.475809] LDISKFS-fs (dm-0): recovery complete [ 1287.484519] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1292.339931] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1296.778689] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1297.406928] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1297.413394] LustreError: 31175:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395a739fc00 x1867760072430208/t42949672962(42949672962) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:0/0 lens 528/448 e 0 to 0 dl 1781236015 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1308.710437] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1308.811733] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:353) [ 1308.813417] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:353) [ 1313.603904] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1315.218922] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1324.006440] Lustre: DEBUG MARKER: == replay-dual test 10: resending a replayed unlink ====== 23:47:13 (1781236033) [ 1331.664746] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1334.702986] Lustre: Failing over lustre-MDT0000 [ 1334.757326] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1334.772133] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 1334.786283] Lustre: Skipped 2 previous similar messages [ 1335.012502] Lustre: server umount lustre-MDT0000 complete [ 1351.650817] Lustre: 3669:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781236047/real 1781236047] req@ffff93959d359680 x1867760093727488/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781236063 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1351.680338] Lustre: 3669:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1351.691651] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1351.709822] LustreError: Skipped 1 previous similar message [ 1358.492700] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1358.499093] LDISKFS-fs (dm-0): recovery complete [ 1358.521668] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1362.655137] LustreError: 33083:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 1362.671927] LustreError: 33083:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9395a72c7c00 x1867760093733248/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1781236073 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1362.714012] LustreError: 33083:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 1362.918612] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939484e40b40 x1867760093735936/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1363.265219] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1363.269355] Lustre: Skipped 3 previous similar messages [ 1363.315854] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1366.043924] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1366.050665] Lustre: Skipped 1 previous similar message [ 1367.424811] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1368.593977] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1368.598711] LustreError: 33118:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff93959d29b480 x1867760072448768/t47244640260(47244640260) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:73/0 lens 528/448 e 0 to 0 dl 1781236088 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1381.442189] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1381.529559] Lustre: lustre-MDT0000: Recovery over after 0:15, of 3 clients 3 recovered and 0 were evicted. [ 1381.539343] Lustre: Skipped 1 previous similar message [ 1381.579237] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:385) [ 1381.582358] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:385) [ 1385.723862] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1387.306224] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1397.031194] Lustre: DEBUG MARKER: == replay-dual test 11: both clients timeout during replay ========================================================== 23:48:26 (1781236106) [ 1403.947228] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1406.719159] Lustre: Failing over lustre-MDT0000 [ 1407.015478] Lustre: server umount lustre-MDT0000 complete [ 1407.458331] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1429.178534] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1429.183701] LDISKFS-fs (dm-0): recovery complete [ 1429.195256] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1436.129048] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939483125e00 x1867760093776128/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1441.205704] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1441.845571] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1441.853249] LustreError: 35061:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff939483126940 x1867760072467328/t51539607554(51539607554) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:145/0 lens 528/448 e 0 to 0 dl 1781236160 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1447.890870] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1454.147190] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1454.327317] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:417) [ 1454.334221] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:417) [ 1456.240879] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 6 sec [ 1463.979035] Lustre: DEBUG MARKER: == replay-dual test 12: open resend timeout ============== 23:49:33 (1781236173) [ 1471.702455] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1475.203330] Lustre: Failing over lustre-MDT0000 [ 1475.741541] Lustre: server umount lustre-MDT0000 complete [ 1480.160863] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1501.214987] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1501.217546] LDISKFS-fs (dm-0): recovery complete [ 1501.232383] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1507.898741] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1507.916837] Lustre: Skipped 1 previous similar message [ 1512.186483] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1512.947883] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1512.962311] Lustre: Skipped 15 previous similar messages [ 1513.042189] Lustre: *** cfs_fail_loc=302, val=2147483648*** [ 1529.396075] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnected, waiting for 3 clients in recovery for 1:24 [ 1529.533913] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:449) [ 1529.551427] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:449) [ 1536.261580] Lustre: DEBUG MARKER: == replay-dual test 13: close resend timeout ============= 23:50:45 (1781236245) [ 1543.331195] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1545.926901] Lustre: Failing over lustre-MDT0000 [ 1546.182737] Lustre: server umount lustre-MDT0000 complete [ 1548.774077] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1548.788095] Lustre: Skipped 14 previous similar messages [ 1567.748378] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1567.751136] LDISKFS-fs (dm-0): recovery complete [ 1567.756645] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1579.599230] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1581.152844] Lustre: *** cfs_fail_loc=115, val=2147483648*** [ 1596.487251] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnected, waiting for 3 clients in recovery for 1:25 [ 1596.630407] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:481) [ 1596.631138] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:481) [ 1603.441138] Lustre: DEBUG MARKER: SKIP: replay-dual test_14b skipping ALWAYS excluded test 14b [ 1604.955285] Lustre: DEBUG MARKER: == replay-dual test 15a: timeout waiting for lost client during replay, 1 client completes ========================================================== 23:51:54 (1781236314) [ 1611.429934] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1614.169424] Lustre: Failing over lustre-MDT0000 [ 1614.457368] Lustre: server umount lustre-MDT0000 complete [ 1633.247383] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781236328/real 1781236328] req@ffff93959d3b9680 x1867760093876992/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781236344 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1633.276384] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 1633.289487] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1633.302141] LustreError: Skipped 3 previous similar messages [ 1636.214430] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1636.217104] LDISKFS-fs (dm-0): recovery complete [ 1636.223930] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1642.466064] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff93959d2f43c0 x1867760093885568/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1644.426350] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1644.435445] Lustre: Skipped 3 previous similar messages [ 1648.104272] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1714.500235] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1714.509476] Lustre: 40584:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 538a27c0-bd46-4012-b5e6-4fa9330a3a43@ [ 1714.532186] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1715.486528] Lustre: lustre-MDT0000: Recovery over after 1:11, of 3 clients 2 recovered and 1 was evicted. [ 1715.503693] Lustre: Skipped 3 previous similar messages [ 1715.562259] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:495 to 0x280000401:513) [ 1715.563554] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:494 to 0x2c0000401:513) [ 1720.453029] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1722.231662] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1733.256816] Lustre: DEBUG MARKER: == replay-dual test 15c: remove multiple OST orphans ===== 23:54:02 (1781236442) [ 1741.874744] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1856.555399] Lustre: Failing over lustre-MDT0000 [ 1856.997217] Lustre: server umount lustre-MDT0000 complete [ 1858.045914] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1858.078246] LustreError: 6524:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 243 previous similar messages [ 1882.256751] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1882.259038] LDISKFS-fs (dm-0): recovery complete [ 1882.278140] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1883.487352] LustreError: 42453:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 1883.506682] LustreError: 42453:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff939599a62940 x1867760094003968/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1781236595 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1883.557574] LustreError: 42453:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 1883.682996] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395918ec3c0 x1867760094006656/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1884.016666] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1884.020769] Lustre: Skipped 4 previous similar messages [ 1884.098863] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1884.109132] Lustre: Skipped 2 previous similar messages [ 1888.404237] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 1954.506827] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1954.511297] Lustre: 42487:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client f185b744-b960-49fc-8754-a6f529e172b8@ [ 1954.520674] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1954.641815] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:495 to 0x280000401:1537) [ 1954.641888] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:494 to 0x2c0000401:1537) [ 1959.448879] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1960.801365] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1970.168556] Lustre: DEBUG MARKER: == replay-dual test 16: fail MDS during recovery (3571) == 23:57:59 (1781236679) [ 1978.509899] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1981.806657] Lustre: Failing over lustre-MDT0000 [ 1982.074905] Lustre: server umount lustre-MDT0000 complete [ 1985.507381] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2004.314309] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2004.325618] LDISKFS-fs (dm-0): recovery complete [ 2004.336225] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2013.168182] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c84b24f [ 2018.210384] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2042.939940] Lustre: Failing over lustre-MDT0000 [ 2042.953956] LustreError: 44811:0:(ldlm_lib.c:2985:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 2042.966426] Lustre: 44350:0:(ldlm_lib.c:2388:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2042.977116] Lustre: 44350:0:(ldlm_lib.c:1898:abort_req_replay_queue()) @@@ aborted: req@ffff9395bb0c8000 x1867760074929792/t0(73014444033) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:746/0 lens 528/0 e 2 to 0 dl 1781236761 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 2042.997302] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 2043.029406] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.38@tcp (stopping) [ 2043.044408] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 2043.057545] LustreError: 44350:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff93948bd7de00 x1867760094086656/t0(0) o700->lustre-MDT0001-osp-MDT0000@0@lo:30/10 lens 264/248 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'tgt_recover_0.0' uid:0 gid:0 projid:4294967295 [ 2043.076606] LustreError: 44350:0:(fid_request.c:213:seq_client_alloc_seq()) cli-cli-lustre-MDT0001-osp-MDT0000: Cannot allocate new meta-sequence: rc = -5 [ 2043.085066] LustreError: 44350:0:(fid_request.c:316:seq_client_alloc_fid()) cli-cli-lustre-MDT0001-osp-MDT0000: Can't allocate new sequence: rc = -5 [ 2043.557192] Lustre: server umount lustre-MDT0000 complete [ 2062.598489] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2074.933119] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2075.654982] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2075.664258] Lustre: Skipped 19 previous similar messages [ 2141.505945] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2141.519034] Lustre: 45268:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 884341bd-44bf-4785-8158-d41b0549d749@ [ 2141.538916] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2142.303425] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1551 to 0x280000401:1569) [ 2142.303432] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1550 to 0x2c0000401:1569) [ 2146.407646] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2148.097592] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2159.873676] Lustre: DEBUG MARKER: == replay-dual test 17: fail OST during recovery (3571) == 00:01:08 (1781236868) [ 2170.099145] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 2172.475544] Lustre: Failing over lustre-OST0000 [ 2172.549387] Lustre: server umount lustre-OST0000 complete [ 2172.909131] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2172.932329] Lustre: Skipped 17 previous similar messages [ 2196.719490] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 2196.721622] LDISKFS-fs (dm-2): recovery complete [ 2196.731336] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2198.182256] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 2198.196594] Lustre: Skipped 3 previous similar messages [ 2203.447832] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2228.828215] Lustre: Failing over lustre-OST0000 [ 2228.835211] LustreError: 47671:0:(ldlm_lib.c:2985:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 2228.843865] Lustre: 47122:0:(ldlm_lib.c:2388:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2228.850888] Lustre: 47122:0:(ldlm_lib.c:2388:target_recovery_overseer()) Skipped 2 previous similar messages [ 2228.864048] LustreError: 47122:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 2228.874318] Lustre: lustre-OST0000: Recovery over after 0:30, of 4 clients 0 recovered and 4 were evicted. [ 2228.885352] Lustre: Skipped 3 previous similar messages [ 2229.032604] Lustre: server umount lustre-OST0000 complete [ 2248.487425] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2250.912408] Lustre: 3665:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781236909/real 1781236909] req@ffff9395b3f76940 x1867760094161664/t0(0) o400->lustre-OST0000-osc-MDT0000@0@lo:28/4 lens 224/224 e 3 to 1 dl 1781236962 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 2250.954882] Lustre: 3665:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2256.088148] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2320.503798] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 2320.510943] Lustre: 48109:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client c50a8049-3535-4c88-ad26-8d71f9a2e45e@ [ 2320.521852] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 2324.905860] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 2327.355957] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 2338.153562] Lustre: DEBUG MARKER: == replay-dual test 18: ldlm_handle_enqueue succeeds on evicted export (3822) ========================================================== 00:04:07 (1781237047) [ 2342.726047] LustreError: 6525:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b sleeping for 40000ms [ 2382.823434] LustreError: 6525:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b awake [ 2398.100880] Lustre: DEBUG MARKER: == replay-dual test 19: resend of open request =========== 00:05:06 (1781237106) [ 2408.130748] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2410.028669] Lustre: *** cfs_fail_loc=157, val=2147483648*** [ 2410.033796] LustreError: 8427:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395bf4ef480 x1867760075056896/t0(0) o101->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:432/0 lens 576/688 e 0 to 0 dl 1781237202 ref 1 fl Interpret:/600/0 rc 0/0 job:'createmany.0' uid:0 gid:0 projid:0 [ 2496.027309] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnecting [ 2499.958561] Lustre: Failing over lustre-MDT0000 [ 2500.334878] Lustre: server umount lustre-MDT0000 complete [ 2500.584028] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2500.618719] LustreError: 7866:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2500.668050] LustreError: 7866:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 110 previous similar messages [ 2516.464156] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2516.477153] LustreError: Skipped 3 previous similar messages [ 2524.551448] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2524.557645] LDISKFS-fs (dm-0): recovery complete [ 2524.567094] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2526.815367] LustreError: 50617:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 2526.822844] LustreError: 50617:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff93948b214b40 x1867760094299008/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1781237238 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2526.854368] LustreError: 50617:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 2527.722454] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395c1c970c0 x1867760094301696/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2528.152346] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2528.156096] Lustre: Skipped 4 previous similar messages [ 2528.220324] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2528.224606] Lustre: Skipped 4 previous similar messages [ 2533.399705] Lustre: 50650:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2533.686558] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1601) [ 2533.688108] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1601) [ 2534.149896] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2544.175592] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2546.528601] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2555.434378] Lustre: DEBUG MARKER: == replay-dual test 20: recovery time is not increasing == 00:07:44 (1781237264) [ 2563.622779] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2566.122804] Lustre: Failing over lustre-MDT0000 [ 2566.377496] Lustre: server umount lustre-MDT0000 complete [ 2589.997536] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2590.001253] LDISKFS-fs (dm-0): recovery complete [ 2590.010535] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2595.811730] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939481584000 x1867760094341504/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2602.087090] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2737.501550] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2737.504386] Lustre: 52484:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client c99b0c7e-ba31-4d6e-b078-80452fc5f68a@ [ 2737.529649] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2737.575804] Lustre: 52484:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2737.588750] Lustre: 52484:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 6 previous similar messages [ 2737.682695] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2737.693209] Lustre: Skipped 12 previous similar messages [ 2737.727569] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1633) [ 2737.727694] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1603 to 0x2c0000401:1633) [ 2742.617686] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2744.651867] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2755.983402] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2758.819436] Lustre: Failing over lustre-MDT0000 [ 2759.277136] Lustre: server umount lustre-MDT0000 complete [ 2782.782366] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2782.792814] LDISKFS-fs (dm-0): recovery complete [ 2782.820383] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2785.794424] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c84dc5d [ 2791.833806] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 2929.500892] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2929.503483] Lustre: 54153:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 4224b1f5-e5af-459c-a306-cc1144184aad@ [ 2929.531971] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2929.580631] Lustre: 54153:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2929.603112] Lustre: 54153:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 2929.732426] Lustre: lustre-MDT0000: Recovery over after 2:20, of 3 clients 2 recovered and 1 was evicted. [ 2929.745788] Lustre: Skipped 3 previous similar messages [ 2929.801353] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1665) [ 2929.806022] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1665) [ 2935.841438] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2938.009246] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2949.816203] Lustre: DEBUG MARKER: == replay-dual test 21a: commit on sharing =============== 00:14:18 (1781237658) [ 2960.413401] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2962.669994] Lustre: Failing over lustre-MDT0000 [ 2963.033391] Lustre: server umount lustre-MDT0000 complete [ 2963.426091] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2963.450449] Lustre: Skipped 14 previous similar messages [ 2979.810387] Lustre: 3668:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781237674/real 1781237674] req@ffff93948bc20f00 x1867760094505344/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781237690 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2979.851785] Lustre: 3668:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2988.816130] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2988.818925] LDISKFS-fs (dm-0): recovery complete [ 2988.835583] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2990.054851] LustreError: 56039:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 2990.074198] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939484654780 x1867760094514944/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2993.181594] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 2993.203456] Lustre: Skipped 4 previous similar messages [ 2996.078428] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3133.501478] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3133.510167] Lustre: 56072:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 3689d3c9-2e41-4330-b8a7-799c0372c94a@ [ 3133.521706] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3133.558422] Lustre: 56072:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3133.570542] Lustre: 56072:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 3133.641809] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1697) [ 3133.642622] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1697) [ 3144.728669] Lustre: DEBUG MARKER: SKIP: replay-dual test_21b skipping SLOW test 21b [ 3146.855565] Lustre: DEBUG MARKER: == replay-dual test 22a: c1 lfs mkdir -i 1 dir1, M1 drop reply [ 3148.527791] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3148.534093] LustreError: 10989:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395a73b6940 x1867760075180032/t4294967346(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:414/0 lens 560/448 e 0 to 0 dl 1781237939 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3151.093132] Lustre: Failing over lustre-MDT0001 [ 3151.566317] Lustre: server umount lustre-MDT0001 complete [ 3152.359802] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3152.388785] LustreError: 6528:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 142 previous similar messages [ 3171.627139] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3172.087173] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3172.093357] Lustre: Skipped 3 previous similar messages [ 3172.168071] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3172.175164] Lustre: Skipped 3 previous similar messages [ 3177.212135] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3177.629803] Lustre: 15136:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9395916eed00 x1867760075180032/t4294967346(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:444/0 lens 560/2880 e 0 to 0 dl 1781237969 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3185.978656] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3187.671810] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3197.689739] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3200.081962] Lustre: Failing over lustre-MDT0000 [ 3200.618773] Lustre: server umount lustre-MDT0000 complete [ 3203.046973] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3219.429548] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3219.445021] LustreError: Skipped 3 previous similar messages [ 3223.369621] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3223.374470] LDISKFS-fs (dm-0): recovery complete [ 3223.401563] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3229.679113] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c84eafe [ 3235.347354] Lustre: 59008:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3235.486675] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1729) [ 3235.491880] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1729) [ 3236.062386] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3245.406610] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3247.094417] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3258.402864] Lustre: DEBUG MARKER: == replay-dual test 22b: c1 lfs mkdir -i 1 d1, M1 drop reply [ 3260.131237] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3260.143604] LustreError: 6523:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395bb0c92c0 x1867760075220992/t8589934617(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:526/0 lens 560/448 e 0 to 0 dl 1781238051 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3263.745355] Lustre: Failing over lustre-MDT0000 [ 3264.204135] Lustre: server umount lustre-MDT0000 complete [ 3268.757159] Lustre: Failing over lustre-MDT0001 [ 3268.761080] LustreError: 6509:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781237980 with bad export cookie 13706422200134068990 [ 3268.780061] LustreError: 6509:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 3269.087217] Lustre: server umount lustre-MDT0001 complete [ 3288.880893] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3288.988451] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3289.415558] LustreError: 60736:0:(llog.c:1646:llog_backup()) MGC192.168.204.138@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3289.435181] Lustre: 60736:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.138@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3293.678538] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c84f275 [ 3300.167643] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3300.186198] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3300.607851] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:36 to 0x280000400:65) [ 3300.611647] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:36 to 0x2c0000400:65) [ 3300.650779] Lustre: 60754:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9395b46d1680 x1867760075220992/t8589934617(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:567/0 lens 560/2880 e 0 to 0 dl 1781238092 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3306.787973] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1761) [ 3306.792855] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1761) [ 3311.729656] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3313.681653] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3315.690581] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3326.076925] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3329.226641] Lustre: Failing over lustre-MDT0000 [ 3329.745038] Lustre: server umount lustre-MDT0000 complete [ 3352.576624] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3352.580082] LDISKFS-fs (dm-0): recovery complete [ 3352.585639] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3358.175643] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff93948bee4780 x1867760094707456/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3363.352610] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3363.835702] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 3363.852675] Lustre: Skipped 24 previous similar messages [ 3363.876338] Lustre: 62837:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3363.884932] Lustre: 62837:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 3364.025798] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1793) [ 3364.027804] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1793) [ 3371.884708] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3373.531902] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3382.525504] Lustre: DEBUG MARKER: == replay-dual test 22c: c1 lfs mkdir -i 1 d1, M1 drop update [ 3383.753636] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3383.756761] LustreError: 8414:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff93948a3f12c0 x1867760094733952/t107374182411(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:581/0 lens 2520/4320 e 0 to 0 dl 1781238106 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3387.653057] Lustre: Failing over lustre-MDT0000 [ 3388.047730] Lustre: server umount lustre-MDT0000 complete [ 3407.603410] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3415.040862] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c85027b [ 3420.733330] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1825) [ 3420.736285] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1825) [ 3420.762500] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3428.925771] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3430.788787] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3441.881257] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3444.409727] Lustre: Failing over lustre-MDT0000 [ 3444.702580] Lustre: server umount lustre-MDT0000 complete [ 3446.246430] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3467.912423] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3467.917966] LDISKFS-fs (dm-0): recovery complete [ 3467.925894] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3474.913577] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939591616d00 x1867760094779776/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3480.321782] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3480.602103] Lustre: 65805:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3480.614646] Lustre: 65805:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 3480.783322] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1857) [ 3480.797845] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1857) [ 3488.769910] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3490.960686] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3500.752801] Lustre: DEBUG MARKER: == replay-dual test 22d: c1 lfs mkdir -i 1 d1, M1 drop update [ 3506.365847] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3506.384679] LustreError: 61310:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff93959d3dcb40 x1867760094810624/t115964117002(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:703/0 lens 2520/4320 e 0 to 0 dl 1781238228 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3510.256100] Lustre: Failing over lustre-MDT0000 [ 3510.732353] Lustre: server umount lustre-MDT0000 complete [ 3515.325429] LustreError: 6508:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781238226 with bad export cookie 13706422200134076711 [ 3515.327405] Lustre: Failing over lustre-MDT0001 [ 3515.339920] LustreError: 6508:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 3515.387627] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.38@tcp (stopping) [ 3516.410903] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 3517.986128] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.38@tcp (stopping) [ 3517.991742] Lustre: Skipped 1 previous similar message [ 3520.418503] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 3520.426194] Lustre: Skipped 2 previous similar messages [ 3521.248662] Lustre: server umount lustre-MDT0001 complete [ 3541.204317] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3541.477952] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3560.416941] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff9395c1e2a940 x1867760094817792/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3566.791026] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3567.393137] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3568.383426] Lustre: lustre-MDT0000: Recovery over after 0:02, of 3 clients 3 recovered and 0 were evicted. [ 3568.395334] Lustre: Skipped 8 previous similar messages [ 3568.434522] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1889) [ 3568.435750] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1889) [ 3582.836092] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 3582.840068] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 3582.843230] Lustre: 67655:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9395a653a1c0 x1867760075310720/t12884901939(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:94/0 lens 560/2880 e 0 to 0 dl 1781238374 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3587.521503] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3589.431564] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3591.210547] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3602.043400] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3604.578919] Lustre: Failing over lustre-MDT0000 [ 3604.851571] Lustre: server umount lustre-MDT0000 complete [ 3605.486526] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3605.500805] Lustre: Skipped 34 previous similar messages [ 3621.926471] Lustre: 3667:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781238317/real 1781238317] req@ffff93958eb25e00 x1867760094853632/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781238333 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3621.981898] Lustre: 3667:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 17 previous similar messages [ 3629.518663] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3629.524553] LDISKFS-fs (dm-0): recovery complete [ 3629.538975] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3631.775296] LustreError: 69716:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 3631.789337] LustreError: 69716:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff9395c1e2b0c0 x1867760094859520/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1781238343 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3631.829043] LustreError: 69716:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 3632.682802] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c85192d [ 3635.234457] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3635.250533] Lustre: Skipped 9 previous similar messages [ 3637.136503] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3638.297419] Lustre: 69751:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3638.312532] Lustre: 69751:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 3638.461067] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1635 to 0x2c0000401:1921) [ 3638.461609] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1667 to 0x280000401:1921) [ 3645.220263] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3646.876939] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3655.247178] Lustre: DEBUG MARKER: == replay-dual test 23a: c1 rmdir d1, M1 drop reply and fail, client2 mkdir d1 ========================================================== 00:26:04 (1781238364) [ 3657.018342] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3657.020694] LustreError: 67655:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395a653b840 x1867760075356672/t17179869210(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:168/0 lens 496/456 e 0 to 0 dl 1781238448 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3660.982644] Lustre: Failing over lustre-MDT0001 [ 3661.491112] Lustre: server umount lustre-MDT0001 complete [ 3663.851783] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3680.399584] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3685.476683] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3686.487070] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:129) [ 3686.488914] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:129) [ 3686.534339] Lustre: 67656:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9394850fc3c0 x1867760075356672/t17179869210(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:198/0 lens 496/2888 e 0 to 0 dl 1781238478 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3692.555196] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3694.224747] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3703.438678] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3705.621181] Lustre: Failing over lustre-MDT0000 [ 3706.056231] Lustre: server umount lustre-MDT0000 complete [ 3728.336717] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3728.339518] LDISKFS-fs (dm-0): recovery complete [ 3728.346567] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3732.977592] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c852543 [ 3737.405791] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3738.643331] Lustre: 72701:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3738.656774] Lustre: 72701:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 4 previous similar messages [ 3738.802070] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1953) [ 3738.807735] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1953) [ 3745.119253] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3747.070688] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3756.197258] Lustre: DEBUG MARKER: == replay-dual test 23b: c1 rmdir d1, M1 drop reply and fail M0/M1, c2 mkdir d1 ========================================================== 00:27:45 (1781238465) [ 3757.902300] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3757.906331] LustreError: 67655:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff9395983f5a40 x1867760075390976/t21474836483(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:269/0 lens 496/456 e 0 to 0 dl 1781238549 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3761.604690] Lustre: Failing over lustre-MDT0000 [ 3762.053616] Lustre: server umount lustre-MDT0000 complete [ 3762.160446] LustreError: 6509:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781238473 with bad export cookie 13706422200134084019 [ 3762.179456] LustreError: 61310:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3762.207091] LustreError: 61310:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 307 previous similar messages [ 3765.466820] LustreError: 6510:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781238477 with bad export cookie 13706422200134083907 [ 3765.477095] Lustre: Failing over lustre-MDT0001 [ 3765.482894] LustreError: 6510:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 1 previous similar message [ 3765.881918] Lustre: server umount lustre-MDT0001 complete [ 3786.955490] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3787.461703] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3787.494635] LustreError: 74460:0:(llog.c:1646:llog_backup()) MGC192.168.204.138@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3787.521924] Lustre: 74460:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.138@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3791.242520] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3791.247433] Lustre: Skipped 11 previous similar messages [ 3791.288598] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3791.294101] Lustre: Skipped 11 previous similar messages [ 3796.993018] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3797.339429] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3797.703204] Lustre: 74478:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff93958a69f480 x1867760075390976/t21474836483(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:309/0 lens 496/2888 e 0 to 0 dl 1781238589 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3797.716396] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:161) [ 3797.741529] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:161) [ 3805.673437] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1985) [ 3805.674575] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1985) [ 3810.067393] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3811.537955] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3813.079281] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3823.059734] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3825.364543] Lustre: Failing over lustre-MDT0000 [ 3825.776848] Lustre: server umount lustre-MDT0000 complete [ 3826.153737] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3846.111448] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3846.123702] LustreError: Skipped 8 previous similar messages [ 3849.219866] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3849.222826] LDISKFS-fs (dm-0): recovery complete [ 3849.228873] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3856.374408] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c8534ee [ 3861.590468] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3862.270311] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2017) [ 3862.271265] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2017) [ 3869.656621] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3871.725479] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3882.069215] Lustre: DEBUG MARKER: == replay-dual test 23c: c1 rmdir d1, M0 drop update reply and fail M0, c2 mkdir d1 ========================================================== 00:29:51 (1781238591) [ 3883.532382] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3883.543055] LustreError: 8414:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff93948515b0c0 x1867760095039872/t137438953491(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:326/0 lens 1984/4320 e 0 to 0 dl 1781238606 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3887.227379] Lustre: Failing over lustre-MDT0000 [ 3887.778807] Lustre: server umount lustre-MDT0000 complete [ 3907.472371] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3907.893972] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.38@tcp (not set up) [ 3907.907925] Lustre: Skipped 1 previous similar message [ 3913.232803] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3913.287591] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2049) [ 3913.290587] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2049) [ 3921.865386] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3924.215500] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3933.483907] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3935.523676] Lustre: Failing over lustre-MDT0000 [ 3936.074802] Lustre: server umount lustre-MDT0000 complete [ 3938.805181] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3958.957937] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3958.960092] LDISKFS-fs (dm-0): recovery complete [ 3958.976971] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3965.409147] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff939586e96d00 x1867760095082368/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3965.438884] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 19 previous similar messages [ 3970.901694] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 3971.059925] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 3971.074846] Lustre: Skipped 46 previous similar messages [ 3971.099712] Lustre: 79532:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3971.108902] Lustre: 79532:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 17 previous similar messages [ 3971.303913] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2081) [ 3971.307757] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2081) [ 3978.716865] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3980.531873] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3989.828964] Lustre: DEBUG MARKER: == replay-dual test 23d: c1 rmdir d1, M0 drop update reply and fail M0/M1, c2 mkdir d1 ========================================================== 00:31:38 (1781238698) [ 3994.690991] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3994.694072] LustreError: 61310:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff93959d3530c0 x1867760095109376/t146028888081(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:437/0 lens 1984/4320 e 0 to 0 dl 1781238717 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3998.323747] Lustre: Failing over lustre-MDT0000 [ 3998.771435] Lustre: server umount lustre-MDT0000 complete [ 4002.484569] LustreError: 10275:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781238714 with bad export cookie 13706422200134091173 [ 4002.492641] Lustre: Failing over lustre-MDT0001 [ 4002.507898] LustreError: 10275:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 4002.548721] LustreError: 80655:0:(ldlm_resource.c:1180:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x81:0x0].0x0 (ffff9395bf53a100) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 4002.601314] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.38@tcp (stopping) [ 4008.273098] Lustre: server umount lustre-MDT0001 complete [ 4028.284151] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4028.498888] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4028.650215] LustreError: 81348:0:(llog.c:1646:llog_backup()) MGC192.168.204.138@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 4028.667781] Lustre: 81348:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.138@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 4053.700330] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4053.726554] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4054.136350] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2113) [ 4054.164526] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2113) [ 4054.355801] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:193) [ 4054.357891] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:193) [ 4054.416304] Lustre: 81370:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9395b46943c0 x1867760075469952/t25769803783(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:565/0 lens 496/2888 e 0 to 0 dl 1781238845 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 4062.274935] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4063.966930] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4065.608325] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4076.340285] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4079.353511] Lustre: Failing over lustre-MDT0000 [ 4079.600373] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 4079.758144] Lustre: server umount lustre-MDT0000 complete [ 4103.252575] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4103.254647] LDISKFS-fs (dm-0): recovery complete [ 4103.270875] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4116.885651] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4117.711898] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2115 to 0x2c0000401:2145) [ 4117.712184] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2115 to 0x280000401:2145) [ 4125.131443] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4126.845876] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4135.982744] Lustre: DEBUG MARKER: == replay-dual test 24: reconstruct on non-existing object ========================================================== 00:34:05 (1781238845) [ 4137.176834] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 4137.184316] LustreError: 82123:0:(ldlm_lib.c:3327:target_send_reply_msg()) @@@ dropping reply req@ffff939585e803c0 x1867760075511296/t154618822673(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:648/0 lens 488/456 e 0 to 0 dl 1781238928 ref 1 fl Interpret:/200/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4221.488132] Lustre: lustre-MDT0000: Client 12ed2e13-20e3-4295-bd2d-2e458f4c4105 (at 192.168.204.38@tcp) reconnecting [ 4221.513642] Lustre: 81370:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff9395983f3480 x1867760075511296/t154618822673(0) o36->12ed2e13-20e3-4295-bd2d-2e458f4c4105@192.168.204.38@tcp:733/0 lens 488/3152 e 0 to 0 dl 1781239013 ref 1 fl Interpret:/202/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4229.172624] Lustre: DEBUG MARKER: == replay-dual test 25: replay|resend ==================== 00:35:38 (1781238938) [ 4231.380697] Lustre: *** cfs_fail_loc=304, val=0*** [ 4233.369042] Lustre: Failing over lustre-OST0000 [ 4233.514332] Lustre: server umount lustre-OST0000 complete [ 4233.697653] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4233.712654] Lustre: Skipped 36 previous similar messages [ 4252.424331] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4254.062865] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 4254.078404] Lustre: Skipped 10 previous similar messages [ 4254.347510] Lustre: lustre-OST0000: Recovery over after 0:01, of 4 clients 4 recovered and 0 were evicted. [ 4254.356940] Lustre: Skipped 12 previous similar messages [ 4259.552343] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4268.285839] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4270.237693] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4280.967261] Lustre: DEBUG MARKER: == replay-dual test 26: dbench and tar with mds failover ========================================================== 00:36:29 (1781238989) [ 4293.147256] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4297.326657] Lustre: DEBUG MARKER: test_26 fail mds1 1 times [ 4299.769638] Lustre: Failing over lustre-MDT0000 [ 4299.801666] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.38@tcp (stopping) [ 4299.813141] Lustre: Skipped 6 previous similar messages [ 4300.212228] Lustre: server umount lustre-MDT0000 complete [ 4319.712482] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1781239015/real 1781239015] req@ffff93948b7fb480 x1867760095301888/t0(0) o400->MGC192.168.204.138@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1781239031 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4319.772366] Lustre: 3666:0:(client.c:2480:ptlrpc_expire_one_request()) Skipped 17 previous similar messages [ 4326.678270] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 4326.680281] LDISKFS-fs (dm-0): recovery complete [ 4326.706561] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4331.039181] LustreError: 87122:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 4331.053309] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff93959d3530c0 x1867760095311232/t0(0) o250->MGC192.168.204.138@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4331.087938] LustreError: 3665:0:(client.c:1390:ptlrpc_import_delay_req()) Skipped 2 previous similar messages [ 4336.643093] Lustre: 87156:0:(ldlm_lib.c:2069:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4336.673344] Lustre: 87156:0:(ldlm_lib.c:2069:extend_recovery_timer()) Skipped 17 previous similar messages [ 4337.169433] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4339.368287] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2169 to 0x2c0000401:2209) [ 4339.368287] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2170 to 0x280000401:2209) [ 4346.273211] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4348.536771] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4361.084872] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4365.112420] Lustre: DEBUG MARKER: test_26 fail mds2 2 times [ 4367.332402] Lustre: Failing over lustre-MDT0001 [ 4367.396133] LustreError: lustre-MDT0001-osp-MDT0000: operation out_update to node 0@lo failed: rc = -107 [ 4367.410465] LustreError: Skipped 2 previous similar messages [ 4367.419621] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 4373.415995] Lustre: server umount lustre-MDT0001 complete [ 4377.578522] LustreError: 81370:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4377.607838] LustreError: 81370:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 273 previous similar messages [ 4397.485492] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4397.496234] LDISKFS-fs (dm-1): recovery complete [ 4397.511455] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4398.185221] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 4398.204214] Lustre: Skipped 9 previous similar messages [ 4398.277924] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 4398.285585] Lustre: Skipped 9 previous similar messages [ 4403.346566] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4405.976872] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:267 to 0x2c0000400:289) [ 4405.981626] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:267 to 0x280000400:289) [ 4412.685394] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4414.743548] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4428.372414] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4432.393573] Lustre: DEBUG MARKER: test_26 fail mds1 3 times [ 4434.526431] Lustre: Failing over lustre-MDT0000 [ 4434.763989] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4434.773902] Lustre: Skipped 7 previous similar messages [ 4434.909407] Lustre: server umount lustre-MDT0000 complete [ 4455.394506] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4455.417213] LustreError: Skipped 5 previous similar messages [ 4465.025970] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4465.037249] LDISKFS-fs (dm-0): recovery complete [ 4465.052374] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4480.004376] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xbe36fe5f0c86a44b [ 4484.598028] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4488.608679] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2299 to 0x2c0000401:2337) [ 4488.615148] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2299 to 0x280000401:2337) [ 4493.244381] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4495.251462] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4542.605685] Lustre: DEBUG MARKER: == replay-dual test 28: lock replay should be ordered: waiting after granted ========================================================== 00:40:51 (1781239251) [ 4560.430585] Lustre: Failing over lustre-OST0000 [ 4560.456185] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4560.463536] LustreError: 6512:0:(client.c:1380:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff9395b51b0b40 x1867760096066816/t0(0) o105->lustre-OST0000@192.168.204.38@tcp:15/16 lens 392/224 e 0 to 0 dl 0 ref 1 fl Rpc:QU/0/ffffffff rc 0/-1 job:'' uid:4294967295 gid:4294967295 projid:4294967295 [ 4560.469625] LustreError: 91797:0:(ldlm_resource.c:1180:ldlm_resource_complain()) filter-lustre-OST0000_UUID: namespace resource [0x280000401:0x93a:0x0].0x0 (ffff9395854bb000) refcount nonzero (2) after lock cleanup; forcing cleanup. [ 4560.495868] LustreError: 91797:0:(ldlm_resource.c:1180:ldlm_resource_complain()) Skipped 1 previous similar message [ 4562.809166] Lustre: server umount lustre-OST0000 complete [ 4583.179561] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4585.544555] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4585.598941] Lustre: lustre-OST0000-osc-MDT0001: Connection restored to 0@lo (at 0@lo) [ 4585.606657] Lustre: Skipped 27 previous similar messages [ 4591.113662] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4599.385260] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4601.405194] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4612.236607] Lustre: DEBUG MARKER: == replay-dual test 29: replay vs update with the same xid ========================================================== 00:42:01 (1781239321) [ 4613.940108] Lustre: DEBUG MARKER: SKIP: replay-dual test_29 needs >= 2 clients [ 4616.012480] Lustre: DEBUG MARKER: == replay-dual test 30: layout lock replay is not blocked on IO ========================================================== 00:42:04 (1781239324) [ 4619.624937] Lustre: Failing over lustre-MDT0000 [ 4620.183802] Lustre: server umount lustre-MDT0000 complete [ 4620.641749] LustreError: 10275:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_convert from 192.168.204.38@tcp arrived at 1781239332 with bad export cookie 13706422200134182075 [ 4638.745320] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4644.244968] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4645.252541] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2363 to 0x2c0000401:2401) [ 4645.256873] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2363 to 0x280000401:2401) [ 4654.600791] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4657.592864] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4666.911781] Lustre: DEBUG MARKER: == replay-dual test 31: deadlock on file_remove_privs and occupied mod rpc slots ========================================================== 00:42:55 (1781239375) [ 4670.635233] Lustre: Failing over lustre-OST0000 [ 4670.842597] Lustre: server umount lustre-OST0000 complete [ 4691.207493] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4699.301737] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4708.332676] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4710.763108] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in IDLE [ 4720.986881] Lustre: DEBUG MARKER: == replay-dual test 32: gap in update llog shouldn't break recovery ========================================================== 00:43:50 (1781239430) [ 4722.257267] Lustre: *** cfs_fail_loc=131d, val=10*** [ 4722.888347] Lustre: *** cfs_fail_loc=131d, val=0*** [ 4722.891469] Lustre: Skipped 9 previous similar messages [ 4723.934001] Lustre: *** cfs_fail_loc=131d, val=4294967282*** [ 4723.944376] Lustre: Skipped 13 previous similar messages [ 4727.046453] Lustre: Failing over lustre-MDT0001 [ 4727.507303] Lustre: server umount lustre-MDT0001 complete [ 4731.974705] Lustre: Failing over lustre-MDT0000 [ 4732.485101] Lustre: server umount lustre-MDT0000 complete [ 4740.720363] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4741.444191] Lustre: *** cfs_fail_loc=131d, val=4294967266*** [ 4741.448910] Lustre: Skipped 15 previous similar messages [ 4746.836346] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4755.723741] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4755.964222] Lustre: *** cfs_fail_loc=131d, val=4294967262*** [ 4755.970572] Lustre: Skipped 3 previous similar messages [ 4761.733665] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:325 to 0x280000400:353) [ 4761.737170] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:325 to 0x2c0000400:353) [ 4761.809403] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2497) [ 4761.809540] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2363 to 0x2c0000401:2433) [ 4762.043269] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4774.187510] Lustre: DEBUG MARKER: == replay-dual test 33: Check for OBD_INCOMPAT_MULTI_RPCS in last_rcvd after abort_recovery ========================================================== 00:44:43 (1781239483) [ 4781.256558] Lustre: Failing over lustre-MDT0001 [ 4781.553735] Lustre: server umount lustre-MDT0001 complete [ 4802.759675] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4807.804741] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4814.387975] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount REPLAY_WAIT mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4816.521073] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in REPLAY_WAIT state after 0 sec [ 4817.458718] Lustre: lustre-MDT0001: Aborting client recovery [ 4817.463762] LustreError: 99106:0:(ldlm_lib.c:2985:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 4817.474058] Lustre: 98585:0:(ldlm_lib.c:2388:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 4817.489296] Lustre: 98585:0:(ldlm_lib.c:2388:target_recovery_overseer()) Skipped 2 previous similar messages [ 4817.502644] Lustre: 98585:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client a4d52f74-feaa-4c5e-8d3d-efab64b573ba@ [ 4817.513091] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 4817.531068] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 4817.556680] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 4817.655913] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:325 to 0x280000400:385) [ 4817.655914] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:325 to 0x2c0000400:385) [ 4822.070463] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4824.395605] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4828.376275] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4836.231275] Lustre: Failing over lustre-MDT0001 [ 4836.738199] Lustre: server umount lustre-MDT0001 complete [ 4838.387105] Lustre: lustre-MDT0001-osp-MDT0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4838.417807] Lustre: Skipped 28 previous similar messages [ 4847.254253] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4852.095375] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug -1 all [ 4853.362664] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:325 to 0x2c0000400:417) [ 4853.363479] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:325 to 0x280000400:417) [ 4859.401269] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4861.342321] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4866.052702] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 4874.743197] Lustre: DEBUG MARKER: == replay-dual test complete, duration 4559 sec ========== 00:46:23 (1781239583) [ 4876.771356] Lustre: DEBUG MARKER: === replay-dual: start cleanup 00:46:25 (1781239585) === [ 4887.903590] Lustre: DEBUG MARKER: === replay-dual: finish cleanup 00:46:36 (1781239596) === [ 4890.421295] Lustre: Failing over lustre-MDT0000 [ 4890.738306] Lustre: server umount lustre-MDT0000 complete [ 4893.666213] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 4893.685327] LustreError: Skipped 5 previous similar messages [ 4918.550774] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4921.373432] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 4921.387921] Lustre: Skipped 10 previous similar messages [ 4923.886329] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 5061.500513] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 5061.514886] Lustre: 102009:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client a4d52f74-feaa-4c5e-8d3d-efab64b573ba@ [ 5061.534557] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 5061.555648] Lustre: lustre-MDT0000: Recovery over after 2:20, of 3 clients 2 recovered and 1 was evicted. [ 5061.579047] Lustre: Skipped 10 previous similar messages [ 5061.616570] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2529) [ 5061.617328] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2363 to 0x2c0000401:2465) [ 5067.931500] Lustre: DEBUG MARKER: oleg438-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 5070.507469] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5077.990436] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 5077.998835] Lustre: Skipped 8 previous similar messages [ 5083.839322] Lustre: server umount lustre-MDT0000 complete [ 5088.224348] LustreError: 96676:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 5088.233803] LustreError: 96676:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 200 previous similar messages [ 5095.012213] LustreError: 6508:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1781239806 with bad export cookie 13706422200134267601 [ 5095.033670] LustreError: MGC192.168.204.138@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 5095.036413] LustreError: 6508:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 5095.060329] LustreError: Skipped 3 previous similar messages [ 5095.549842] Lustre: server umount lustre-MDT0001 complete [ 5121.260539] Lustre: server umount lustre-OST0000 complete [ 5142.309880] Lustre: server umount lustre-OST0001 complete [ 5165.762196] Lustre: DEBUG MARKER: oleg438-server.virtnet: executing unload_modules_local [ 5169.679876] Key type lgssc unregistered [ 5170.119644] LNet: 104831:0:(lib-ptl.c:967:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 5170.132801] LNetError: 104831:0:(acceptor.c:246:lnet_acceptor_remove_socket()) Interface ens2 not found [ 5170.172234] LNet: Removed LNI 192.168.204.138@tcp [ 5171.507262] Key type .llcrypt unregistered [ 5171.510147] Key type ._llcrypt unregistered