[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-10.fc44 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 494638061 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2399.998 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE247C 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE2318 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 0022D8 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE238C 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE241C 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2454 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe2318-0xbffe238b] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe2317] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe238c-0xbffe241b] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe241c-0xbffe2453] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2454-0xbffe247b] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524584K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001010] APIC: Switch to symmetric I/O mode setup [ 0.002470] x2apic enabled [ 0.003010] Switched APIC routing to physical x2apic. [ 0.005010] kvm-guest: setup PV IPIs [ 0.007999] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.008000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x229835b7123, max_idle_ns: 440795242976 ns [ 0.008016] Calibrating delay loop (skipped) preset value.. 4799.99 BogoMIPS (lpj=2399998) [ 0.009028] pid_max: default: 32768 minimum: 301 [ 0.010141] LSM: Security Framework initializing [ 0.011040] Yama: becoming mindful. [ 0.012026] SELinux: Initializing. [ 0.013045] *** VALIDATE selinux *** [ 0.019632] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.023392] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.024149] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.026063] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.027104] *** VALIDATE tmpfs *** [ 0.028364] *** VALIDATE proc *** [ 0.029178] *** VALIDATE cgroup *** [ 0.030005] *** VALIDATE cgroup2 *** [ 0.031234] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.032158] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.033008] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.034032] Spectre V2 : User space: Vulnerable [ 0.035008] Speculative Store Bypass: Vulnerable [ 0.039133] debug: unmapping init [mem 0xffffffff8c059000-0xffffffff8c060fff] [ 0.042000] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.042775] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.043025] ... version: 2 [ 0.044012] ... bit width: 48 [ 0.045011] ... generic registers: 4 [ 0.046013] ... value mask: 0000ffffffffffff [ 0.047017] ... max period: 00007fffffffffff [ 0.048021] ... fixed-purpose events: 3 [ 0.049017] ... event mask: 000000070000000f [ 0.050441] rcu: Hierarchical SRCU implementation. [ 0.052687] smp: Bringing up secondary CPUs ... [ 0.053717] x86: Booting SMP configuration: [ 0.054040] .... node #0, CPUs: #1 #2 #3 [ 0.057534] smp: Brought up 1 node, 4 CPUs [ 0.059011] smpboot: Max logical packages: 1 [ 0.060022] smpboot: Total of 4 processors activated (19199.98 BogoMIPS) [ 0.138568] node 0 deferred pages initialised in 75ms [ 0.142060] devtmpfs: initialized [ 0.143233] x86/mm: Memory block size: 128MB [ 0.146834] gcov: version magic: 0x41383552 [ 0.148030] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.151072] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.152269] pinctrl core: initialized pinctrl subsystem [ 0.153168] [ 0.153765] ************************************************************* [ 0.154017] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.155014] ** ** [ 0.156012] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.157013] ** ** [ 0.158013] ** This means that this kernel is built to expose internal ** [ 0.159014] ** IOMMU data structures, which may compromise security on ** [ 0.160017] ** your system. ** [ 0.161015] ** ** [ 0.162017] ** If you see this message and you are not debugging the ** [ 0.163016] ** kernel, report this immediately to your vendor! ** [ 0.164017] ** ** [ 0.165016] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.166015] ************************************************************* [ 0.167740] NET: Registered protocol family 16 [ 0.168467] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.169074] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.170074] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.171489] cpuidle: using governor menu [ 0.172173] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.173529] PCI: Using configuration type 1 for base access [ 0.174127] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.181125] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.184023] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.187084] cryptd: max_cpu_qlen set to 1000 [ 0.189251] ACPI: Added _OSI(Module Device) [ 0.191014] ACPI: Added _OSI(Processor Device) [ 0.193013] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.195012] ACPI: Added _OSI(Processor Aggregator Device) [ 0.201065] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.209377] ACPI: Interpreter enabled [ 0.210043] ACPI: PM: (supports S0 S3 S4 S5) [ 0.210940] ACPI: Using IOAPIC for interrupt routing [ 0.211158] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.212385] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.221226] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.222045] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.223017] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.224105] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.226025] acpiphp: Slot [2] registered [ 0.226844] acpiphp: Slot [5] registered [ 0.227087] acpiphp: Slot [6] registered [ 0.228086] acpiphp: Slot [7] registered [ 0.228935] acpiphp: Slot [8] registered [ 0.229062] acpiphp: Slot [9] registered [ 0.229864] acpiphp: Slot [10] registered [ 0.230054] acpiphp: Slot [3] registered [ 0.230845] acpiphp: Slot [4] registered [ 0.231055] acpiphp: Slot [11] registered [ 0.231999] acpiphp: Slot [12] registered [ 0.232938] acpiphp: Slot [13] registered [ 0.233065] acpiphp: Slot [14] registered [ 0.233957] acpiphp: Slot [15] registered [ 0.234076] acpiphp: Slot [16] registered [ 0.234912] acpiphp: Slot [17] registered [ 0.235060] acpiphp: Slot [18] registered [ 0.235798] acpiphp: Slot [19] registered [ 0.236055] acpiphp: Slot [20] registered [ 0.236867] acpiphp: Slot [21] registered [ 0.237073] acpiphp: Slot [22] registered [ 0.237928] acpiphp: Slot [23] registered [ 0.238071] acpiphp: Slot [24] registered [ 0.238769] acpiphp: Slot [25] registered [ 0.239101] acpiphp: Slot [26] registered [ 0.240065] acpiphp: Slot [27] registered [ 0.240956] acpiphp: Slot [28] registered [ 0.241067] acpiphp: Slot [29] registered [ 0.241913] acpiphp: Slot [30] registered [ 0.242064] acpiphp: Slot [31] registered [ 0.243051] PCI host bridge to bus 0000:00 [ 0.244027] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.246029] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.249028] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.251019] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.252030] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.255036] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.256286] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.258810] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.260989] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.268014] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.271045] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.274039] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.276023] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.277013] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.279805] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.284215] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.287064] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.291000] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.295013] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.304015] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.307011] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.311785] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.316020] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.320013] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.336022] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.352000] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.364020] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.373027] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.393017] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.403032] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.408018] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.419019] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.442019] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.457509] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.467018] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.477017] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.495017] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.507219] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.519016] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.527021] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.545015] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.554413] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.563015] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.569017] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.589022] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.603199] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.605484] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.607524] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.609552] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.612293] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.617111] iommu: Default domain type: Passthrough [ 0.618431] SCSI subsystem initialized [ 0.619154] ACPI: bus type USB registered [ 0.621115] usbcore: registered new interface driver usbfs [ 0.622089] usbcore: registered new interface driver hub [ 0.624093] usbcore: registered new device driver usb [ 0.626131] pps_core: LinuxPPS API ver. 1 registered [ 0.627008] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.629054] PTP clock support registered [ 0.630185] EDAC MC: Ver: 3.0.0 [ 0.632133] PCI: Using ACPI for IRQ routing [ 0.634025] NetLabel: Initializing [ 0.635013] NetLabel: domain hash size = 128 [ 0.637014] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.639084] NetLabel: unlabeled traffic allowed by default [ 0.641454] vgaarb: loaded [ 0.643419] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.644009] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.653000] clocksource: Switched to clocksource kvm-clock [ 0.760942] VFS: Disk quotas dquot_6.6.0 [ 0.762608] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 0.764923] *** VALIDATE ramfs *** [ 0.766048] *** VALIDATE hugetlbfs *** [ 0.768272] pnp: PnP ACPI init [ 0.770433] pnp: PnP ACPI: found 6 devices [ 0.788528] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 0.792557] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 0.794855] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 0.797325] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 0.800396] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 0.803436] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 0.806326] NET: Registered protocol family 2 [ 0.810223] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 0.817123] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 0.820548] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 0.826251] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 0.829817] TCP: Hash tables configured (established 65536 bind 65536) [ 0.832588] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 0.835699] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.838330] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 0.841229] NET: Registered protocol family 1 [ 0.843677] RPC: Registered named UNIX socket transport module. [ 0.845708] RPC: Registered udp transport module. [ 0.847293] RPC: Registered tcp transport module. [ 0.848843] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 0.851048] NET: Registered protocol family 44 [ 0.852643] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 0.854716] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 0.856684] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 0.858869] PCI: CLS 0 bytes, default 64 [ 0.861602] Unpacking initramfs... [ 2.261450] debug: unmapping init [mem 0xffff8973bcc54000-0xffff8973bffbffff] [ 2.264526] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 2.266385] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 2.268758] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x229835b7123, max_idle_ns: 440795242976 ns [ 2.761876] Initialise system trusted keyrings [ 2.763575] Key type blacklist registered [ 2.765274] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 2.774359] zbud: loaded [ 2.777532] *** VALIDATE nfs *** [ 2.778760] *** VALIDATE nfs4 *** [ 2.780678] pstore: using deflate compression [ 2.785539] Platform Keyring initialized [ 2.900782] NET: Registered protocol family 38 [ 2.902728] Key type asymmetric registered [ 2.904306] Asymmetric key parser 'x509' registered [ 2.905655] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 2.908424] io scheduler mq-deadline registered [ 2.909851] io scheduler kyber registered [ 2.911431] io scheduler bfq registered [ 2.913196] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 2.916024] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 2.918790] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 2.921470] ACPI: Power Button [PWRF] [ 2.926909] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 2.933987] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 2.947119] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 2.953961] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 2.968051] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 2.994265] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 3.022277] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 3.027404] Non-volatile memory driver v1.3 [ 3.029137] Linux agpgart interface v0.103 [ 3.063738] virtio_blk virtio1: [vda] 145912 512-byte logical blocks (74.7 MB/71.2 MiB) [ 3.067502] vda: detected capacity change from 0 to 74706944 [ 3.094770] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 3.098878] vdb: detected capacity change from 0 to 1073741824 [ 3.114706] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.118462] vdc: detected capacity change from 0 to 2621440000 [ 3.134612] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 3.137468] vdd: detected capacity change from 0 to 2621440000 [ 3.152105] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.155221] vde: detected capacity change from 0 to 4294967296 [ 3.172868] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 3.175461] vdf: detected capacity change from 0 to 4294967296 [ 3.182529] libphy: Fixed MDIO Bus: probed [ 3.187575] usbcore: registered new interface driver usbserial_generic [ 3.190459] usbserial: USB Serial support registered for generic [ 3.193102] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 3.198386] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 3.200352] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 3.203536] mousedev: PS/2 mouse device common for all mice [ 3.206689] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 3.208500] rtc_cmos 00:05: RTC can wake from S4 [ 3.213306] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 3.213967] rtc_cmos 00:05: registered as rtc0 [ 3.219413] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 3.223201] intel_pstate: CPU model not supported [ 3.225735] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 3.227442] hid: raw HID events driver (C) Jiri Kosina [ 3.231123] usbcore: registered new interface driver usbhid [ 3.232875] usbhid: USB HID core driver [ 3.234504] drop_monitor: Initializing network drop monitor service [ 3.236618] Initializing XFRM netlink socket [ 3.238520] NET: Registered protocol family 10 [ 3.241597] Segment Routing with IPv6 [ 3.243322] NET: Registered protocol family 17 [ 3.246532] mpls_gso: MPLS GSO support [ 3.251540] RAS: Correctable Errors collector initialized. [ 3.254727] AVX version of gcm_enc/dec engaged. [ 3.256965] AES CTR mode by8 optimization enabled [ 3.333539] sched_clock: Marking stable (3333517525, 0)->(4227542011, -894024486) [ 3.336674] registered taskstats version 1 [ 3.338586] Loading compiled-in X.509 certificates [ 3.341255] zswap: loaded using pool lzo/zbud [ 3.367658] Key type big_key registered [ 3.381597] Key type encrypted registered [ 3.383789] ima: No TPM chip found, activating TPM-bypass! [ 3.386145] ima: Allocated hash algorithm: sha1 [ 3.388337] ima: No architecture policies found [ 3.390690] evm: Initialising EVM extended attributes: [ 3.393270] evm: security.selinux [ 3.394771] evm: security.ima [ 3.396286] evm: security.capability [ 3.397949] evm: HMAC attrs: 0x1 [ 3.400642] rtc_cmos 00:05: setting system clock to 2026-08-18 17:40:41 UTC (1787074841) [ 3.407846] debug: unmapping init [mem 0xffffffff8d003000-0xffffffff8d1fffff] [ 3.411613] debug: unmapping init [mem 0xffffffff8bd82000-0xffffffff8c058fff] [ 3.421113] Write protecting the kernel read-only data: 28672k [ 3.425463] debug: unmapping init [mem 0xffffffff8a403000-0xffffffff8a5fffff] [ 3.429088] debug: unmapping init [mem 0xffffffff8ad14000-0xffffffff8adfffff] [ 3.462894] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 3.473761] systemd[1]: Detected virtualization kvm. [ 3.478882] systemd[1]: Detected architecture x86-64. [ 3.480948] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 3.513597] systemd[1]: No hostname configured. [ 3.515333] systemd[1]: Set hostname to . [ 3.517360] random: systemd: uninitialized urandom read (16 bytes read) [ 3.519819] systemd[1]: Initializing machine ID from random generator. [ 3.659231] random: systemd: uninitialized urandom read (16 bytes read) [ 3.661049] systemd[1]: Reached target Local File Systems. [ OK ] Reached target Local File Systems. [ 3.664588] random: systemd: uninitialized urandom read (16 bytes read) [ 3.666717] systemd[1]: Listening on udev Kernel Socket. [ OK ] Listening on udev Kernel Socket. [ 3.669753] systemd[1]: Listening on Journal Socket (/dev/log). [ OK ] Listening on Journal Socket (/dev/log). [ OK ] Reached target Initrd Root Device. [ OK ] Listening on udev Control Socket. [ OK ] Listening on Journal Socket. Starting Journal Service... Starting Create Volatile Files and Directories... Starting Create list of required st…ce nodes for the current kernel... [ OK ] Reached target Timers. [ OK ] Started Memstrack Anylazing Service. [ OK ] Reached target Slices. Starting Apply Kernel Variables... Starting Setup Virtual Console... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Sockets. [ OK ] Reached target Swap. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Started Create Volatile Files and Directories. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. [ OK ] Started Setup Virtual Console. Starting dracut cmdline hook... Starting Create Static Device Nodes in /dev... [ OK ] Started Journal Service. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 4.298427] device-mapper: uevent: version 1.0.3 [ 4.300958] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. Starting dracut initqueue hook... [ 5.001415] virtio_net virtio0 ens2: renamed from eth0 [ 5.030691] random: fast init done [ 5.121138] scsi host0: ata_piix [ 5.177329] scsi host1: ata_piix [ 5.206674] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 5.209381] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 9.909630] random: crng init done [ 9.918848] random: 7 urandom warning(s) missed due to ratelimiting [ 9.975161] dracut-initqueue[582]: RTNETLINK answers: File exists Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Mounting /sysroot... [ 11.637158] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped target Timers. [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Initrd Root Device. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Paths. [ OK ] Stopped target Sockets. [ OK ] Stopped target Slices. [ OK ] Stopped target System Initialization. [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped target Swap. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. Stopping udev Kernel Device Manager... [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Local File Systems. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Closed udev Kernel Socket. [ OK ] Closed udev Control Socket. Starting Cleanup udevd DB... [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 14.721949] printk: systemd: 25 output lines suppressed due to ratelimiting [ 15.857819] SELinux: Disabled at runtime. [ 16.031920] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 16.052890] systemd[1]: Detected virtualization kvm. [ 16.057744] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 17.615868] systemd[1]: initrd-switch-root.service: Succeeded. [ 17.633376] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 17.649911] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 17.661299] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 17.675871] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 17.707399] systemd[1]: Starting Journal Service... Starting Journal Service... [ 17.720278] systemd[1]: Stopped target Switch Root. [ OK ] Stopped target Switch Root. [ OK ] Created slice system-getty.slice. [ OK ] Created slice User and Session Slice. [ OK ] Created slice system-serial\x2dgetty.slice. [ OK ] Listening on udev Kernel Socket. [ OK ] Listening on udev Control Socket. [ OK ] Stopped target Initrd File Systems. [ OK ] Stopped target Initrd Root File System. [ OK ] Started Dispatch Password Requests to Console Directory Watch. Mounting Kernel Debug File System... [ OK ] Listening on Process Core Dump Socket. [ OK ] Listening on initctl Compatibility Named Pipe. [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. Mounting Huge Pages File System... [ OK ] Reached target Slices. [ OK ] Listening on RPCbind Server Activation Socket. [ OK ] Reached target RPC Port Mapper. Starting udev Coldplug all Devices... Activating swap /dev/disk/by-label/SWAP... Starting Create list of required st…ce nodes for the current kernel... [ OK ] Reached target rpc_pipefs.target. Starting Apply Kernel Variables... Starting Remount Root and Kernel File Systems... [ 18.289291] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS [ OK ] Created slice system-sshd\x2dkeygen.slice. Mounting POSIX Message Queue File System... [ OK ] Started Forward Password Requests to Wall Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Mounted Kernel Debug File System. [ OK ] Mounted Huge Pages File System. [ OK ] Started Journal Service. [ OK ] Activated swap /dev/disk/by-label/SWAP. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. [ OK ] Mounted POSIX Message Queue File System. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... [ OK ] Reached target Swap. Starting Flush Journal to Persistent Storage... [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started udev Coldplug all Devices. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Reached target Local File Systems (Pre). Mounting /home/green/git/lustre-release... Mounting /mnt... Starting udev Kernel Device Manager... [ OK ] Mounted /mnt. [ 19.473146] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Kernel Device Manager. [ 20.783916] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 21.063810] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 21.357252] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 21.555929] EDAC sbridge: Ver: 1.1.2 [* ] A start job is running for Configur…-only root support (8s / no limit) [** ] A start job is running for Configur…-only root support (8s / no limit)[ 26.279995] Key type dns_resolver registered [*** ] A start job is running for Configur…-only root support (8s / no limit) [ *** ] A start job is running for Configur…-only root support (9s / no limit)[ 26.975924] NFS: Registering the id_resolver key type [ 26.977773] Key type id_resolver registered [ 26.979586] Key type id_legacy registered [ *** ] A start job is running for Configur…-only root support (9s / no limit) [ OK ] Started Configure read-only root support. Starting Load/Save Random Seed... [ OK ] Reached target Local File Systems. Starting Mark the need to relabel after reboot... Starting Rebuild Dynamic Linker Cache... Starting Create Volatile Files and Directories... [ OK ] Started Load/Save Random Seed. [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Create Volatile Files and Directories. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started dnf makecache --timer. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Started daily update of the root trust anchor for DNSSEC. [ OK ] Reached target Timers. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. [ OK ] Started irqbalance daemon. [ OK ] Started D-Bus System Message Bus. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. Starting Network Manager... Starting Restore /run/initramfs on shutdown... Starting Login Service... [ OK ] Reached target sshd-keygen.target. [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Network Manager. [ OK ] Reached target Network. Starting Dynamic System Tuning Daemon... Starting OpenSSH server daemon... Starting GSSAPI Proxy Daemon... Starting Network Manager Wait Online... [ OK ] Started Login Service. Starting Hostname Service... [ OK ] Started OpenSSH server daemon. [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... [ OK ] Started Permit User Sessions. [ OK ] Started Serial Getty on ttyS1. [ OK ] Started Serial Getty on ttyS0. [ OK ] Started Command Scheduler. [ OK ] Started Getty on tty1. [ OK ] Reached target Login Prompts. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting Notify NFS peers of a restart... Starting Crash recovery kernel arming... Starting System Logging Service... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Starting Authorization Manager... Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg120-server login: [ 48.020087] hrtimer: interrupt took 2337150 ns [ 91.853644] libcfs: loading out-of-tree module taints kernel. [ 91.888621] Key type ._llcrypt registered [ 91.889921] Key type .llcrypt registered [ 91.976078] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_hostid [ 111.465906] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing load_modules_local [ 112.755649] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 112.763113] alg: No test for adler32 (adler32-zlib) [ 114.081638] Lustre: Lustre: Build Version: 2.17.57_1_g711d631 [ 114.779211] LNet: Added LNI 192.168.201.120@tcp [8/256/0/180] [ 116.535627] Key type lgssc registered [ 118.669633] Lustre: Echo OBD driver; http://www.lustre.org/ [ 137.533368] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 185.060506] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing load_modules_local [ 199.238577] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 199.341163] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 200.733425] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 200.781660] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 200.917250] Lustre: lustre-MDT0000: new disk, initializing [ 201.153222] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 201.189922] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 206.757124] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 220.650426] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 220.726464] Lustre: 6508:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 220.750622] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 220.756446] Lustre: Skipped 1 previous similar message [ 220.834058] Lustre: lustre-MDT0001: new disk, initializing [ 220.897362] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 220.925971] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 220.940483] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 224.861612] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 229.527349] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 241.551588] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 241.909708] Lustre: lustre-OST0000: new disk, initializing [ 241.915571] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 241.933928] Lustre: 8450:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 241.963078] Lustre: lustre-OST0000: Not available for connect from 0@lo (not set up) [ 242.093525] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 247.287104] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 247.307350] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 247.377117] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 250.414978] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 266.773939] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 266.882549] Lustre: lustre-OST0001: new disk, initializing [ 266.885850] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 266.891886] Lustre: 9521:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 266.962480] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 274.006637] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 274.016599] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 274.083972] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 274.158543] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 287.386633] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 295.438768] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 302.530490] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing check_logdir /tmp/testlogs/ [ 308.344747] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing yml_node [ 314.594344] Lustre: DEBUG MARKER: Client: 2.17.57.1 [ 317.789677] Lustre: DEBUG MARKER: MDS: 2.17.57.1 [ 321.508754] Lustre: DEBUG MARKER: OSS: 2.17.57.1 [ 323.658193] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-single ============----- Tue Aug 18 13:45:59 EDT 2026 [ 348.764275] Lustre: DEBUG MARKER: excepting tests: 110f 131b 59 36 [ 350.883147] Lustre: DEBUG MARKER: === replay-single: start setup 13:46:27 (1787075187) === [ 359.101394] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing check_config_client /mnt/lustre [ 383.090792] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 386.548840] Lustre: 13334:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 391.255770] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 398.062435] Lustre: DEBUG MARKER: === replay-single: finish setup 13:47:14 (1787075234) === [ 400.805870] Lustre: DEBUG MARKER: == replay-single test 100a: DNE: create striped dir, drop update rep from MDT1, fail MDT1 ========================================================== 13:47:17 (1787075237) [ 402.112560] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 402.117521] LustreError: 6519:0:(ldlm_lib.c:3346:target_send_reply_msg()) @@@ dropping reply req@ffff897427a35c00 x1873883904993536/t4294967300(0) o1000->lustre-MDT0000-mdtlov_UUID@0@lo:66/0 lens 264/4320 e 0 to 0 dl 1787075251 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up1-0.0' uid:0 gid:0 projid:4294967295 [ 403.746290] Lustre: Failing over lustre-MDT0001 [ 404.225578] Lustre: server umount lustre-MDT0001 complete [ 405.488919] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 405.509900] Lustre: lustre-MDT0001-osp-MDT0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 407.520471] Lustre: lustre-MDT0001-lwp-OST0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 407.545962] Lustre: Skipped 1 previous similar message [ 409.726235] LustreError: 6514:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.201.20@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 409.750487] LustreError: 6514:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 6 previous similar messages [ 412.641346] LustreError: 6515:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 412.655862] LustreError: 6515:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 2 previous similar messages [ 414.821740] LustreError: 6516:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.201.20@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 417.764461] LustreError: 8432:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 417.786146] LustreError: 8432:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 2 previous similar messages [ 418.292386] Lustre: 7800:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075240/real 1787075240] req@ffff897427a35880 x1873883904993536/t0(0) o1000->lustre-MDT0001-osp-MDT0000@0@lo:24/4 lens 264/4320 e 0 to 1 dl 1787075256 ref 2 fl Rpc:XQr/200/ffffffff rc 0/-1 job:'osp_up1-0.0' uid:0 gid:0 projid:4294967295 [ 422.882108] LustreError: 6515:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 422.910899] LustreError: 6515:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 3 previous similar messages [ 423.675581] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 424.227259] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 425.057058] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 428.893568] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 429.546681] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 429.571364] Lustre: lustre-MDT0001: Recovery over after 0:04, of 2 clients 2 recovered and 0 were evicted. [ 439.836726] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 441.957055] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 452.692654] Lustre: DEBUG MARKER: == replay-single test 100b: DNE: create striped dir, fail MDT0 ========================================================== 13:48:09 (1787075289) [ 453.868139] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 453.870070] LustreError: 6514:0:(ldlm_lib.c:3346:target_send_reply_msg()) @@@ dropping reply req@ffff897307fe6a00 x1873883886046464/t4294967361(0) o36->284ced2d-a060-449c-99a0-b921118b22d2@192.168.201.20@tcp:156/0 lens 560/544 e 0 to 0 dl 1787075341 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 455.570069] Lustre: Failing over lustre-MDT0000 [ 455.790140] LustreError: 6514:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.201.20@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 457.882148] Lustre: server umount lustre-MDT0000 complete [ 460.258218] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 460.262262] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 460.279046] Lustre: Skipped 3 previous similar messages [ 475.616356] LustreError: 6514:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 475.660168] LustreError: 6514:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 19 previous similar messages [ 476.644454] Lustre: 3640:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075298/real 1787075298] req@ffff897306643b80 x1873883905038208/t0(0) o400->MGC192.168.201.120@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1787075314 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 476.677883] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 476.891286] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 486.912689] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974248f6680 x1873883905047296/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 487.205781] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 488.753988] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 492.433919] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 492.515259] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 492.537281] Lustre: Skipped 2 previous similar messages [ 492.596891] Lustre: lustre-MDT0000: Recovery over after 0:04, of 2 clients 2 recovered and 0 were evicted. [ 492.613218] Lustre: 6516:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff89743e11fb80 x1873883886046464/t4294967361(0) o36->284ced2d-a060-449c-99a0-b921118b22d2@192.168.201.20@tcp:195/0 lens 560/3152 e 0 to 0 dl 1787075380 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 492.650536] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:6 to 0x280000401:33) [ 492.651723] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:6 to 0x2c0000401:33) [ 502.167472] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 503.982332] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 513.268278] Lustre: DEBUG MARKER: == replay-single test 100c: DNE: create striped dir, abort_recov_mdt mds2 ========================================================== 13:49:09 (1787075349) [ 522.173867] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 524.587587] Lustre: Failing over lustre-MDT0001 [ 525.130694] Lustre: server umount lustre-MDT0001 complete [ 528.360795] Lustre: lustre-MDT0001-lwp-OST0001: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 528.369233] LustreError: 13604:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 528.376805] Lustre: Skipped 1 previous similar message [ 528.400725] LustreError: 13604:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 16 previous similar messages [ 528.420826] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 538.947805] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 538.953592] LDISKFS-fs (dm-1): recovery complete [ 538.980168] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 539.402123] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 539.462541] Lustre: lustre-MDT0001: Aborting MDT recovery [ 539.483546] LustreError: 17671:0:(lod_dev.c:511:lod_sub_recovery_thread()) lustre-MDT0000-osp-MDT0001: get update log duration 0, retries 0, failed: rc = -108 [ 539.750174] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 544.765917] Lustre: lustre-MDT0001-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 544.770944] Lustre: Skipped 3 previous similar messages [ 544.808987] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 544.896967] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 544.923170] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 544.975627] Lustre: lustre-MDT0001: Recovery over after 0:05, of 2 clients 2 recovered and 0 were evicted. [ 545.012572] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:46 to 0x2c0000400:65) [ 545.014076] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:46 to 0x280000400:65) [ 562.351061] Lustre: Failing over lustre-MDT0001 [ 562.842351] Lustre: server umount lustre-MDT0001 complete [ 565.217132] Lustre: lustre-MDT0001-osp-MDT0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 565.240292] Lustre: Skipped 1 previous similar message [ 582.067099] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 582.488590] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 583.892444] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 587.762882] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 587.778144] Lustre: Skipped 2 previous similar messages [ 587.799828] Lustre: lustre-MDT0001: Recovery over after 0:04, of 2 clients 2 recovered and 0 were evicted. [ 587.863587] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 587.869488] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 588.089987] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 598.645316] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 600.933996] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 612.260595] Lustre: DEBUG MARKER: == replay-single test 100d: DNE: cancel update logs upon recovery abort ========================================================== 13:50:48 (1787075448) [ 631.502154] Lustre: Failing over lustre-MDT0001 [ 631.830962] Lustre: server umount lustre-MDT0001 complete [ 633.824266] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 633.826695] Lustre: lustre-MDT0001-lwp-OST0000: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 633.827050] LustreError: 6516:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 633.827057] LustreError: 6516:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 25 previous similar messages [ 633.887752] Lustre: Skipped 3 previous similar messages [ 643.941038] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 644.349434] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 644.349737] Lustre: lustre-MDT0001: Aborting client recovery [ 644.358521] LustreError: 20213:0:(ldlm_lib.c:3004:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 644.362209] Lustre: 20237:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 644.366159] Lustre: 20237:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 284ced2d-a060-449c-99a0-b921118b22d2@ [ 644.373393] Lustre: lustre-MDT0001: disconnecting 2 stale clients [ 644.380592] Lustre: lustre-MDT0001-osd: cancel update llog [0x2400013a0:0x3:0x0] [ 644.388539] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000bd1:0x3:0x0] [ 644.431686] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:129) [ 644.435751] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:129) [ 649.748478] Lustre: lustre-MDT0001-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 649.769717] Lustre: Skipped 2 previous similar messages [ 649.787808] LustreError: lustre-MDT0001-osp-MDT0000: This client was evicted by lustre-MDT0001; in progress operations using this service will fail. [ 651.042608] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 679.014515] Lustre: DEBUG MARKER: == replay-single test 100e: DNE: create striped dir on MDT0 and MDT1, fail MDT0, MDT1 ========================================================== 13:51:55 (1787075515) [ 689.228697] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 698.959422] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 701.800611] Lustre: Failing over lustre-MDT0000 [ 702.358609] Lustre: server umount lustre-MDT0000 complete [ 706.016950] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 706.028854] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 706.067288] Lustre: Skipped 3 previous similar messages [ 707.165821] Lustre: Failing over lustre-MDT0001 [ 707.166759] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787075545 with bad export cookie 15188512954156389848 [ 707.167056] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 707.183288] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 707.542772] Lustre: server umount lustre-MDT0001 complete [ 735.066330] LDISKFS-fs (dm-1): 5 truncates cleaned up [ 735.071921] LDISKFS-fs (dm-1): recovery complete [ 735.092991] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 735.397770] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 735.400123] LDISKFS-fs (dm-0): recovery complete [ 735.414064] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 735.428164] LustreError: 23153:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 735.436913] Lustre: 23153:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 752.098409] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff897308fd8700 x1873883905404672/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 752.724378] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 752.731363] Lustre: Skipped 1 previous similar message [ 752.768811] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 752.782985] Lustre: Skipped 2 previous similar messages [ 753.092990] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_connect to node 0@lo failed: rc = -114 [ 753.100930] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 753.104922] Lustre: Skipped 2 previous similar messages [ 753.378436] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 758.878785] Lustre: lustre-MDT0000: Recovery over after 0:05, of 2 clients 2 recovered and 0 were evicted. [ 758.899086] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:44 to 0x2c0000401:65) [ 758.906249] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:44 to 0x280000401:65) [ 758.936954] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:161) [ 758.948534] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:161) [ 760.166264] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 761.081475] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 766.432195] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075549/real 1787075549] req@ffff897308fd8a80 x1873883905400064/t0(0) o400->lustre-MDT0001-lwp-OST0001@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787075604 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 766.477610] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 771.616559] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075554/real 1787075554] req@ffff8974084b3100 x1873883905400576/t0(0) o400->lustre-MDT0001-lwp-OST0001@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787075609 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 771.641489] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 2 previous similar messages [ 775.510171] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 777.695501] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075560/real 1787075560] req@ffff897308fd9f80 x1873883905401088/t0(0) o400->lustre-MDT0001-lwp-OST0000@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787075615 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 777.767736] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 778.032303] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 779.983493] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 786.911631] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075569/real 1787075569] req@ffff897309359180 x1873883905401984/t0(0) o400->lustre-MDT0001-lwp-OST0001@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787075624 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 786.933992] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 5 previous similar messages [ 792.463985] Lustre: DEBUG MARKER: == replay-single test 101: Shouldn't reassign precreated objs to other files after recovery ========================================================== 13:53:48 (1787075628) [ 801.975869] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 833.264516] Lustre: Failing over lustre-MDT0000 [ 833.728339] Lustre: server umount lustre-MDT0000 complete [ 835.045061] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 835.053754] LustreError: 23178:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 835.055699] Lustre: Skipped 4 previous similar messages [ 835.082990] LustreError: 23178:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 43 previous similar messages [ 849.858846] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 849.860897] LDISKFS-fs (dm-0): recovery complete [ 849.867350] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 849.979162] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 850.264743] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 850.267714] Lustre: lustre-MDT0000: Aborting client recovery [ 850.273818] Lustre: Skipped 1 previous similar message [ 850.282768] LustreError: 25616:0:(ldlm_lib.c:3004:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 850.290449] Lustre: 25649:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 850.302353] Lustre: 25649:0:(ldlm_lib.c:2404:target_recovery_overseer()) Skipped 2 previous similar messages [ 850.310143] Lustre: 25649:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 284ced2d-a060-449c-99a0-b921118b22d2@ [ 850.332081] Lustre: 25649:0:(genops.c:1622:class_disconnect_stale_exports()) Skipped 1 previous similar message [ 850.338677] Lustre: lustre-MDT0000: disconnecting 2 stale clients [ 850.349411] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 850.363519] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 850.422443] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:44 to 0x2c0000401:609) [ 850.439647] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:44 to 0x280000401:609) [ 855.530679] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 855.545419] Lustre: Skipped 5 previous similar messages [ 855.558201] LustreError: lustre-MDT0000-osp-MDT0001: This client was evicted by lustre-MDT0000; in progress operations using this service will fail. [ 856.304879] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 958.341753] Lustre: DEBUG MARKER: == replay-single test 102a: check resend (request lost) with multiple modify RPCs in flight ========================================================== 13:56:35 (1787075795) [ 959.959491] Lustre: *** cfs_fail_loc=159, val=0*** [ 959.966303] Lustre: Skipped 1 previous similar message [ 1016.017947] Lustre: lustre-MDT0000: Client 284ced2d-a060-449c-99a0-b921118b22d2 (at 192.168.201.20@tcp) reconnecting [ 1024.390372] Lustre: DEBUG MARKER: == replay-single test 102b: check resend (reply lost) with multiple modify RPCs in flight ========================================================== 13:57:40 (1787075860) [ 1026.339343] Lustre: *** cfs_fail_loc=15a, val=0*** [ 1081.458373] Lustre: lustre-MDT0001: Client 284ced2d-a060-449c-99a0-b921118b22d2 (at 192.168.201.20@tcp) reconnecting [ 1081.493899] Lustre: 23179:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff8973091f1c00 x1873883889067776/t25769807836(0) o36->284ced2d-a060-449c-99a0-b921118b22d2@192.168.201.20@tcp:29/0 lens 488/3152 e 0 to 0 dl 1787075969 ref 1 fl Interpret:/202/0 rc 0/0 job:'chmod.0' uid:0 gid:0 projid:4294967295 [ 1081.550532] Lustre: 23179:0:(mdt_recovery.c:102:mdt_req_from_lrd()) Skipped 3 previous similar messages [ 1091.218801] Lustre: DEBUG MARKER: == replay-single test 102c: check replay w/o reconstruction with multiple mod RPCs in flight ========================================================== 13:58:47 (1787075927) [ 1099.460992] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1100.490669] Lustre: *** cfs_fail_loc=15a, val=0*** [ 1100.506050] Lustre: Skipped 6 previous similar messages [ 1104.606596] Lustre: Failing over lustre-MDT0000 [ 1104.942690] Lustre: server umount lustre-MDT0000 complete [ 1106.404953] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1106.413154] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1106.426871] Lustre: Skipped 1 previous similar message [ 1106.437732] LustreError: 23579:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1106.466940] LustreError: 23579:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 11 previous similar messages [ 1125.855701] Lustre: 3640:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787075947/real 1787075947] req@ffff897425f09c00 x1873883906026240/t0(0) o400->MGC192.168.201.120@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1787075963 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1125.894086] Lustre: 3640:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1125.905652] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1127.831710] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 1127.835323] LDISKFS-fs (dm-0): recovery complete [ 1127.845596] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1136.096847] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974384a6a00 x1873883906034688/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1136.453529] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1136.467785] Lustre: Skipped 2 previous similar messages [ 1136.520586] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1136.525902] Lustre: Skipped 2 previous similar messages [ 1137.475451] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1137.480076] Lustre: Skipped 1 previous similar message [ 1141.734658] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 1141.759000] Lustre: Skipped 3 previous similar messages [ 1141.861201] Lustre: lustre-MDT0000: Recovery over after 0:04, of 2 clients 2 recovered and 0 were evicted. [ 1141.872351] Lustre: Skipped 1 previous similar message [ 1141.893177] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1141.943901] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:617 to 0x280000401:641) [ 1141.954218] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:617 to 0x2c0000401:641) [ 1152.286799] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1153.887205] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1162.980786] Lustre: DEBUG MARKER: == replay-single test 102d: check replay [ 1164.491895] Lustre: *** cfs_fail_loc=15a, val=0*** [ 1164.500162] Lustre: Skipped 6 previous similar messages [ 1169.370321] Lustre: Failing over lustre-MDT0001 [ 1169.675603] Lustre: server umount lustre-MDT0001 complete [ 1172.448186] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 1189.396333] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1189.787230] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 1191.551676] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1195.020572] Lustre: lustre-MDT0001: Recovery over after 0:04, of 2 clients 2 recovered and 0 were evicted. [ 1195.020885] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1195.073345] Lustre: 23180:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff897305701880 x1873883889138176/t25769807890(0) o36->284ced2d-a060-449c-99a0-b921118b22d2@192.168.201.20@tcp:143/0 lens 488/3152 e 0 to 0 dl 1787076083 ref 1 fl Interpret:/202/0 rc 0/0 job:'chmod.0' uid:0 gid:0 projid:4294967295 [ 1195.085121] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:705) [ 1195.086391] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:705) [ 1195.104635] Lustre: 23180:0:(mdt_recovery.c:102:mdt_req_from_lrd()) Skipped 3 previous similar messages [ 1206.968247] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 1209.729103] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1220.380943] Lustre: DEBUG MARKER: == replay-single test 103: Check otr_next_id overflow ==== 14:00:56 (1787076056) [ 1226.036165] Lustre: Failing over lustre-MDT0000 [ 1226.353507] Lustre: server umount lustre-MDT0000 complete [ 1246.671071] Lustre: 3639:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787076068/real 1787076068] req@ffff8974384a7480 x1873883906106368/t0(0) o400->MGC192.168.201.120@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1787076084 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1246.697881] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1247.044473] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1255.913503] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xd2c8700f42ec83f6 [ 1255.934794] Lustre: MGC192.168.201.120@tcp: Connection restored to 0@lo (at 0@lo) [ 1255.945478] Lustre: Skipped 6 previous similar messages [ 1256.594738] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1258.176216] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1261.633050] Lustre: lustre-MDT0000: Recovery over after 0:03, of 2 clients 2 recovered and 0 were evicted. [ 1261.747955] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:673) [ 1261.753977] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:673) [ 1262.207418] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1276.925983] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1278.931654] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1288.944671] Lustre: DEBUG MARKER: == replay-single test 110a: DNE: create striped dir, fail MDT1 ========================================================== 14:02:05 (1787076125) [ 1297.564537] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1299.676133] Lustre: Failing over lustre-MDT0000 [ 1299.969987] Lustre: server umount lustre-MDT0000 complete [ 1302.507470] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1302.516154] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1302.528409] Lustre: Skipped 12 previous similar messages [ 1319.400143] Lustre: 3639:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787076140/real 1787076140] req@ffff8974383eed80 x1873883906156544/t0(0) o400->MGC192.168.201.120@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1787076156 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1319.439680] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1323.360494] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1323.363760] LDISKFS-fs (dm-0): recovery complete [ 1323.374814] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1329.663559] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xd2c8700f42ec8921 [ 1330.030760] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1330.274602] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1335.387798] Lustre: lustre-MDT0000: Recovery over after 0:05, of 2 clients 2 recovered and 0 were evicted. [ 1335.443470] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:705) [ 1335.443744] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:705) [ 1336.294036] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1347.665817] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1349.361983] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1358.137919] Lustre: DEBUG MARKER: == replay-single test 110b: DNE: create striped dir, fail MDT1 and client ========================================================== 14:03:14 (1787076194) [ 1366.092652] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1368.893557] Lustre: Failing over lustre-MDT0000 [ 1369.259470] Lustre: server umount lustre-MDT0000 complete [ 1386.962387] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1392.421751] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1392.424644] LDISKFS-fs (dm-0): recovery complete [ 1392.432088] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1397.215663] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff897420961f80 x1873883906209792/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1397.592915] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1397.599045] Lustre: Skipped 3 previous similar messages [ 1397.655442] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1402.771851] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1402.861730] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 1402.873291] Lustre: Skipped 9 previous similar messages [ 1413.403280] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 1415.743462] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:56 [ 1420.905895] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:51 [ 1426.022615] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:46 [ 1431.137069] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:41 [ 1436.265522] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:36 [ 1446.498625] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:26 [ 1446.526191] Lustre: Skipped 1 previous similar message [ 1466.979908] Lustre: lustre-MDT0000: Denying connection for new client b9ed249c-f62c-4614-b176-9337a464e0d5 (at 192.168.201.20@tcp), waiting for 2 known clients (0 recovered, 1 in progress, and 0 evicted) to recover in 0:05 [ 1466.996900] Lustre: Skipped 3 previous similar messages [ 1472.500645] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1472.506699] Lustre: 35046:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 284ced2d-a060-449c-99a0-b921118b22d2@ [ 1472.541016] Lustre: 35046:0:(genops.c:1622:class_disconnect_stale_exports()) Skipped 1 previous similar message [ 1472.552569] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1472.574704] Lustre: lustre-MDT0000: Recovery over after 1:10, of 2 clients 1 recovered and 1 was evicted. [ 1472.645393] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:737) [ 1472.646771] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:737) [ 1486.881070] Lustre: DEBUG MARKER: == replay-single test 110c: DNE: create striped dir, fail MDT2 ========================================================== 14:05:23 (1787076323) [ 1495.216734] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 1497.337690] Lustre: Failing over lustre-MDT0001 [ 1497.754248] Lustre: server umount lustre-MDT0001 complete [ 1500.130050] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 1521.910540] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 1521.914518] LDISKFS-fs (dm-1): recovery complete [ 1521.923888] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1522.325099] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 1523.314387] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1523.330621] Lustre: Skipped 1 previous similar message [ 1527.246732] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1527.914719] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:737) [ 1527.918064] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:737) [ 1537.281894] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 1539.221915] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1548.423816] Lustre: DEBUG MARKER: == replay-single test 110d: DNE: create striped dir, fail MDT2 and client ========================================================== 14:06:24 (1787076384) [ 1555.263329] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 1557.246116] Lustre: Failing over lustre-MDT0001 [ 1557.534933] Lustre: server umount lustre-MDT0001 complete [ 1580.018429] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 1580.021613] LDISKFS-fs (dm-1): recovery complete [ 1580.033358] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1585.406094] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1594.621327] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 1596.872759] Lustre: lustre-MDT0001: Denying connection for new client 52aaf10a-a617-4cfc-8773-557c87a75860 (at 192.168.201.20@tcp), waiting for 2 known clients (1 recovered, 0 in progress, and 0 evicted) to recover in 0:58 [ 1596.884109] Lustre: Skipped 1 previous similar message [ 1655.500201] Lustre: lustre-MDT0001: recovery is timed out, evict stale exports [ 1655.508088] Lustre: 38876:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client b9ed249c-f62c-4614-b176-9337a464e0d5@ [ 1655.520960] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 1655.574061] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:769) [ 1655.580517] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:769) [ 1666.995421] Lustre: DEBUG MARKER: == replay-single test 110e: DNE: create striped dir, uncommit on MDT2, fail client/MDT1/MDT2 ========================================================== 14:08:23 (1787076503) [ 1674.825198] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 1681.759632] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1683.482723] Lustre: Failing over lustre-MDT0000 [ 1683.714419] Lustre: server umount lustre-MDT0000 complete [ 1686.786602] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787076524 with bad export cookie 15188512954156617642 [ 1686.789487] Lustre: Failing over lustre-MDT0001 [ 1686.790725] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1686.796714] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 1687.020402] Lustre: server umount lustre-MDT0001 complete [ 1710.343043] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 1710.346927] LDISKFS-fs (dm-1): recovery complete [ 1710.347597] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1710.351170] LDISKFS-fs (dm-0): recovery complete [ 1710.358365] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1710.368407] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1710.515159] LustreError: 41830:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 1710.521481] Lustre: 41830:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 1711.519195] LustreError: 41831:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 1711.525562] LustreError: 41831:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff8974383ed180 x1873883906371456/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1787076549 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1711.540873] LustreError: 41831:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 1711.583616] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974251cfb80 x1873883906374016/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1711.584115] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1711.604289] LustreError: 41838:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1711.619473] Lustre: Skipped 13 previous similar messages [ 1711.648283] LustreError: 41838:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 176 previous similar messages [ 1711.827963] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 1711.830639] Lustre: Skipped 1 previous similar message [ 1712.054352] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 1712.072531] Lustre: Skipped 9 previous similar messages [ 1716.793188] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1716.926949] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1717.433721] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:769) [ 1717.436852] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:769) [ 1727.472383] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 1729.608670] Lustre: lustre-MDT0001: Denying connection for new client ab46c86e-1e5a-4e4e-ba62-9d0d3ce9f6bd (at 192.168.201.20@tcp), waiting for 2 known clients (1 recovered, 0 in progress, and 0 evicted) to recover in 0:57 [ 1729.618264] Lustre: Skipped 11 previous similar messages [ 1743.311192] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787076526/real 1787076526] req@ffff8974383ed500 x1873883906370432/t0(0) o400->lustre-MDT0000-lwp-OST0000@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787076581 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1743.348654] Lustre: 3642:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1787.500452] Lustre: lustre-MDT0001: recovery is timed out, evict stale exports [ 1787.507895] Lustre: 41883:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 52aaf10a-a617-4cfc-8773-557c87a75860@ [ 1787.527132] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 1787.564714] Lustre: lustre-MDT0001: Recovery over after 1:10, of 2 clients 1 recovered and 1 was evicted. [ 1787.570473] Lustre: Skipped 3 previous similar messages [ 1787.606538] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:801) [ 1787.607034] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:801) [ 1799.661823] Lustre: DEBUG MARKER: SKIP: replay-single test_110f skipping excluded test 110f [ 1801.876624] Lustre: DEBUG MARKER: == replay-single test 110g: DNE: create striped dir, uncommit on MDT1, fail client/MDT1/MDT2 ========================================================== 14:10:38 (1787076638) [ 1809.950513] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1817.512940] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 1819.458121] Lustre: Failing over lustre-MDT0000 [ 1819.638044] Lustre: server umount lustre-MDT0000 complete [ 1823.902527] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787076662 with bad export cookie 15188512954156622031 [ 1823.903477] Lustre: Failing over lustre-MDT0001 [ 1823.909809] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1823.914662] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 1824.383710] Lustre: server umount lustre-MDT0001 complete [ 1849.678508] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 1849.682327] LDISKFS-fs (dm-1): recovery complete [ 1849.685227] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1849.694729] LDISKFS-fs (dm-0): recovery complete [ 1849.705585] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1849.710100] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1849.933443] LustreError: 45264:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 1849.942789] Lustre: 45264:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 1869.792271] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff897420961880 x1873883906448512/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1870.384767] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_connect to node 0@lo failed: rc = -114 [ 1870.398588] LustreError: Skipped 2 previous similar messages [ 1875.321853] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1875.464153] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1875.942019] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1875.955267] Lustre: Skipped 4 previous similar messages [ 1876.386883] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:833) [ 1876.397853] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:833) [ 1885.650990] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 1888.119342] Lustre: lustre-MDT0000: Denying connection for new client abf2c67d-c65d-4b20-b163-c8164aede4e3 (at 192.168.201.20@tcp), waiting for 2 known clients (1 recovered, 0 in progress, and 0 evicted) to recover in 0:57 [ 1888.138527] Lustre: Skipped 11 previous similar messages [ 1945.500167] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1945.505449] Lustre: 45361:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client ab46c86e-1e5a-4e4e-ba62-9d0d3ce9f6bd@ [ 1945.524996] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1945.606163] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:801) [ 1945.607914] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:801) [ 1960.111972] Lustre: DEBUG MARKER: == replay-single test 111a: DNE: unlink striped dir, fail MDT1 ========================================================== 14:13:16 (1787076796) [ 1969.272791] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1971.924056] Lustre: Failing over lustre-MDT0000 [ 1972.254144] Lustre: server umount lustre-MDT0000 complete [ 1990.113953] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1995.225496] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1995.227576] LDISKFS-fs (dm-0): recovery complete [ 1995.241708] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2000.354165] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974084b1f80 x1873883906525696/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2000.547460] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2000.551615] Lustre: Skipped 6 previous similar messages [ 2000.594421] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2000.596154] Lustre: Skipped 3 previous similar messages [ 2005.256629] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2006.182451] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:833) [ 2006.185872] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:833) [ 2016.214036] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2018.142825] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2029.327444] Lustre: DEBUG MARKER: == replay-single test 111b: DNE: unlink striped dir, fail MDT2 ========================================================== 14:14:25 (1787076865) [ 2039.185698] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2041.774318] Lustre: Failing over lustre-MDT0001 [ 2041.829134] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 2042.162541] Lustre: server umount lustre-MDT0001 complete [ 2067.159979] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2067.163872] LDISKFS-fs (dm-1): recovery complete [ 2067.174414] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2067.424558] Lustre: lustre-MDT0001: Not available for connect from 0@lo (not set up) [ 2067.431136] Lustre: Skipped 3 previous similar messages [ 2073.992612] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2084.877718] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2143.502979] Lustre: lustre-MDT0001: recovery is timed out, evict stale exports [ 2143.507580] Lustre: 49546:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client abf2c67d-c65d-4b20-b163-c8164aede4e3@ [ 2143.526676] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 2143.619495] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:865) [ 2143.619709] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:865) [ 2153.196420] Lustre: DEBUG MARKER: == replay-single test 111c: DNE: unlink striped dir, uncommit on MDT1, fail client/MDT1/MDT2 ========================================================== 14:16:29 (1787076989) [ 2161.786737] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2171.245498] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2173.500487] Lustre: Failing over lustre-MDT0000 [ 2174.022188] Lustre: server umount lustre-MDT0000 complete [ 2178.789470] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787077016 with bad export cookie 15188512954156626763 [ 2178.791377] Lustre: Failing over lustre-MDT0001 [ 2178.799186] LustreError: 7467:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 2179.183610] Lustre: server umount lustre-MDT0001 complete [ 2195.936371] Lustre: 3639:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787077018/real 1787077018] req@ffff8974073d4000 x1873883906625280/t0(0) o400->lustre-MDT0001-lwp-OST0001@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787077034 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2195.965079] Lustre: 3639:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 31 previous similar messages [ 2205.870044] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2205.874985] LDISKFS-fs (dm-0): recovery complete [ 2205.880097] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2205.886260] LDISKFS-fs (dm-1): recovery complete [ 2205.896177] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2205.938195] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2226.970221] Lustre: lustre-MDT0001-osp-MDT0000: Connection restored to 0@lo (at 0@lo) [ 2226.980594] Lustre: Skipped 18 previous similar messages [ 2227.025960] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:897) [ 2227.026158] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:897) [ 2230.386981] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2230.501908] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2242.708403] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2245.213210] Lustre: lustre-MDT0000: Denying connection for new client f546faae-d5af-400b-b5f6-2496da318e99 (at 192.168.201.20@tcp), waiting for 2 known clients (1 recovered, 0 in progress, and 0 evicted) to recover in 0:51 [ 2245.243436] Lustre: Skipped 22 previous similar messages [ 2296.501591] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2296.514270] Lustre: 52556:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client d02df14b-9957-4995-8855-b7f07b756cf0@ [ 2296.532071] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2296.597631] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:865) [ 2296.598330] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:865) [ 2312.107770] Lustre: DEBUG MARKER: == replay-single test 111d: DNE: unlink striped dir, uncommit on MDT2, fail client/MDT1/MDT2 ========================================================== 14:19:08 (1787077148) [ 2322.163428] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2332.455664] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2334.779576] Lustre: Failing over lustre-MDT0000 [ 2335.185252] Lustre: server umount lustre-MDT0000 complete [ 2337.763177] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2337.778290] Lustre: Skipped 20 previous similar messages [ 2337.784417] LustreError: 8441:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2337.822134] LustreError: 8441:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 78 previous similar messages [ 2339.625786] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787077177 with bad export cookie 15188512954156629668 [ 2339.633064] Lustre: Failing over lustre-MDT0001 [ 2339.639036] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2339.639046] LustreError: Skipped 1 previous similar message [ 2339.643034] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 2340.003604] Lustre: server umount lustre-MDT0001 complete [ 2367.086232] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2367.088529] LDISKFS-fs (dm-0): recovery complete [ 2367.106521] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2367.157160] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2367.159940] LDISKFS-fs (dm-1): recovery complete [ 2367.186204] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2384.415850] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff89743e56a680 x1873883906703488/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2385.427604] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_connect to node 0@lo failed: rc = -114 [ 2385.442928] LustreError: Skipped 2 previous similar messages [ 2389.616023] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 1 client reconnects [ 2389.631198] Lustre: Skipped 5 previous similar messages [ 2390.281775] Lustre: lustre-MDT0000: Recovery over after 0:01, of 1 clients 1 recovered and 0 were evicted. [ 2390.297793] Lustre: Skipped 6 previous similar messages [ 2390.379126] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:897) [ 2390.379286] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:897) [ 2391.504577] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2391.604086] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2403.074613] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2459.500245] Lustre: lustre-MDT0001: recovery is timed out, evict stale exports [ 2459.505656] Lustre: 55985:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client f546faae-d5af-400b-b5f6-2496da318e99@ [ 2459.534511] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 2459.699848] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:929) [ 2459.707509] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:929) [ 2473.435166] Lustre: DEBUG MARKER: == replay-single test 111e: DNE: unlink striped dir, uncommit on MDT2, fail MDT1/MDT2 ========================================================== 14:21:49 (1787077309) [ 2483.000350] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2492.580593] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2494.920430] Lustre: Failing over lustre-MDT0000 [ 2495.214456] Lustre: server umount lustre-MDT0000 complete [ 2499.467258] LustreError: 6499:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787077337 with bad export cookie 15188512954156631873 [ 2499.481785] Lustre: Failing over lustre-MDT0001 [ 2499.821751] Lustre: server umount lustre-MDT0001 complete [ 2526.638941] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2526.643329] LDISKFS-fs (dm-1): recovery complete [ 2526.668301] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2526.689177] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2526.691039] LDISKFS-fs (dm-0): recovery complete [ 2526.710155] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2526.923489] LustreError: 59268:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 2526.928971] Lustre: 59268:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 2544.607712] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff89743e21ed80 x1873883906778752/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2544.956669] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 2544.963675] Lustre: Skipped 5 previous similar messages [ 2550.558193] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2551.177882] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2551.262248] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:961) [ 2551.263534] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:961) [ 2551.409654] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:929) [ 2551.420512] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:929) [ 2561.675193] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2563.497755] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2565.368716] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2573.682512] Lustre: DEBUG MARKER: == replay-single test 111f: DNE: unlink striped dir, uncommit on MDT1, fail MDT1/MDT2 ========================================================== 14:23:30 (1787077410) [ 2582.182763] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2590.872413] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2593.063666] Lustre: Failing over lustre-MDT0000 [ 2593.317448] Lustre: server umount lustre-MDT0000 complete [ 2597.993464] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787077436 with bad export cookie 15188512954156634225 [ 2598.001519] Lustre: Failing over lustre-MDT0001 [ 2598.007921] LustreError: 19651:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 2601.595391] Lustre: lustre-MDT0001: Not available for connect from 192.168.201.20@tcp (stopping) [ 2601.609563] Lustre: Skipped 1 previous similar message [ 2604.551746] Lustre: server umount lustre-MDT0001 complete [ 2630.406498] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2630.408828] LDISKFS-fs (dm-1): recovery complete [ 2630.418027] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2630.712823] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2630.714035] LDISKFS-fs (dm-0): recovery complete [ 2630.720769] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2630.735900] LustreError: 62718:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 2630.746072] Lustre: 62718:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 2643.203189] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 2643.210571] Lustre: Skipped 7 previous similar messages [ 2648.024153] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2648.403882] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2649.611506] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:993) [ 2649.614805] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:993) [ 2652.003429] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:961) [ 2652.003487] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:961) [ 2658.875595] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2660.820878] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2662.174561] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2671.258228] Lustre: DEBUG MARKER: == replay-single test 111g: DNE: unlink striped dir, fail MDT1/MDT2 ========================================================== 14:25:07 (1787077507) [ 2680.491447] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2689.185844] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2691.851924] Lustre: Failing over lustre-MDT0000 [ 2692.169865] Lustre: server umount lustre-MDT0000 complete [ 2696.109855] LustreError: 6500:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787077534 with bad export cookie 15188512954156636542 [ 2696.119108] Lustre: Failing over lustre-MDT0001 [ 2696.128587] LustreError: 6500:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 2696.452413] Lustre: server umount lustre-MDT0001 complete [ 2715.487700] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787077537/real 1787077537] req@ffff897307d6e300 x1873883906882816/t0(0) o400->lustre-MDT0001-lwp-OST0001@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787077553 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2715.519719] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 35 previous similar messages [ 2722.647964] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2722.651854] LDISKFS-fs (dm-1): recovery complete [ 2722.668963] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2722.796982] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2722.801646] LDISKFS-fs (dm-0): recovery complete [ 2722.816813] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2723.040847] LustreError: 66185:0:(llog.c:1655:llog_backup()) MGC192.168.201.120@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 2723.053577] Lustre: 66185:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.201.120@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 2741.223265] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xd2c8700f42ece2bc [ 2747.044937] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2747.543810] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2747.987085] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1025) [ 2747.990509] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:1025) [ 2748.099988] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:993) [ 2748.110161] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:993) [ 2758.493340] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2760.586676] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2762.246922] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2771.098469] Lustre: DEBUG MARKER: == replay-single test 112a: DNE: cross MDT rename, fail MDT1 ========================================================== 14:26:47 (1787077607) [ 2772.558933] Lustre: DEBUG MARKER: SKIP: replay-single test_112a needs >= 4 MDTs [ 2774.485975] Lustre: DEBUG MARKER: == replay-single test 112b: DNE: cross MDT rename, fail MDT2 ========================================================== 14:26:50 (1787077610) [ 2776.043112] Lustre: DEBUG MARKER: SKIP: replay-single test_112b needs >= 4 MDTs [ 2778.111861] Lustre: DEBUG MARKER: == replay-single test 112c: DNE: cross MDT rename, fail MDT3 ========================================================== 14:26:54 (1787077614) [ 2779.612714] Lustre: DEBUG MARKER: SKIP: replay-single test_112c needs >= 4 MDTs [ 2781.759838] Lustre: DEBUG MARKER: == replay-single test 112d: DNE: cross MDT rename, fail MDT4 ========================================================== 14:26:57 (1787077617) [ 2783.821070] Lustre: DEBUG MARKER: SKIP: replay-single test_112d needs >= 4 MDTs [ 2785.750914] Lustre: DEBUG MARKER: == replay-single test 112e: DNE: cross MDT rename, fail MDT1 and MDT2 ========================================================== 14:27:02 (1787077622) [ 2787.517864] Lustre: DEBUG MARKER: SKIP: replay-single test_112e needs >= 4 MDTs [ 2789.404925] Lustre: DEBUG MARKER: == replay-single test 112f: DNE: cross MDT rename, fail MDT1 and MDT3 ========================================================== 14:27:05 (1787077625) [ 2791.070466] Lustre: DEBUG MARKER: SKIP: replay-single test_112f needs >= 4 MDTs [ 2792.992584] Lustre: DEBUG MARKER: == replay-single test 112g: DNE: cross MDT rename, fail MDT1 and MDT4 ========================================================== 14:27:09 (1787077629) [ 2794.951958] Lustre: DEBUG MARKER: SKIP: replay-single test_112g needs >= 4 MDTs [ 2796.812424] Lustre: DEBUG MARKER: == replay-single test 112h: DNE: cross MDT rename, fail MDT2 and MDT3 ========================================================== 14:27:13 (1787077633) [ 2798.574630] Lustre: DEBUG MARKER: SKIP: replay-single test_112h needs >= 4 MDTs [ 2801.020772] Lustre: DEBUG MARKER: == replay-single test 112i: DNE: cross MDT rename, fail MDT2 and MDT4 ========================================================== 14:27:17 (1787077637) [ 2802.625197] Lustre: DEBUG MARKER: SKIP: replay-single test_112i needs >= 4 MDTs [ 2804.356722] Lustre: DEBUG MARKER: == replay-single test 112j: DNE: cross MDT rename, fail MDT3 and MDT4 ========================================================== 14:27:20 (1787077640) [ 2805.931723] Lustre: DEBUG MARKER: SKIP: replay-single test_112j needs >= 4 MDTs [ 2807.618289] Lustre: DEBUG MARKER: == replay-single test 112k: DNE: cross MDT rename, fail MDT1,MDT2,MDT3 ========================================================== 14:27:24 (1787077644) [ 2809.676418] Lustre: DEBUG MARKER: SKIP: replay-single test_112k needs >= 4 MDTs [ 2811.496041] Lustre: DEBUG MARKER: == replay-single test 112l: DNE: cross MDT rename, fail MDT1,MDT2,MDT4 ========================================================== 14:27:28 (1787077648) [ 2813.439318] Lustre: DEBUG MARKER: SKIP: replay-single test_112l needs >= 4 MDTs [ 2815.681578] Lustre: DEBUG MARKER: == replay-single test 112m: DNE: cross MDT rename, fail MDT1,MDT3,MDT4 ========================================================== 14:27:31 (1787077651) [ 2817.471550] Lustre: DEBUG MARKER: SKIP: replay-single test_112m needs >= 4 MDTs [ 2819.798410] Lustre: DEBUG MARKER: == replay-single test 112n: DNE: cross MDT rename, fail MDT2,MDT3,MDT4 ========================================================== 14:27:35 (1787077655) [ 2821.801530] Lustre: DEBUG MARKER: SKIP: replay-single test_112n needs >= 4 MDTs [ 2824.324461] Lustre: DEBUG MARKER: == replay-single test 115: failover for create/unlink striped directory ========================================================== 14:27:40 (1787077660) [ 2833.190828] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 2835.669599] Lustre: Failing over lustre-MDT0001 [ 2835.907208] Lustre: server umount lustre-MDT0001 complete [ 2859.332621] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 2859.337079] LDISKFS-fs (dm-1): recovery complete [ 2859.359232] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2864.788892] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2865.133172] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 2865.150390] Lustre: Skipped 30 previous similar messages [ 2865.413565] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:1057) [ 2865.423606] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1057) [ 2874.204740] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 2876.090781] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2886.233595] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2889.301460] Lustre: Failing over lustre-MDT0000 [ 2889.733357] Lustre: server umount lustre-MDT0000 complete [ 2906.079234] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2906.091927] LustreError: Skipped 3 previous similar messages [ 2912.984377] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2912.988806] LDISKFS-fs (dm-0): recovery complete [ 2912.997276] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2917.151227] LustreError: 71676:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 2917.169533] LustreError: 71676:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff89740fa68a80 x1873883907011968/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1787077754 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2917.205747] LustreError: 71676:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 2922.873822] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2923.204491] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:1025) [ 2923.205400] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1025) [ 2933.349883] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2934.699154] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2944.696859] Lustre: DEBUG MARKER: == replay-single test 116a: large update log master MDT recovery ========================================================== 14:29:40 (1787077780) [ 2952.693260] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2953.792843] Lustre: *** cfs_fail_loc=1702, val=0*** [ 2956.624746] Lustre: Failing over lustre-MDT0000 [ 2957.203232] Lustre: server umount lustre-MDT0000 complete [ 2958.454663] LustreError: 66210:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.201.20@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2958.476571] LustreError: 66210:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 145 previous similar messages [ 2958.825549] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2958.843358] Lustre: Skipped 33 previous similar messages [ 2980.864118] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 2980.875960] LDISKFS-fs (dm-0): recovery complete [ 2980.894076] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2984.417890] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974073d6680 x1873883907071104/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2989.152776] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 2990.268659] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1057) [ 2990.275390] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:1057) [ 2997.550375] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 2998.883348] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3006.898954] Lustre: DEBUG MARKER: == replay-single test 116b: large update log slave MDT recovery ========================================================== 14:30:43 (1787077843) [ 3014.979463] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 3016.064560] Lustre: *** cfs_fail_loc=1702, val=0*** [ 3018.579900] Lustre: Failing over lustre-MDT0001 [ 3018.923980] Lustre: server umount lustre-MDT0001 complete [ 3020.774889] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3020.783309] LustreError: Skipped 5 previous similar messages [ 3041.703259] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 3041.706603] LDISKFS-fs (dm-1): recovery complete [ 3041.726638] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3042.405658] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 3042.419084] Lustre: Skipped 9 previous similar messages [ 3047.119407] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3047.425889] Lustre: lustre-MDT0001: Recovery over after 0:05, of 2 clients 2 recovered and 0 were evicted. [ 3047.434039] Lustre: Skipped 10 previous similar messages [ 3047.458246] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1089) [ 3047.466461] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:670 to 0x2c0000400:1089) [ 3055.854186] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 3057.323282] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3066.631580] Lustre: DEBUG MARKER: == replay-single test 117: DNE: cross MDT unlink, fail MDT1 and MDT2 ========================================================== 14:31:42 (1787077902) [ 3068.549117] Lustre: DEBUG MARKER: SKIP: replay-single test_117 needs >= 4 MDTs [ 3070.176612] Lustre: DEBUG MARKER: == replay-single test 118: invalidate osp update will not cause update log corruption ========================================================== 14:31:46 (1787077906) [ 3071.880734] Lustre: *** cfs_fail_loc=1705, val=0*** [ 3081.172670] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3083.647979] Lustre: Failing over lustre-MDT0000 [ 3083.945543] Lustre: server umount lustre-MDT0000 complete [ 3107.361609] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 3107.364474] LDISKFS-fs (dm-0): recovery complete [ 3107.374105] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3111.922901] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xd2c8700f42ed0b18 [ 3111.943755] LustreError: 77773:0:(ldlm_resource.c:1207:ldlm_resource_complain()) MGC192.168.201.120@tcp: namespace resource [0x65727473756c:0x5:0x0].0x0 (ffff89741ffbe900) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3117.781303] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:1089) [ 3117.786150] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1089) [ 3117.998577] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3128.420275] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3129.855205] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3137.150670] Lustre: DEBUG MARKER: == replay-single test 119: timeout of normal replay does not cause DNE replay fails ========================================================== 14:32:53 (1787077973) [ 3145.486570] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3148.110423] Lustre: Failing over lustre-MDT0000 [ 3148.144114] LustreError: 3641:0:(client.c:1381:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff897425c7fb80 x1873883907199104/t0(0) o103->lustre-MDT0001-osp-MDT0000@0@lo:17/18 lens 328/224 e 0 to 0 dl 0 ref 1 fl Rpc:QU/200/ffffffff rc 0/-1 job:'jbd2/dm-0-8.0' uid:0 gid:0 projid:4294967295 [ 3148.256435] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3148.258601] Lustre: Skipped 4 previous similar messages [ 3148.535677] Lustre: server umount lustre-MDT0000 complete [ 3162.785929] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 3162.790215] LDISKFS-fs (dm-0): recovery complete [ 3162.799371] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3163.124421] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 3163.127401] Lustre: Skipped 10 previous similar messages [ 3165.282484] Lustre: 66210:0:(ldlm_lib.c:2085:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 60, extend: 0 [ 3167.253623] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3168.225572] Lustre: 8441:0:(ldlm_lib.c:2085:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 60, extend: 0 [ 3168.232805] LustreError: 79906:0:(ldlm_lib.c:2690:replay_request_or_update()) cfs_fail_timeout id 714 sleeping for 65000ms [ 3174.038278] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3233.239157] LustreError: 79906:0:(ldlm_lib.c:2690:replay_request_or_update()) cfs_fail_timeout id 714 awake [ 3233.250034] Lustre: 79906:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp [ 3233.285710] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3233.297714] Lustre: 79906:0:(ldlm_lib.c:1914:abort_req_replay_queue()) @@@ aborted: req@ffff8973097a8a80 x1873883889554304/t0(81604378629) o36->98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp:634/0 lens 528/0 e 7 to 0 dl 1787078084 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 3233.340091] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3233.351305] Lustre: 79906:0:(ldlm_lib.c:2085:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 60, extend: 1 [ 3233.357230] Lustre: lustre-MDT0000: Denying connection for new client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp), waiting for 2 known clients (1 recovered, 0 in progress, and 1 evicted) already passed deadline 0:07 [ 3233.389408] Lustre: Skipped 21 previous similar messages [ 3233.437879] Lustre: 79906:0:(ldlm_lib.c:2394:target_recovery_overseer()) lustre-MDT0000 recovery is aborted by hard timeout [ 3233.446602] Lustre: 79906:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 3233.457179] Lustre: 79906:0:(ldlm_lib.c:2404:target_recovery_overseer()) Skipped 2 previous similar messages [ 3233.468315] Lustre: lustre-MDT0000-osd: cancel update llog [0x200001b70:0x1:0x0] [ 3233.492858] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240002b13:0x1:0x0] [ 3233.521892] Lustre: 79906:0:(ldlm_lib.c:2951:target_recovery_thread()) too long recovery - read logs [ 3233.536092] LustreError: dumping log to /tmp/lustre-log.1787078071.79906 [ 3233.791934] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1121) [ 3233.795694] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:1121) [ 3238.976355] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 58 sec [ 3249.579495] Lustre: DEBUG MARKER: == replay-single test 120: DNE fail abort should stop both normal and DNE replay ========================================================== 14:34:46 (1787078086) [ 3256.501781] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3264.334909] Lustre: Failing over lustre-MDT0000 [ 3264.493032] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3264.734498] Lustre: server umount lustre-MDT0000 complete [ 3277.517720] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3277.520795] LDISKFS-fs (dm-0): recovery complete [ 3277.528153] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3277.942890] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 3277.953739] Lustre: Skipped 9 previous similar messages [ 3278.016458] Lustre: lustre-MDT0000: Aborting client recovery [ 3278.026932] LustreError: 81866:0:(ldlm_lib.c:3004:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 3278.035177] LustreError: 81899:0:(lod_dev.c:511:lod_sub_recovery_thread()) lustre-MDT0001-osp-MDT0000: get update log duration 0, retries 0, failed: rc = -108 [ 3278.044088] Lustre: 81900:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 3278.044099] Lustre: 81900:0:(ldlm_lib.c:2404:target_recovery_overseer()) Skipped 1 previous similar message [ 3278.066780] Lustre: 81900:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 98986846-94b1-4b41-aaaa-108b03b874d0@ [ 3278.075469] Lustre: lustre-MDT0000: disconnecting 2 stale clients [ 3278.082224] Lustre: lustre-MDT0000-osd: cancel update llog [0x200009870:0x3:0x0] [ 3278.099732] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x2400090a2:0x1:0x0] [ 3278.179837] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1153) [ 3278.211201] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:657 to 0x280000401:1153) [ 3283.130749] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3283.453195] LustreError: lustre-MDT0000-osp-MDT0001: This client was evicted by lustre-MDT0000; in progress operations using this service will fail. [ 3303.943341] Lustre: DEBUG MARKER: == replay-single test 121: lock replay timed out and race ========================================================== 14:35:40 (1787078140) [ 3307.158365] Lustre: Failing over lustre-MDT0000 [ 3307.625936] Lustre: server umount lustre-MDT0000 complete [ 3313.638353] Lustre: *** cfs_fail_loc=721, val=0*** [ 3313.643541] Lustre: Skipped 5 previous similar messages [ 3314.155428] Lustre: *** cfs_fail_loc=721, val=0*** [ 3314.159054] Lustre: Skipped 3 previous similar messages [ 3318.410078] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3318.525849] Lustre: *** cfs_fail_loc=721, val=0*** [ 3318.529556] Lustre: Skipped 4 previous similar messages [ 3323.023962] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3323.872424] Lustre: *** cfs_fail_loc=721, val=1*** [ 3323.874146] Lustre: Skipped 96 previous similar messages [ 3323.923502] Lustre: *** cfs_fail_loc=721, val=1*** [ 3323.931021] Lustre: Skipped 6 previous similar messages [ 3328.991965] Lustre: *** cfs_fail_loc=721, val=1*** [ 3328.994187] Lustre: Skipped 26 previous similar messages [ 3334.260908] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:54 [ 3339.232596] Lustre: *** cfs_fail_loc=721, val=1*** [ 3339.239860] Lustre: Skipped 36 previous similar messages [ 3350.632506] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:37 [ 3354.079213] Lustre: 3638:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787078162/real 1787078162] req@ffff897412b4fb80 x1873883907341312/t0(0) o400->lustre-MDT0000-osp-MDT0001@0@lo:24/4 lens 224/224 e 0 to 1 dl 1787078192 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 3354.106218] Lustre: 3638:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 15 previous similar messages [ 3354.113165] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3354.119748] Lustre: *** cfs_fail_loc=721, val=1*** [ 3354.124410] Lustre: 83319:0:(tgt_handler.c:581:tgt_handle_recovery()) @@@ obsoleted, rq_xid=1873883889695744, exp_last_xid=1873883889697535 req@ffff897425d4e300 x1873883889695744/t0(0) o101->98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp:0/0 lens 328/0 e 0 to 0 dl 1787078168 ref 1 fl Complete:/240/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 3355.745108] Lustre: *** cfs_fail_loc=721, val=1*** [ 3355.750324] Lustre: Skipped 64 previous similar messages [ 3365.987925] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:22 [ 3382.376422] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:06 [ 3384.288428] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3384.296336] Lustre: *** cfs_fail_loc=721, val=1*** [ 3384.299081] Lustre: 83319:0:(tgt_handler.c:581:tgt_handle_recovery()) @@@ obsoleted, rq_xid=1873883889695616, exp_last_xid=1873883889697535 req@ffff897425d4e680 x1873883889695616/t0(0) o101->98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp:0/0 lens 328/0 e 0 to 0 dl 1787078168 ref 1 fl Complete:/240/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 3389.350400] Lustre: *** cfs_fail_loc=721, val=1*** [ 3389.354776] Lustre: Skipped 120 previous similar messages [ 3398.767669] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:05 [ 3414.128564] Lustre: lustre-MDT0000: Recovery already passed deadline 0:09. If you do not want to wait more, you may force taget eviction via 'lctl --device lustre-MDT0000 abort_recovery. [ 3414.495850] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3414.504644] Lustre: *** cfs_fail_loc=721, val=1*** [ 3414.506917] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3430.507714] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:14 [ 3444.704027] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3444.733466] Lustre: *** cfs_fail_loc=721, val=1*** [ 3446.878730] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:18 [ 3454.944066] Lustre: *** cfs_fail_loc=721, val=1*** [ 3454.951599] Lustre: Skipped 252 previous similar messages [ 3474.911887] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3474.917232] Lustre: *** cfs_fail_loc=721, val=1*** [ 3474.922322] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3474.928430] Lustre: 83319:0:(ldlm_lib.c:2085:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3474.946393] Lustre: 83319:0:(ldlm_lib.c:2085:extend_recovery_timer()) Skipped 20 previous similar messages [ 3493.993772] Lustre: lustre-MDT0000: Client 98986846-94b1-4b41-aaaa-108b03b874d0 (at 192.168.201.20@tcp) reconnected, waiting for 2 clients in recovery for 0:04 [ 3494.007403] Lustre: Skipped 2 previous similar messages [ 3505.123219] Lustre: lustre-MDT0000: Received new MDS connection from 0@lo, keep former export from same NID [ 3505.136341] Lustre: 83319:0:(ldlm_lib.c:2085:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3505.148246] Lustre: 83319:0:(ldlm_lib.c:2394:target_recovery_overseer()) lustre-MDT0000 recovery is aborted by hard timeout [ 3505.154462] Lustre: 83319:0:(ldlm_lib.c:2394:target_recovery_overseer()) Skipped 1 previous similar message [ 3505.170056] Lustre: 83319:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 3505.180321] Lustre: 83319:0:(ldlm_lib.c:2404:target_recovery_overseer()) Skipped 2 previous similar messages [ 3505.185176] Lustre: 83319:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp [ 3505.195467] Lustre: 83319:0:(genops.c:1622:class_disconnect_stale_exports()) Skipped 1 previous similar message [ 3505.208176] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3505.218836] LustreError: 83319:0:(ldlm_lib.c:1934:abort_lock_replay_queue()) @@@ aborted: req@ffff897412b5ad80 x1873883889702016/t0(0) o101->98986846-94b1-4b41-aaaa-108b03b874d0@192.168.201.20@tcp:0/0 lens 328/0 e 0 to 0 dl 1787078215 ref 1 fl Complete:/240/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 3505.253348] Lustre: lustre-MDT0000-osd: cancel update llog [0x20000a040:0x1:0x0] [ 3505.286397] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x2400090a3:0x1:0x0] [ 3505.313473] Lustre: 83319:0:(ldlm_lib.c:2951:target_recovery_thread()) too long recovery - read logs [ 3505.326446] LustreError: dumping log to /tmp/lustre-log.1787078343.83319 [ 3505.455561] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1185) [ 3505.462259] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1155 to 0x280000401:1185) [ 3506.143848] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 3506.148258] Lustre: Skipped 29 previous similar messages [ 3520.872575] Lustre: DEBUG MARKER: == replay-single test 130a: DoM file create (setstripe) replay ========================================================== 14:39:17 (1787078357) [ 3527.540780] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3529.274143] Lustre: Failing over lustre-MDT0000 [ 3529.565725] Lustre: server umount lustre-MDT0000 complete [ 3547.103394] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3547.123562] LustreError: Skipped 5 previous similar messages [ 3550.262947] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3550.268235] LDISKFS-fs (dm-0): recovery complete [ 3550.285630] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3560.948319] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3563.137578] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1217) [ 3563.141945] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1155 to 0x280000401:1217) [ 3569.909406] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3571.450966] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3579.088606] Lustre: DEBUG MARKER: == replay-single test 130b: DoM file create (inherited) replay ========================================================== 14:40:15 (1787078415) [ 3586.880203] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3588.657555] Lustre: Failing over lustre-MDT0000 [ 3588.900739] Lustre: server umount lustre-MDT0000 complete [ 3589.089752] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3589.106613] Lustre: Skipped 23 previous similar messages [ 3589.110043] LustreError: 83718:0:(ldlm_lib.c:1192:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3589.131409] LustreError: 83718:0:(ldlm_lib.c:1192:target_handle_connect()) Skipped 152 previous similar messages [ 3609.037784] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3609.041106] LDISKFS-fs (dm-0): recovery complete [ 3609.048981] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3619.298411] LustreError: 3638:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff8974138d9500 x1873883907468288/t0(0) o250->MGC192.168.201.120@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3623.280971] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3625.039603] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1249) [ 3625.043745] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1155 to 0x280000401:1249) [ 3631.581353] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3632.989777] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3640.396474] Lustre: DEBUG MARKER: == replay-single test 131a: DoM file write lock replay === 14:41:17 (1787078477) [ 3646.607948] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3648.333110] Lustre: Failing over lustre-MDT0000 [ 3648.584758] Lustre: server umount lustre-MDT0000 complete [ 3650.529710] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 3650.538843] LustreError: Skipped 3 previous similar messages [ 3668.952409] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3668.955278] LDISKFS-fs (dm-0): recovery complete [ 3668.961552] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3677.803319] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 3677.808216] Lustre: Skipped 5 previous similar messages [ 3681.118197] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3682.853836] Lustre: lustre-MDT0000: Recovery over after 0:05, of 2 clients 2 recovered and 0 were evicted. [ 3682.868143] Lustre: Skipped 5 previous similar messages [ 3682.907322] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1155 to 0x280000401:1281) [ 3682.908832] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:657 to 0x2c0000401:1281) [ 3689.008047] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3690.311736] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3697.689812] Lustre: DEBUG MARKER: SKIP: replay-single test_131b skipping excluded test 131b [ 3699.105408] Lustre: DEBUG MARKER: == replay-single test 132a: PFL new component instantiate replay ========================================================== 14:42:15 (1787078535) [ 3705.519344] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3707.709231] Lustre: Failing over lustre-MDT0000 [ 3707.958936] Lustre: server umount lustre-MDT0000 complete [ 3728.924072] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3728.932522] LDISKFS-fs (dm-0): recovery complete [ 3728.940994] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3733.997207] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0xd2c8700f42ed5666 [ 3734.117022] Lustre: lustre-MDT0000: Not available for connect from 192.168.201.20@tcp (not set up) [ 3737.547467] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3739.717302] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1313) [ 3739.722560] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1284 to 0x280000401:1313) [ 3745.113553] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3746.245748] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3753.224374] Lustre: DEBUG MARKER: == replay-single test 133: check resend of ongoing requests for lwp during failover ========================================================== 14:43:10 (1787078590) [ 3756.980438] Lustre: *** cfs_fail_loc=123, val=2147483648*** [ 3756.984806] Lustre: Skipped 247 previous similar messages [ 3759.604456] Lustre: Failing over lustre-MDT0000 [ 3759.853532] Lustre: server umount lustre-MDT0000 complete [ 3773.536841] Lustre: lustre-MDT0001: Client 0f7ac2a5-466f-4108-a0e6-15c5379c72ae (at 192.168.201.20@tcp) reconnecting [ 3777.205730] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3786.911368] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 3786.918341] Lustre: Skipped 8 previous similar messages [ 3789.742826] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3792.374495] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000300000400-0x0000000340000400]:1:mdt [ 3792.377205] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000300000400-0x0000000340000400]:1:mdt] [ 3792.411659] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1345) [ 3792.415492] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1284 to 0x280000401:1345) [ 3797.862952] Lustre: DEBUG MARKER: == replay-single test 134: replay creation of a file created in a pool ========================================================== 14:43:54 (1787078634) [ 3810.976617] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3812.609731] Lustre: Failing over lustre-MDT0000 [ 3812.816316] Lustre: server umount lustre-MDT0000 complete [ 3830.374653] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 3830.376202] LDISKFS-fs (dm-0): recovery complete [ 3830.382548] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3832.729094] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3835.926341] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1284 to 0x280000401:1377) [ 3835.928393] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1377) [ 3840.028380] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3841.281305] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3856.128111] Lustre: DEBUG MARKER: == replay-single test 135: Server failure in lock replay phase ========================================================== 14:44:53 (1787078693) [ 3863.653878] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 3865.277959] Lustre: Failing over lustre-OST0000 [ 3865.346658] Lustre: server umount lustre-OST0000 complete [ 3870.476766] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing load_module ../libcfs/libcfs/libcfs [ 3878.889476] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 3878.891128] LDISKFS-fs (dm-2): recovery complete [ 3878.898567] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 3878.977885] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 3878.982166] Lustre: Skipped 7 previous similar messages [ 3880.526902] Lustre: *** cfs_fail_loc=32d, val=20*** [ 3880.528981] Lustre: Skipped 1 previous similar message [ 3882.613114] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3887.599532] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount REPLAY_LOCKS osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid 1475 0 [ 3889.010036] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in REPLAY_LOCKS state after 0 sec [ 3890.248454] Lustre: Failing over lustre-OST0000 [ 3890.264777] LustreError: 97714:0:(ldlm_lib.c:3004:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 3890.272569] Lustre: 96933:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 3890.280268] LustreError: 96933:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 3890.434416] Lustre: server umount lustre-OST0000 complete [ 3904.599843] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing load_module ../libcfs/libcfs/libcfs [ 3908.366452] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 3911.891975] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3913.853960] LustreError: 98987:0:(ldlm_lib.c:3004:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 3913.858125] Lustre: 98449:0:(ldlm_lib.c:2404:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 3913.864205] LustreError: 98449:0:(ldlm_lib.c:1934:abort_lock_replay_queue()) @@@ aborted: req@ffff897424fb7800 x1873883889883776/t0(0) o101->0f7ac2a5-466f-4108-a0e6-15c5379c72ae@192.168.201.20@tcp:554/0 lens 328/0 e 0 to 0 dl 1787078759 ref 1 fl Complete:/240/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 3913.874224] LustreError: 98449:0:(ldlm_lib.c:1934:abort_lock_replay_queue()) Skipped 17 previous similar messages [ 3913.877594] LustreError: 98449:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 3913.882208] Lustre: lustre-OST0000: Not available for connect from 192.168.201.20@tcp (stopping) [ 3913.953822] Lustre: server umount lustre-OST0000 complete [ 3930.079114] Lustre: lustre-OST0001 is waiting for obd_unlinked_exports more than 8 seconds. The obd refcount = 3. Is it stuck? [ 3930.150245] Lustre: server umount lustre-OST0001 complete [ 3934.544920] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 3937.589662] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3942.364260] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 3945.574148] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3957.795598] LustreError: lustre-OST0000-osc-MDT0001: This client was evicted by lustre-OST0000; in progress operations using this service will fail. [ 3957.801048] LustreError: lustre-OST0001-osc-MDT0001: This client was evicted by lustre-OST0001; in progress operations using this service will fail. [ 3957.815555] LustreError: lustre-OST0000-osc-MDT0000: This client was evicted by lustre-OST0000; in progress operations using this service will fail. [ 3957.822485] LustreError: lustre-OST0001-osc-MDT0000: This client was evicted by lustre-OST0001; in progress operations using this service will fail. [ 3960.932910] Lustre: DEBUG MARKER: == replay-single test 136: MDS to disconnect all OSPs first, then cleanup ldlm ========================================================== 14:46:38 (1787078798) [ 3961.857737] Lustre: DEBUG MARKER: SKIP: replay-single test_136 needs > 2 MDTs [ 3962.730390] Lustre: DEBUG MARKER: == replay-single test 137a: DNE: create under striped dir, fail MDT1 ========================================================== 14:46:40 (1787078800) [ 3966.722129] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3967.756996] Lustre: Failing over lustre-MDT0000 [ 3968.016677] Lustre: server umount lustre-MDT0000 complete [ 3984.278113] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3984.280080] LDISKFS-fs (dm-0): recovery complete [ 3984.285420] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3986.848137] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3989.557969] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1409) [ 3989.558965] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1399 to 0x280000401:1441) [ 3992.466895] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 3993.365659] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3999.275648] Lustre: DEBUG MARKER: == replay-single test 137b: DNE: create under striped dir, fail MDT2 ========================================================== 14:47:16 (1787078836) [ 4003.464359] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4004.579390] Lustre: Failing over lustre-MDT0001 [ 4004.724127] Lustre: server umount lustre-MDT0001 complete [ 4020.934564] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4020.935801] LDISKFS-fs (dm-1): recovery complete [ 4020.941766] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4023.848473] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4026.409250] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:1093 to 0x2c0000400:1121) [ 4026.409256] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1121) [ 4029.789909] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4030.552984] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4035.772047] Lustre: DEBUG MARKER: == replay-single test 137c: DNE: create under striped dir, fail MDT1/MDT2 ========================================================== 14:47:53 (1787078873) [ 4039.440475] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4043.841940] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4045.226836] Lustre: Failing over lustre-MDT0001 [ 4045.413993] Lustre: server umount lustre-MDT0001 complete [ 4047.536577] Lustre: Failing over lustre-MDT0000 [ 4047.752453] Lustre: server umount lustre-MDT0000 complete [ 4064.510433] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4064.512332] LDISKFS-fs (dm-0): recovery complete [ 4064.522280] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4064.528575] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4064.531580] LDISKFS-fs (dm-1): recovery complete [ 4064.549278] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4067.306855] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4067.401575] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4067.977520] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1787078890/real 1787078890] req@ffff897309021c00 x1873883907788160/t0(0) o400->lustre-MDT0000-lwp-OST0000@0@lo:12/10 lens 224/224 e 0 to 1 dl 1787078906 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4067.988767] Lustre: 3641:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 12 previous similar messages [ 4067.992711] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:1093 to 0x2c0000400:1153) [ 4067.995136] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1153) [ 4068.036972] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1441) [ 4068.037524] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1399 to 0x280000401:1473) [ 4072.795844] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid 1475 0 [ 4073.680549] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4074.681644] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4080.188909] Lustre: DEBUG MARKER: == replay-single test 200: Dropping one OBD_PING should not cause disconnect ========================================================== 14:48:37 (1787078917) [ 4081.095028] Lustre: DEBUG MARKER: SKIP: replay-single test_200 Need remote client [ 4082.143629] Lustre: DEBUG MARKER: == replay-single test 201: MDT umount cascading disconnects timeouts ========================================================== 14:48:39 (1787078919) [ 4084.775276] LustreError: 107302:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 sleeping for 8000ms [ 4092.791129] LustreError: 107302:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 awake [ 4092.796354] Lustre: Failing over lustre-MDT0001 [ 4092.804122] LustreError: 99654:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 sleeping for 8000ms [ 4092.809947] LustreError: 99654:0:(tgt_handler.c:1143:tgt_disconnect()) Skipped 2 previous similar messages [ 4092.896111] Lustre: lustre-MDT0001: Not available for connect from 0@lo (stopping) [ 4092.900616] Lustre: Skipped 10 previous similar messages [ 4098.655436] LustreError: 99646:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 sleeping for 8000ms [ 4098.659641] LustreError: 99646:0:(tgt_handler.c:1143:tgt_disconnect()) Skipped 1 previous similar message [ 4100.815116] LustreError: 100944:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 awake [ 4100.857862] Lustre: server umount lustre-MDT0001 complete [ 4105.921052] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4106.703114] LustreError: 99646:0:(tgt_handler.c:1143:tgt_disconnect()) cfs_fail_timeout id 245 awake [ 4106.711052] LustreError: 99646:0:(tgt_handler.c:1143:tgt_disconnect()) Skipped 2 previous similar messages [ 4108.666255] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4111.334736] Lustre: lustre-MDT0001-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 4111.338205] Lustre: Skipped 42 previous similar messages [ 4111.372847] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:669 to 0x280000400:1185) [ 4111.372847] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:1093 to 0x2c0000400:1185) [ 4114.024695] Lustre: DEBUG MARKER: == replay-single test 202: pfl replay should recovery layout generation ========================================================== 14:49:11 (1787078951) [ 4118.540131] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4119.626340] Lustre: Failing over lustre-MDT0000 [ 4119.829496] Lustre: server umount lustre-MDT0000 complete [ 4136.736376] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4136.738767] LDISKFS-fs (dm-0): recovery complete [ 4136.747567] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4139.028619] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 4142.152738] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1283 to 0x2c0000401:1473) [ 4142.158615] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1475 to 0x280000401:1505) [ 4145.682410] Lustre: DEBUG MARKER: oleg120-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid 1475 0 [ 4146.731872] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4152.386880] Lustre: DEBUG MARKER: == replay-single test 203: resend can hit original request ========================================================== 14:49:49 (1787078989) [ 4153.096831] LustreError: 108823:0:(mdt_handler.c:2174:mdt_getattr_name_lock()) cfs_fail_timeout id 2403 sleeping for 2000ms [ 4155.183129] LustreError: 108823:0:(mdt_handler.c:2174:mdt_getattr_name_lock()) cfs_fail_timeout id 2403 awake [ 4155.188773] LustreError: 108823:0:(mdt_handler.c:2174:mdt_getattr_name_lock()) Skipped 1 previous similar message [ 4155.195505] Lustre: 108823:0:(service.c:2628:ptlrpc_server_handle_request()) @@@ pause req after reply req@ffff8974259ace00 x1873883889986816/t0(0) o101->0f7ac2a5-466f-4108-a0e6-15c5379c72ae@192.168.201.20@tcp:81/0 lens 592/1888 e 0 to 0 dl 1787079041 ref 1 fl Complete:/600/0 rc 0/0 job:'stat.0' uid:0 gid:0 projid:0 [ 4158.241093] Lustre: 108823:0:(service.c:2630:ptlrpc_server_handle_request()) @@@ continue req@ffff8974259ace00 x1873883889986816/t0(0) o101->0f7ac2a5-466f-4108-a0e6-15c5379c72ae@192.168.201.20@tcp:81/0 lens 592/1888 e 0 to 0 dl 1787079041 ref 1 fl Complete:/600/0 rc 0/0 job:'stat.0' uid:0 gid:0 projid:0 [ 4160.264233] Lustre: DEBUG MARKER: == replay-single test complete, duration 3836 sec ======== 14:49:57 (1787078997) [ 4161.048718] Lustre: DEBUG MARKER: === replay-single: start cleanup 14:49:58 (1787078998) === [ 4165.378078] Lustre: DEBUG MARKER: === replay-single: finish cleanup 14:50:02 (1787079002) === [ 4167.303318] LustreError: 6501:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787079005 with bad export cookie 15188512954156686396 [ 4167.313180] LustreError: 6501:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 4167.320544] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4167.323590] Lustre: Skipped 5 previous similar messages [ 4173.493563] Lustre: server umount lustre-MDT0000 complete [ 4178.472100] LustreError: 26199:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1787079016 with bad export cookie 15188512954156686291 [ 4178.475618] LustreError: MGC192.168.201.120@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4178.478985] LustreError: 26199:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 4178.494874] LustreError: Skipped 8 previous similar messages [ 4178.677040] Lustre: server umount lustre-MDT0001 complete [ 4193.100596] Lustre: server umount lustre-OST0000 complete [ 4207.858292] Lustre: server umount lustre-OST0001 complete [ 4216.259136] Lustre: DEBUG MARKER: oleg120-server.virtnet: executing unload_modules_local [ 4217.680214] Key type lgssc unregistered [ 4217.838280] LNet: 114558:0:(lib-ptl.c:964:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 4217.843387] LNetError: 114558:0:(acceptor.c:252:lnet_acceptor_remove_socket()) Interface ens2 not found [ 4217.856522] LNet: Removed LNI 192.168.201.120@tcp [ 4218.394261] Key type .llcrypt unregistered [ 4218.395685] Key type ._llcrypt unregistered