[ 0.000000] Linux version 4.18.0rh8.10-debug (green@maintenance) (gcc version 8.5.0 20210514 (Red Hat 8.5.0-26) (GCC)) #2 SMP Mon Jul 14 01:24:22 EDT 2025 [ 0.000000] Command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] x86/fpu: Supporting XSAVE feature 0x001: 'x87 floating point registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x002: 'SSE registers' [ 0.000000] x86/fpu: Supporting XSAVE feature 0x004: 'AVX registers' [ 0.000000] x86/fpu: xstate_offset[2]: 576, xstate_sizes[2]: 256 [ 0.000000] x86/fpu: Enabled xstate features 0x7, context size is 832 bytes, using 'standard' format. [ 0.000000] signal: max sigframe size: 1776 [ 0.000000] BIOS-provided physical RAM map: [ 0.000000] BIOS-e820: [mem 0x0000000000000000-0x000000000009fbff] usable [ 0.000000] BIOS-e820: [mem 0x000000000009fc00-0x000000000009ffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000000f0000-0x00000000000fffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000000100000-0x00000000bffcdfff] usable [ 0.000000] BIOS-e820: [mem 0x00000000bffce000-0x00000000bfffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000feffc000-0x00000000feffffff] reserved [ 0.000000] BIOS-e820: [mem 0x00000000fffc0000-0x00000000ffffffff] reserved [ 0.000000] BIOS-e820: [mem 0x0000000100000000-0x0000000146dfffff] usable [ 0.000000] NX (Execute Disable) protection: active [ 0.000000] SMBIOS 2.8 present. [ 0.000000] DMI: QEMU Standard PC (i440FX + PIIX, 1996), BIOS 1.17.0-10.fc44 06/10/2025 [ 0.000000] Hypervisor detected: KVM [ 0.000000] kvm-clock: Using msrs 4b564d01 and 4b564d00 [ 0.000000] kvm-clock: using sched offset of 702173940 cycles [ 0.000000] clocksource: kvm-clock: mask: 0xffffffffffffffff max_cycles: 0x1cd42e4dffb, max_idle_ns: 881590591483 ns [ 0.000000] tsc: Detected 2400.000 MHz processor [ 0.000000] last_pfn = 0x146e00 max_arch_pfn = 0x400000000 [ 0.000000] x86/PAT: Configuration [0-7]: WB WC UC- UC WB WP UC- WT [ 0.000000] last_pfn = 0xbffce max_arch_pfn = 0x400000000 [ 0.000000] found SMP MP-table at [mem 0x000f54b0-0x000f54bf] [ 0.000000] RAMDISK: [mem 0xbcc54000-0xbffbffff] [ 0.000000] ACPI: Early table checksum verification disabled [ 0.000000] ACPI: RSDP 0x00000000000F52D0 000014 (v00 BOCHS ) [ 0.000000] ACPI: RSDT 0x00000000BFFE247C 000034 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACP 0x00000000BFFE2318 000074 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: DSDT 0x00000000BFFE0040 0022D8 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: FACS 0x00000000BFFE0000 000040 [ 0.000000] ACPI: APIC 0x00000000BFFE238C 000090 (v03 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: HPET 0x00000000BFFE241C 000038 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: WAET 0x00000000BFFE2454 000028 (v01 BOCHS BXPC 00000001 BXPC 00000001) [ 0.000000] ACPI: Reserving FACP table memory at [mem 0xbffe2318-0xbffe238b] [ 0.000000] ACPI: Reserving DSDT table memory at [mem 0xbffe0040-0xbffe2317] [ 0.000000] ACPI: Reserving FACS table memory at [mem 0xbffe0000-0xbffe003f] [ 0.000000] ACPI: Reserving APIC table memory at [mem 0xbffe238c-0xbffe241b] [ 0.000000] ACPI: Reserving HPET table memory at [mem 0xbffe241c-0xbffe2453] [ 0.000000] ACPI: Reserving WAET table memory at [mem 0xbffe2454-0xbffe247b] [ 0.000000] No NUMA configuration found [ 0.000000] Faking a node at [mem 0x0000000000000000-0x0000000146dfffff] [ 0.000000] NODE_DATA(0) allocated [mem 0x1465a3000-0x1465cdfff] [ 0.000000] Reserving 256MB of memory at 2752MB for crashkernel (System RAM: 4205MB) [ 0.000000] Zone ranges: [ 0.000000] DMA [mem 0x0000000000001000-0x0000000000ffffff] [ 0.000000] DMA32 [mem 0x0000000001000000-0x00000000ffffffff] [ 0.000000] Normal [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Device empty [ 0.000000] Movable zone start for each node [ 0.000000] Early memory node ranges [ 0.000000] node 0: [mem 0x0000000000001000-0x000000000009efff] [ 0.000000] node 0: [mem 0x0000000000100000-0x00000000bffcdfff] [ 0.000000] node 0: [mem 0x0000000100000000-0x0000000146dfffff] [ 0.000000] Zeroed struct page in unavailable ranges: 4756 pages [ 0.000000] Initmem setup node 0 [mem 0x0000000000001000-0x0000000146dfffff] [ 0.000000] ACPI: PM-Timer IO Port: 0x608 [ 0.000000] ACPI: LAPIC_NMI (acpi_id[0xff] dfl dfl lint[0x1]) [ 0.000000] IOAPIC[0]: apic_id 0, version 17, address 0xfec00000, GSI 0-23 [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 0 global_irq 2 dfl dfl) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 5 global_irq 5 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 9 global_irq 9 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 10 global_irq 10 high level) [ 0.000000] ACPI: INT_SRC_OVR (bus 0 bus_irq 11 global_irq 11 high level) [ 0.000000] Using ACPI (MADT) for SMP configuration information [ 0.000000] ACPI: HPET id: 0x8086a201 base: 0xfed00000 [ 0.000000] TSC deadline timer available [ 0.000000] smpboot: Allowing 4 CPUs, 0 hotplug CPUs [ 0.000000] kvm-guest: KVM setup pv remote TLB flush [ 0.000000] kvm-guest: setup PV sched yield [ 0.000000] PM: Registered nosave memory: [mem 0x00000000-0x00000fff] [ 0.000000] PM: Registered nosave memory: [mem 0x0009f000-0x0009ffff] [ 0.000000] PM: Registered nosave memory: [mem 0x000a0000-0x000effff] [ 0.000000] PM: Registered nosave memory: [mem 0x000f0000-0x000fffff] [ 0.000000] PM: Registered nosave memory: [mem 0xbffce000-0xbfffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xc0000000-0xfeffbfff] [ 0.000000] PM: Registered nosave memory: [mem 0xfeffc000-0xfeffffff] [ 0.000000] PM: Registered nosave memory: [mem 0xff000000-0xfffbffff] [ 0.000000] PM: Registered nosave memory: [mem 0xfffc0000-0xffffffff] [ 0.000000] [mem 0xc0000000-0xfeffbfff] available for PCI devices [ 0.000000] Booting paravirtualized kernel on KVM [ 0.000000] clocksource: refined-jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1910969940391419 ns [ 0.000000] setup_percpu: NR_CPUS:8192 nr_cpumask_bits:4 nr_cpu_ids:4 nr_node_ids:1 [ 0.000000] percpu: Embedded 63 pages/cpu s221184 r8192 d28672 u524288 [ 0.000000] kvm-guest: PV spinlocks enabled [ 0.000000] PV qspinlock hash table entries: 256 (order: 0, 4096 bytes, linear) [ 0.000000] Built 1 zonelists, mobility grouping on. Total pages: 1059606 [ 0.000000] Policy zone: Normal [ 0.000000] Kernel command line: rd.shell root=nbd:192.168.200.253:rocky8.10:ext4:ro:-p,-b4096 ro crashkernel=256M panic=1 nomodeset ipmtu=9000 ip=dhcp rd.neednet=1 init_on_free=off mitigations=off console=ttyS1,115200 audit=0 [ 0.000000] Specific versions of hardware are certified with Red Hat Enterprise Linux 8. Please see the list of hardware certified with Red Hat Enterprise Linux 8 at https://catalog.redhat.com. [ 0.000000] audit: disabled (until reboot) [ 0.000000] software IO TLB: area num 4. [ 0.000000] Memory: 2829652K/4306352K available (18435K kernel code, 11221K rwdata, 7248K rodata, 2908K init, 18040K bss, 524584K reserved, 0K cma-reserved) [ 0.000000] SLUB: HWalign=64, Order=0-3, MinObjects=0, CPUs=4, Nodes=1 [ 0.000000] kmemleak: Kernel memory leak detector disabled [ 0.000000] ftrace: allocating 41240 entries in 162 pages [ 0.000000] ftrace: allocated 162 pages with 3 groups [ 0.000000] rcu: Hierarchical RCU implementation. [ 0.000000] rcu: RCU event tracing is enabled. [ 0.000000] rcu: RCU restricting CPUs from NR_CPUS=8192 to nr_cpu_ids=4. [ 0.000000] rcu: RCU callback double-/use-after-free debug enabled. [ 0.000000] Rude variant of Tasks RCU enabled. [ 0.000000] Tracing variant of Tasks RCU enabled. [ 0.000000] rcu: RCU calculated value of scheduler-enlistment delay is 100 jiffies. [ 0.000000] rcu: Adjusting geometry for rcu_fanout_leaf=16, nr_cpu_ids=4 [ 0.000000] NR_IRQS: 524544, nr_irqs: 456, preallocated irqs: 16 [ 0.000000] random: get_random_bytes called from start_kernel+0x622/0x9a8 with crng_init=0 [ 0.001000] Console: colour *CGA 80x25 [ 0.001000] printk: console [ttyS1] enabled [ 0.001000] ACPI: Core revision 20220331 [ 0.001000] clocksource: hpet: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 19112604467 ns [ 0.001010] APIC: Switch to symmetric I/O mode setup [ 0.003331] x2apic enabled [ 0.004012] Switched APIC routing to physical x2apic. [ 0.006011] kvm-guest: setup PV IPIs [ 0.010670] ..TIMER: vector=0x30 apic1=0 pin1=2 apic2=-1 pin2=-1 [ 0.011000] clocksource: tsc-early: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 0.011023] Calibrating delay loop (skipped) preset value.. 4800.00 BogoMIPS (lpj=2400000) [ 0.012015] pid_max: default: 32768 minimum: 301 [ 0.014055] LSM: Security Framework initializing [ 0.015232] Yama: becoming mindful. [ 0.016056] SELinux: Initializing. [ 0.017306] *** VALIDATE selinux *** [ 0.028336] Dentry cache hash table entries: 1048576 (order: 11, 8388608 bytes, vmalloc) [ 0.036847] Inode-cache hash table entries: 524288 (order: 10, 4194304 bytes, vmalloc) [ 0.037176] Mount-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.038134] Mountpoint-cache hash table entries: 16384 (order: 5, 131072 bytes, vmalloc) [ 0.039123] *** VALIDATE tmpfs *** [ 0.040581] *** VALIDATE proc *** [ 0.042234] *** VALIDATE cgroup *** [ 0.043014] *** VALIDATE cgroup2 *** [ 0.044358] x86/cpu: User Mode Instruction Prevention (UMIP) activated [ 0.046101] Last level iTLB entries: 4KB 0, 2MB 0, 4MB 0 [ 0.047008] Last level dTLB entries: 4KB 0, 2MB 0, 4MB 0, 1GB 0 [ 0.049023] Spectre V2 : User space: Vulnerable [ 0.050009] Speculative Store Bypass: Vulnerable [ 0.053764] debug: unmapping init [mem 0xffffffffafe59000-0xffffffffafe60fff] [ 0.056000] smpboot: CPU0: Intel(R) Xeon(R) CPU E5-2695 v2 @ 2.40GHz (family: 0x6, model: 0x3e, stepping: 0x4) [ 0.056713] Performance Events: IvyBridge events, full-width counters, Intel PMU driver. [ 0.057022] ... version: 2 [ 0.058012] ... bit width: 48 [ 0.059010] ... generic registers: 4 [ 0.060012] ... value mask: 0000ffffffffffff [ 0.061014] ... max period: 00007fffffffffff [ 0.062016] ... fixed-purpose events: 3 [ 0.063014] ... event mask: 000000070000000f [ 0.064375] rcu: Hierarchical SRCU implementation. [ 0.066653] smp: Bringing up secondary CPUs ... [ 0.067595] x86: Booting SMP configuration: [ 0.068030] .... node #0, CPUs: #1 #2 #3 [ 0.079223] smp: Brought up 1 node, 4 CPUs [ 0.081024] smpboot: Max logical packages: 1 [ 0.082014] smpboot: Total of 4 processors activated (19200.00 BogoMIPS) [ 0.122419] node 0 deferred pages initialised in 38ms [ 0.126112] devtmpfs: initialized [ 0.127400] x86/mm: Memory block size: 128MB [ 0.132707] gcov: version magic: 0x41383552 [ 0.139299] clocksource: jiffies: mask: 0xffffffff max_cycles: 0xffffffff, max_idle_ns: 1911260446275000 ns [ 0.140084] futex hash table entries: 1024 (order: 4, 65536 bytes, vmalloc) [ 0.141365] pinctrl core: initialized pinctrl subsystem [ 0.142298] [ 0.143000] ************************************************************* [ 0.143022] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.146015] ** ** [ 0.152015] ** IOMMU DebugFS SUPPORT HAS BEEN ENABLED IN THIS KERNEL ** [ 0.157013] ** ** [ 0.161014] ** This means that this kernel is built to expose internal ** [ 0.167014] ** IOMMU data structures, which may compromise security on ** [ 0.171015] ** your system. ** [ 0.175019] ** ** [ 0.179017] ** If you see this message and you are not debugging the ** [ 0.182016] ** kernel, report this immediately to your vendor! ** [ 0.186016] ** ** [ 0.189014] ** NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE NOTICE ** [ 0.193014] ************************************************************* [ 0.197058] NET: Registered protocol family 16 [ 0.200527] DMA: preallocated 512 KiB GFP_KERNEL pool for atomic allocations [ 0.205141] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA pool for atomic allocations [ 0.208078] DMA: preallocated 512 KiB GFP_KERNEL|GFP_DMA32 pool for atomic allocations [ 0.213023] cpuidle: using governor menu [ 0.215764] acpiphp: ACPI Hot Plug PCI Controller Driver version: 0.5 [ 0.220085] PCI: Using configuration type 1 for base access [ 0.222124] core: PMU erratum BJ122, BV98, HSD29 worked around, HT is on [ 0.231232] HugeTLB registered 1.00 GiB page size, pre-allocated 0 pages [ 0.232000] HugeTLB registered 2.00 MiB page size, pre-allocated 0 pages [ 0.234707] cryptd: max_cpu_qlen set to 1000 [ 0.235425] ACPI: Added _OSI(Module Device) [ 0.236016] ACPI: Added _OSI(Processor Device) [ 0.237000] ACPI: Added _OSI(3.0 _SCP Extensions) [ 0.240040] ACPI: Added _OSI(Processor Aggregator Device) [ 0.246587] ACPI: 1 ACPI AML tables successfully acquired and loaded [ 0.256181] ACPI: Interpreter enabled [ 0.258135] ACPI: PM: (supports S0 S3 S4 S5) [ 0.260016] ACPI: Using IOAPIC for interrupt routing [ 0.264486] PCI: Using host bridge windows from ACPI; if necessary, use "pci=nocrs" and report a bug [ 0.270633] ACPI: Enabled 2 GPEs in block 00 to 0F [ 0.284103] ACPI: PCI Root Bridge [PCI0] (domain 0000 [bus 00-ff]) [ 0.287114] acpi PNP0A03:00: _OSC: OS supports [ASPM ClockPM Segments MSI HPX-Type3] [ 0.288023] acpi PNP0A03:00: _OSC: not requesting OS control; OS requires [ExtendedConfig ASPM ClockPM MSI] [ 0.293075] acpi PNP0A03:00: fail to add MMCONFIG information, can't access extended PCI configuration space under this bridge. [ 0.299966] acpiphp: Slot [2] registered [ 0.300382] acpiphp: Slot [5] registered [ 0.303198] acpiphp: Slot [6] registered [ 0.306167] acpiphp: Slot [7] registered [ 0.308193] acpiphp: Slot [8] registered [ 0.309415] acpiphp: Slot [9] registered [ 0.311129] acpiphp: Slot [10] registered [ 0.312127] acpiphp: Slot [3] registered [ 0.313094] acpiphp: Slot [4] registered [ 0.315116] acpiphp: Slot [11] registered [ 0.316290] acpiphp: Slot [12] registered [ 0.318329] acpiphp: Slot [13] registered [ 0.320232] acpiphp: Slot [14] registered [ 0.322111] acpiphp: Slot [15] registered [ 0.323103] acpiphp: Slot [16] registered [ 0.324117] acpiphp: Slot [17] registered [ 0.325340] acpiphp: Slot [18] registered [ 0.327331] acpiphp: Slot [19] registered [ 0.329110] acpiphp: Slot [20] registered [ 0.330110] acpiphp: Slot [21] registered [ 0.332124] acpiphp: Slot [22] registered [ 0.334137] acpiphp: Slot [23] registered [ 0.336143] acpiphp: Slot [24] registered [ 0.337358] acpiphp: Slot [25] registered [ 0.339118] acpiphp: Slot [26] registered [ 0.341154] acpiphp: Slot [27] registered [ 0.342088] acpiphp: Slot [28] registered [ 0.343080] acpiphp: Slot [29] registered [ 0.345109] acpiphp: Slot [30] registered [ 0.347110] acpiphp: Slot [31] registered [ 0.349082] PCI host bridge to bus 0000:00 [ 0.350019] pci_bus 0000:00: root bus resource [io 0x0000-0x0cf7 window] [ 0.353026] pci_bus 0000:00: root bus resource [io 0x0d00-0xffff window] [ 0.356024] pci_bus 0000:00: root bus resource [mem 0x000a0000-0x000bffff window] [ 0.359026] pci_bus 0000:00: root bus resource [mem 0xc0000000-0xfebfffff window] [ 0.362029] pci_bus 0000:00: root bus resource [mem 0xe0000000000-0xe007fffffff window] [ 0.365028] pci_bus 0000:00: root bus resource [bus 00-ff] [ 0.367187] pci 0000:00:00.0: [8086:1237] type 00 class 0x060000 [ 0.370181] pci 0000:00:01.0: [8086:7000] type 00 class 0x060100 [ 0.373281] pci 0000:00:01.1: [8086:7010] type 00 class 0x010180 [ 0.387026] pci 0000:00:01.1: reg 0x20: [io 0xc320-0xc32f] [ 0.393057] pci 0000:00:01.1: legacy IDE quirk: reg 0x10: [io 0x01f0-0x01f7] [ 0.396092] pci 0000:00:01.1: legacy IDE quirk: reg 0x14: [io 0x03f6] [ 0.400022] pci 0000:00:01.1: legacy IDE quirk: reg 0x18: [io 0x0170-0x0177] [ 0.404018] pci 0000:00:01.1: legacy IDE quirk: reg 0x1c: [io 0x0376] [ 0.407935] pci 0000:00:01.3: [8086:7113] type 00 class 0x068000 [ 0.411157] pci 0000:00:01.3: quirk: [io 0x0600-0x063f] claimed by PIIX4 ACPI [ 0.414045] pci 0000:00:01.3: quirk: [io 0x0700-0x070f] claimed by PIIX4 SMB [ 0.416889] pci 0000:00:02.0: [1af4:1000] type 00 class 0x020000 [ 0.423014] pci 0000:00:02.0: reg 0x10: [io 0xc300-0xc31f] [ 0.437014] pci 0000:00:02.0: reg 0x20: [mem 0xe0000000000-0xe0000003fff 64bit pref] [ 0.442014] pci 0000:00:02.0: reg 0x30: [mem 0xfeb80000-0xfebbffff pref] [ 0.451750] pci 0000:00:05.0: [1af4:1001] type 00 class 0x010000 [ 0.464016] pci 0000:00:05.0: reg 0x10: [io 0xc000-0xc07f] [ 0.479028] pci 0000:00:05.0: reg 0x14: [mem 0xfebc0000-0xfebc0fff] [ 0.513016] pci 0000:00:05.0: reg 0x20: [mem 0xe0000004000-0xe0000007fff 64bit pref] [ 0.526890] pci 0000:00:06.0: [1af4:1001] type 00 class 0x010000 [ 0.536020] pci 0000:00:06.0: reg 0x10: [io 0xc080-0xc0ff] [ 0.547322] pci 0000:00:06.0: reg 0x14: [mem 0xfebc1000-0xfebc1fff] [ 0.580098] pci 0000:00:06.0: reg 0x20: [mem 0xe0000008000-0xe000000bfff 64bit pref] [ 0.598000] pci 0000:00:07.0: [1af4:1001] type 00 class 0x010000 [ 0.610017] pci 0000:00:07.0: reg 0x10: [io 0xc100-0xc17f] [ 0.627026] pci 0000:00:07.0: reg 0x14: [mem 0xfebc2000-0xfebc2fff] [ 0.653014] pci 0000:00:07.0: reg 0x20: [mem 0xe000000c000-0xe000000ffff 64bit pref] [ 0.676043] pci 0000:00:08.0: [1af4:1001] type 00 class 0x010000 [ 0.687017] pci 0000:00:08.0: reg 0x10: [io 0xc180-0xc1ff] [ 0.697153] pci 0000:00:08.0: reg 0x14: [mem 0xfebc3000-0xfebc3fff] [ 0.716020] pci 0000:00:08.0: reg 0x20: [mem 0xe0000010000-0xe0000013fff 64bit pref] [ 0.733417] pci 0000:00:09.0: [1af4:1001] type 00 class 0x010000 [ 0.744019] pci 0000:00:09.0: reg 0x10: [io 0xc200-0xc27f] [ 0.757025] pci 0000:00:09.0: reg 0x14: [mem 0xfebc4000-0xfebc4fff] [ 0.775021] pci 0000:00:09.0: reg 0x20: [mem 0xe0000014000-0xe0000017fff 64bit pref] [ 0.798052] pci 0000:00:0a.0: [1af4:1001] type 00 class 0x010000 [ 0.814018] pci 0000:00:0a.0: reg 0x10: [io 0xc280-0xc2ff] [ 0.826017] pci 0000:00:0a.0: reg 0x14: [mem 0xfebc5000-0xfebc5fff] [ 0.864017] pci 0000:00:0a.0: reg 0x20: [mem 0xe0000018000-0xe000001bfff 64bit pref] [ 0.884511] ACPI: PCI: Interrupt link LNKA configured for IRQ 10 [ 0.887392] ACPI: PCI: Interrupt link LNKB configured for IRQ 10 [ 0.889401] ACPI: PCI: Interrupt link LNKC configured for IRQ 11 [ 0.892608] ACPI: PCI: Interrupt link LNKD configured for IRQ 11 [ 0.894251] ACPI: PCI: Interrupt link LNKS configured for IRQ 9 [ 0.902069] iommu: Default domain type: Passthrough [ 0.904533] SCSI subsystem initialized [ 0.906234] ACPI: bus type USB registered [ 0.908135] usbcore: registered new interface driver usbfs [ 0.910116] usbcore: registered new interface driver hub [ 0.911084] usbcore: registered new device driver usb [ 0.913157] pps_core: LinuxPPS API ver. 1 registered [ 0.915012] pps_core: Software ver. 5.3.6 - Copyright 2005-2007 Rodolfo Giometti [ 0.918066] PTP clock support registered [ 0.920145] EDAC MC: Ver: 3.0.0 [ 0.923164] PCI: Using ACPI for IRQ routing [ 0.924000] NetLabel: Initializing [ 0.926011] NetLabel: domain hash size = 128 [ 0.928022] NetLabel: protocols = UNLABELED CIPSOv4 CALIPSO [ 0.931080] NetLabel: unlabeled traffic allowed by default [ 0.933414] vgaarb: loaded [ 0.935039] hpet0: at MMIO 0xfed00000, IRQs 2, 8, 0 [ 0.937013] hpet0: 3 comparators, 64-bit 100.000000 MHz counter [ 0.943748] clocksource: Switched to clocksource kvm-clock [ 1.138963] VFS: Disk quotas dquot_6.6.0 [ 1.144368] VFS: Dquot-cache hash table entries: 512 (order 0, 4096 bytes) [ 1.153177] *** VALIDATE ramfs *** [ 1.158248] *** VALIDATE hugetlbfs *** [ 1.164592] pnp: PnP ACPI init [ 1.171927] pnp: PnP ACPI: found 6 devices [ 1.226384] clocksource: acpi_pm: mask: 0xffffff max_cycles: 0xffffff, max_idle_ns: 2085701024 ns [ 1.233510] pci_bus 0000:00: resource 4 [io 0x0000-0x0cf7 window] [ 1.242245] pci_bus 0000:00: resource 5 [io 0x0d00-0xffff window] [ 1.246820] pci_bus 0000:00: resource 6 [mem 0x000a0000-0x000bffff window] [ 1.253329] pci_bus 0000:00: resource 7 [mem 0xc0000000-0xfebfffff window] [ 1.260496] pci_bus 0000:00: resource 8 [mem 0xe0000000000-0xe007fffffff window] [ 1.267562] NET: Registered protocol family 2 [ 1.277531] IP idents hash table entries: 131072 (order: 8, 1048576 bytes, vmalloc) [ 1.295826] tcp_listen_portaddr_hash hash table entries: 4096 (order: 5, 163840 bytes, vmalloc) [ 1.305050] TCP established hash table entries: 65536 (order: 7, 524288 bytes, vmalloc) [ 1.320147] TCP bind hash table entries: 65536 (order: 9, 2097152 bytes, vmalloc) [ 1.328514] TCP: Hash tables configured (established 65536 bind 65536) [ 1.337450] MPTCP token hash table entries: 8192 (order: 6, 393216 bytes, vmalloc) [ 1.346095] UDP hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.357394] UDP-Lite hash table entries: 4096 (order: 6, 393216 bytes, vmalloc) [ 1.361867] NET: Registered protocol family 1 [ 1.364913] RPC: Registered named UNIX socket transport module. [ 1.367447] RPC: Registered udp transport module. [ 1.369435] RPC: Registered tcp transport module. [ 1.371899] RPC: Registered tcp NFSv4.1 backchannel transport module. [ 1.374734] NET: Registered protocol family 44 [ 1.376630] pci 0000:00:00.0: Limiting direct PCI/PCI transfers [ 1.378669] pci 0000:00:01.0: PIIX3: Enabling Passive Release [ 1.381609] pci 0000:00:01.0: Activating ISA DMA hang workarounds [ 1.385566] PCI: CLS 0 bytes, default 64 [ 1.390326] Unpacking initramfs... [ 4.132136] debug: unmapping init [mem 0xffff97be7cc54000-0xffff97be7ffbffff] [ 4.136442] PCI-DMA: Using software bounce buffering for IO (SWIOTLB) [ 4.139170] software IO TLB: mapped [mem 0x00000000a8000000-0x00000000ac000000] (64MB) [ 4.142408] clocksource: tsc: mask: 0xffffffffffffffff max_cycles: 0x22983777dd9, max_idle_ns: 440795300422 ns [ 5.100121] Initialise system trusted keyrings [ 5.102219] Key type blacklist registered [ 5.104726] workingset: timestamp_bits=36 max_order=20 bucket_order=0 [ 5.120799] zbud: loaded [ 5.126252] *** VALIDATE nfs *** [ 5.128326] *** VALIDATE nfs4 *** [ 5.132228] pstore: using deflate compression [ 5.140805] Platform Keyring initialized [ 5.376612] NET: Registered protocol family 38 [ 5.379614] Key type asymmetric registered [ 5.382244] Asymmetric key parser 'x509' registered [ 5.385915] Block layer SCSI generic (bsg) driver version 0.4 loaded (major 247) [ 5.394030] io scheduler mq-deadline registered [ 5.397909] io scheduler kyber registered [ 5.400627] io scheduler bfq registered [ 5.404598] atomic64_test: passed for x86-64 platform with CX8 and with SSE [ 5.411377] shpchp: Standard Hot Plug PCI Controller Driver version: 0.4 [ 5.419231] input: Power Button as /devices/LNXSYSTM:00/LNXPWRBN:00/input/input0 [ 5.429357] ACPI: Power Button [PWRF] [ 5.442744] ACPI: \_SB_.LNKB: Enabled at IRQ 10 [ 5.458251] ACPI: \_SB_.LNKA: Enabled at IRQ 11 [ 5.503080] ACPI: \_SB_.LNKC: Enabled at IRQ 11 [ 5.523109] ACPI: \_SB_.LNKD: Enabled at IRQ 10 [ 5.567646] Serial: 8250/16550 driver, 4 ports, IRQ sharing enabled [ 5.612197] 00:03: ttyS1 at I/O 0x2f8 (irq = 3, base_baud = 115200) is a 16550A [ 5.660922] 00:04: ttyS0 at I/O 0x3f8 (irq = 4, base_baud = 115200) is a 16550A [ 5.674474] Non-volatile memory driver v1.3 [ 5.677949] Linux agpgart interface v0.103 [ 5.756577] virtio_blk virtio1: [vda] 149376 512-byte logical blocks (76.5 MB/72.9 MiB) [ 5.763556] vda: detected capacity change from 0 to 76480512 [ 5.785526] virtio_blk virtio2: [vdb] 2097152 512-byte logical blocks (1.07 GB/1.00 GiB) [ 5.790857] vdb: detected capacity change from 0 to 1073741824 [ 5.820280] virtio_blk virtio3: [vdc] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 5.827396] vdc: detected capacity change from 0 to 2621440000 [ 5.855545] virtio_blk virtio4: [vdd] 5120000 512-byte logical blocks (2.62 GB/2.44 GiB) [ 5.858245] vdd: detected capacity change from 0 to 2621440000 [ 5.879364] virtio_blk virtio5: [vde] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 5.883958] vde: detected capacity change from 0 to 4294967296 [ 5.928518] virtio_blk virtio6: [vdf] 8388608 512-byte logical blocks (4.29 GB/4.00 GiB) [ 5.932101] vdf: detected capacity change from 0 to 4294967296 [ 5.946653] libphy: Fixed MDIO Bus: probed [ 5.957498] usbcore: registered new interface driver usbserial_generic [ 5.960827] usbserial: USB Serial support registered for generic [ 5.963946] i8042: PNP: PS/2 Controller [PNP0303:KBD,PNP0f13:MOU] at 0x60,0x64 irq 1,12 [ 5.970512] serio: i8042 KBD port at 0x60,0x64 irq 1 [ 5.972856] serio: i8042 AUX port at 0x60,0x64 irq 12 [ 5.975852] mousedev: PS/2 mouse device common for all mice [ 5.979846] input: AT Translated Set 2 keyboard as /devices/platform/i8042/serio0/input/input1 [ 5.981289] rtc_cmos 00:05: RTC can wake from S4 [ 5.987760] rtc_cmos 00:05: registered as rtc0 [ 5.989960] rtc_cmos 00:05: alarms up to one day, y3k, 242 bytes nvram, hpet irqs [ 5.993761] intel_pstate: CPU model not supported [ 5.998505] hid: raw HID events driver (C) Jiri Kosina [ 6.001116] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input4 [ 6.001954] usbcore: registered new interface driver usbhid [ 6.011556] usbhid: USB HID core driver [ 6.014919] input: VirtualPS/2 VMware VMMouse as /devices/platform/i8042/serio1/input/input3 [ 6.017061] drop_monitor: Initializing network drop monitor service [ 6.024913] Initializing XFRM netlink socket [ 6.027115] NET: Registered protocol family 10 [ 6.031359] Segment Routing with IPv6 [ 6.033403] NET: Registered protocol family 17 [ 6.035957] mpls_gso: MPLS GSO support [ 6.043741] RAS: Correctable Errors collector initialized. [ 6.046522] AVX version of gcm_enc/dec engaged. [ 6.048653] AES CTR mode by8 optimization enabled [ 6.157111] sched_clock: Marking stable (6157087749, 0)->(7405015290, -1247927541) [ 6.164212] registered taskstats version 1 [ 6.168515] Loading compiled-in X.509 certificates [ 6.171820] zswap: loaded using pool lzo/zbud [ 6.216561] Key type big_key registered [ 6.237552] Key type encrypted registered [ 6.240803] ima: No TPM chip found, activating TPM-bypass! [ 6.246197] ima: Allocated hash algorithm: sha1 [ 6.249522] ima: No architecture policies found [ 6.252230] evm: Initialising EVM extended attributes: [ 6.255456] evm: security.selinux [ 6.260717] evm: security.ima [ 6.262198] evm: security.capability [ 6.263879] evm: HMAC attrs: 0x1 [ 6.267524] rtc_cmos 00:05: setting system clock to 2026-08-10 21:53:54 UTC (1786398834) [ 6.284799] debug: unmapping init [mem 0xffffffffb0e03000-0xffffffffb0ffffff] [ 6.289820] debug: unmapping init [mem 0xffffffffafb82000-0xffffffffafe58fff] [ 6.298235] Write protecting the kernel read-only data: 28672k [ 6.302945] debug: unmapping init [mem 0xffffffffae203000-0xffffffffae3fffff] [ 6.306321] debug: unmapping init [mem 0xffffffffaeb14000-0xffffffffaebfffff] [ 6.353756] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 6.363770] systemd[1]: Detected virtualization kvm. [ 6.366371] systemd[1]: Detected architecture x86-64. [ 6.369185] systemd[1]: Running in initial RAM disk. Welcome to Rocky Linux 8.10 (Green Obsidian) dracut-049-233.git20240115.el8 (Initramfs)! [ 6.404161] systemd[1]: No hostname configured. [ 6.406465] systemd[1]: Set hostname to . [ 6.408894] random: systemd: uninitialized urandom read (16 bytes read) [ 6.412377] systemd[1]: Initializing machine ID from random generator. [ 6.557558] random: ln: uninitialized urandom read (6 bytes read) [ 6.763507] random: systemd: uninitialized urandom read (16 bytes read) [ 6.768198] systemd[1]: Listening on udev Control Socket. [ OK ] Listening on udev Control Socket. [ 6.782915] systemd[1]: Reached target Initrd Root Device. [ OK ] Reached target Initrd Root Device. [ 6.791490] systemd[1]: Reached target Swap. [ OK ] Reached target Swap. [ OK ] Listening on Journal Socket (/dev/log). [ OK ] Reached target Local File Systems. [ OK ] Listening on Journal Socket. Starting Journal Service... Starting Create list of required st…ce nodes for the current kernel... [ OK ] Reached target Timers. [ OK ] Started Memstrack Anylazing Service. [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Slices. Starting Setup Virtual Console... [ OK ] Reached target Local Encrypted Volumes. [ OK ] Listening on udev Kernel Socket. [ OK ] Reached target Sockets. Starting Apply Kernel Variables... Starting Create Volatile Files and Directories... [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Setup Virtual Console. [ OK ] Started Apply Kernel Variables. [ OK ] Started Create Volatile Files and Directories. Starting dracut cmdline hook... Starting Create Static Device Nodes in /dev... [ OK ] Started Journal Service. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Started dracut cmdline hook. Starting dracut pre-udev hook... [ 8.746755] device-mapper: uevent: version 1.0.3 [ 8.752089] device-mapper: ioctl: 4.46.0-ioctl (2022-02-22) initialised: dm-devel@redhat.com [ OK ] Started dracut pre-udev hook. Starting udev Kernel Device Manager... [ OK ] Started udev Kernel Device Manager. Starting dracut pre-trigger hook... [ OK ] Started dracut pre-trigger hook. Starting udev Coldplug all Devices... Mounting Kernel Configuration File System... [ OK ] Mounted Kernel Configuration File System. [ OK ] Started udev Coldplug all Devices. Starting dracut initqueue hook... [ OK ] Reached target System Initialization. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ 10.775213] virtio_net virtio0 ens2: renamed from eth0 [ 10.859686] random: fast init done [ 11.237549] scsi host0: ata_piix [ 11.350676] scsi host1: ata_piix [ 11.355396] ata1: PATA max MWDMA2 cmd 0x1f0 ctl 0x3f6 bmdma 0xc320 irq 14 [ 11.368203] ata2: PATA max MWDMA2 cmd 0x170 ctl 0x376 bmdma 0xc328 irq 15 [ 16.771824] random: crng init done [ 16.774540] random: 7 urandom warning(s) missed due to ratelimiting [ 18.310902] dracut-initqueue[577]: RTNETLINK answers: File exists Starting nbd nbd0... [ OK ] Started nbd nbd0. [ OK ] Started dracut initqueue hook. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Mounting /sysroot... [ 20.254360] EXT4-fs (nbd0): mounted filesystem with ordered data mode. Opts: (null) [ OK ] Mounted /sysroot. [ OK ] Reached target Initrd Root File System. Starting Reload Configuration from the Real Root... [ OK ] Started Reload Configuration from the Real Root. [ OK ] Reached target Initrd File Systems. [ OK ] Reached target Initrd Default Target. Starting dracut pre-pivot and cleanup hook... [ OK ] Started dracut pre-pivot and cleanup hook. Starting Cleaning Up and Shutting Down Daemons... Stopping Hardware RNG Entropy Gatherer Daemon... [ OK ] Stopped dracut pre-pivot and cleanup hook. [ OK ] Stopped target Initrd Default Target. [ OK ] Stopped target Initrd Root Device. [ OK ] Stopped target Remote File Systems. [ OK ] Stopped target Remote File Systems (Pre). [ OK ] Stopped dracut initqueue hook. [ OK ] Stopped target Timers. [ OK ] Stopped Hardware RNG Entropy Gatherer Daemon. [ OK ] Stopped target Basic System. [ OK ] Stopped target Sockets. [ OK ] Stopped target Slices. [ OK ] Stopped target Paths. [ OK ] Stopped target System Initialization. [ OK ] Stopped Create Volatile Files and Directories. [ OK ] Stopped target Swap. [ OK ] Stopped target Local Encrypted Volumes. [ OK ] Stopped Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped udev Coldplug all Devices. [ OK ] Stopped dracut pre-trigger hook. [ OK ] Stopped target Local File Systems. Stopping udev Kernel Device Manager... [ OK ] Stopped Apply Kernel Variables. [ OK ] Stopped udev Kernel Device Manager. [ OK ] Started Cleaning Up and Shutting Down Daemons. [ OK ] Stopped dracut pre-udev hook. [ OK ] Stopped dracut cmdline hook. [ OK ] Stopped Create Static Device Nodes in /dev. [ OK ] Stopped Create list of required sta…vice nodes for the current kernel. [ OK ] Closed udev Kernel Socket. [ OK ] Closed udev Control Socket. Starting Cleanup udevd DB... [ OK ] Started Cleanup udevd DB. [ OK ] Reached target Switch Root. Starting Switch Root... [ 23.496210] printk: systemd: 25 output lines suppressed due to ratelimiting [ 24.130984] SELinux: Disabled at runtime. [ 24.237129] systemd[1]: systemd 239 (239-82.el8_10.5) running in system mode. (+PAM +AUDIT +SELINUX +IMA -APPARMOR +SMACK +SYSVINIT +UTMP +LIBCRYPTSETUP +GCRYPT +GNUTLS +ACL +XZ +LZ4 +SECCOMP +BLKID +ELFUTILS +KMOD +IDN2 -IDN +PCRE2 default-hierarchy=legacy) [ 24.260493] systemd[1]: Detected virtualization kvm. [ 24.263059] systemd[1]: Detected architecture x86-64. Welcome to Rocky Linux 8.10 (Green Obsidian)! [ 25.525614] systemd[1]: initrd-switch-root.service: Succeeded. [ 25.533707] systemd[1]: Stopped Switch Root. [ OK ] Stopped Switch Root. [ 25.551210] systemd[1]: systemd-journald.service: Service has no hold-off time (RestartSec=0), scheduling restart. [ 25.563753] systemd[1]: systemd-journald.service: Scheduled restart job, restart counter is at 1. [ 25.574401] systemd[1]: Stopped Journal Service. [ OK ] Stopped Journal Service. [ 25.587168] systemd[1]: Starting Journal Service... Starting Journal Service... [ 25.598901] systemd[1]: Mounting POSIX Message Queue File System... Mounting POSIX Message Queue File System... [ OK ] Created slice system-sshd\x2dkeygen.slice. Mounting Kernel Debug File System... [ OK ] Started Dispatch Password Requests to Console Directory Watch. [ OK ] Stopped target Switch Root. [ OK ] Listening on udev Kernel Socket. Activating swap /dev/disk/by-label/SWAP... [ OK ] Listening on udev Control Socket. [ OK ] Stopped target Initrd Root File System.[ 25.834375] Adding 1048572k swap on /dev/vdb. Priority:-2 extents:1 across:1048572k FS Starting Remount Root and Kernel File Systems... Mounting Huge Pages File System... [ OK ] Listening on initctl Compatibility Named Pipe. [ OK ] Created slice system-getty.slice. [ OK ] Listening on Process Core Dump Socket. [FAILED] Failed to set up automount Arbitrar…rmats File System Automount Point. See 'systemctl status proc-sys-fs-binfmt_misc.automount' for details. [ OK ] Stopped target Initrd File Systems. [ OK ] Reached target rpc_pipefs.target. [ OK ] Created slice system-serial\x2dgetty.slice. [ OK ] Started Forward Password Requests to Wall Directory Watch. [ OK ] Reached target Paths. [ OK ] Reached target Local Encrypted Volumes. [ OK ] Created slice User and Session Slice. [ OK ] Reached target Slices. Starting Create list of required st…ce nodes for the current kernel... Starting udev Coldplug all Devices... [ OK ] Listening on RPCbind Server Activation Socket. [ OK ] Reached target RPC Port Mapper. Starting Apply Kernel Variables... [ OK ] Started Journal Service. [ OK ] Mounted POSIX Message Queue File System. [ OK ] Mounted Kernel Debug File System. [ OK ] Activated swap /dev/disk/by-label/SWAP. [FAILED] Failed to start Remount Root and Kernel File Systems. See 'systemctl status systemd-remount-fs.service' for details. [ OK ] Mounted Huge Pages File System. [ OK ] Started Create list of required sta…vice nodes for the current kernel. [ OK ] Started Apply Kernel Variables. Starting Configure read-only root support... Starting Create Static Device Nodes in /dev... [ OK ] Reached target Swap. Starting Flush Journal to Persistent Storage... [ OK ] Started Flush Journal to Persistent Storage. [ OK ] Started udev Coldplug all Devices. [ OK ] Started Create Static Device Nodes in /dev. [ OK ] Reached target Local File Systems (Pre). Mounting /mnt... Mounting /home/green/git/lustre-release... Starting udev Kernel Device Manager... [ OK ] Mounted /mnt. [ 27.172712] squashfs: version 4.0 (2009/01/31) Phillip Lougher [ OK ] Mounted /home/green/git/lustre-release. [ OK ] Started udev Kernel Device Manager. [ 28.302940] piix4_smbus 0000:00:01.3: SMBus Host Controller at 0x700, revision 0 [ 28.556828] input: PC Speaker as /devices/platform/pcspkr/input/input5 [ 29.154318] RAPL PMU: API unit is 2^-32 Joules, 0 fixed counters, 10737418240 ms ovfl timer [ 29.292842] EDAC sbridge: Ver: 1.1.2 [* ] A start job is running for Configur…-only root support (7s / no limit) [** ] A start job is running for Configur…-only root support (8s / no limit) [*** ] A start job is running for Configur…-only root support (8s / no limit)[ 34.571361] Key type dns_resolver registered [ *** ] A start job is running for Configur…-only root support (9s / no limit)[ 35.194944] NFS: Registering the id_resolver key type [ 35.206803] Key type id_resolver registered [ 35.212107] Key type id_legacy registered [ *** ] A start job is running for Configur…only root support (10s / no limit) [ ***] A start job is running for Configur…only root support (10s / no limit) [ OK ] Started Configure read-only root support. Starting Load/Save Random Seed... [ OK ] Reached target Local File Systems. Starting Rebuild Dynamic Linker Cache... Starting Mark the need to relabel after reboot... Starting Create Volatile Files and Directories... [ OK ] Started Load/Save Random Seed. [ OK ] Started Mark the need to relabel after reboot. [ OK ] Started Create Volatile Files and Directories. Starting RPC Bind... Starting Update UTMP about System Boot/Shutdown... [ OK ] Started Update UTMP about System Boot/Shutdown. [ OK ] Started RPC Bind. [ OK ] Started Rebuild Dynamic Linker Cache. Starting Update is Completed... [ OK ] Started Update is Completed. [ OK ] Reached target System Initialization. [ OK ] Started dnf makecache --timer. [ OK ] Started daily update of the root trust anchor for DNSSEC. [ OK ] Started Daily Cleanup of Temporary Directories. [ OK ] Reached target Timers. [ OK ] Listening on D-Bus System Message Bus Socket. [ OK ] Reached target Sockets. [ OK ] Reached target Basic System. [ OK ] Started Hardware RNG Entropy Gatherer Daemon. [ OK ] Started D-Bus System Message Bus. Starting Restore /run/initramfs on shutdown... Starting Login Service... Starting Network Manager... [ OK ] Reached target sshd-keygen.target. [ OK ] Started irqbalance daemon. [ OK ] Started Restore /run/initramfs on shutdown. [ OK ] Started Network Manager. [ OK ] Reached target Network. Starting OpenSSH server daemon... Starting GSSAPI Proxy Daemon... Starting Dynamic System Tuning Daemon... Starting Network Manager Wait Online... [ OK ] Started GSSAPI Proxy Daemon. [ OK ] Started Login Service. [ OK ] Reached target NFS client services. [ OK ] Reached target Remote File Systems (Pre). [ OK ] Reached target Remote File Systems. Starting Permit User Sessions... Starting Hostname Service... [ OK ] Started OpenSSH server daemon. [ OK ] Started Permit User Sessions. [ OK ] Started Getty on tty1. [ OK ] Started Serial Getty on ttyS1. [ OK ] Started Command Scheduler. [ OK ] Started Serial Getty on ttyS0. [ OK ] Reached target Login Prompts. [ OK ] Started Hostname Service. Starting Network Manager Script Dispatcher Service... [ OK ] Started Network Manager Script Dispatcher Service. [ OK ] Started Network Manager Wait Online. [ OK ] Reached target Network is Online. Starting System Logging Service... Starting Notify NFS peers of a restart... Starting Crash recovery kernel arming... [ OK ] Started Notify NFS peers of a restart. [ OK ] Started System Logging Service. Rocky Linux 8.10 (Green Obsidian) Kernel 4.18.0rh8.10-debug on an x86_64 oleg436-server login: [ 86.815037] hrtimer: interrupt took 3027640 ns [ 92.736963] libcfs: loading out-of-tree module taints kernel. [ 92.797842] Key type ._llcrypt registered [ 92.802926] Key type .llcrypt registered [ 92.917470] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_hostid [ 112.263412] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing load_modules_local [ 114.007384] libcfs: HW NUMA nodes: 1, HW CPU cores: 4, npartitions: 1 [ 114.055800] alg: No test for adler32 (adler32-zlib) [ 115.305529] Lustre: Lustre: Build Version: 2.17.54_107_gf473267 [ 116.021532] LNet: Added LNI 192.168.204.136@tcp [8/256/0/180] [ 117.815181] Key type lgssc registered [ 119.865948] Lustre: Echo OBD driver; http://www.lustre.org/ [ 135.300033] ZFS: Loaded module v2.3.2-1, ZFS pool version 5000, ZFS filesystem version 5 [ 172.651162] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing load_modules_local [ 185.631038] Lustre: lustre-MDT0000: mounting server target with '-t lustre' deprecated, use '-t lustre_tgt' [ 185.685323] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 186.897498] Lustre: Setting parameter lustre-MDT0000.mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity in log lustre-MDT0000 [ 186.919597] Lustre: ctl-lustre-MDT0000: No data found on store. Initialize space. [ 186.979937] Lustre: lustre-MDT0000: new disk, initializing [ 187.071959] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 187.106462] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000200000400-0x0000000240000400]:0:mdt [ 192.793845] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 206.225491] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 206.387993] Lustre: 6501:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify lustre-MDT0001/mdt.identity_upcall=/home/green/git/lustre-release/lustre/utils/l_getidentity (mode = 0) failed: rc = -17 [ 206.439451] Lustre: srv-lustre-MDT0001: No data found on store. Initialize space. [ 206.444640] Lustre: Skipped 1 previous similar message [ 206.538508] Lustre: lustre-MDT0001: new disk, initializing [ 206.653990] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 206.735066] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000240000400-0x0000000280000400]:1:mdt [ 206.759321] Lustre: cli-ctl-lustre-MDT0001: Allocated super-sequence [0x0000000240000400-0x0000000280000400]:1:mdt] [ 212.060942] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 216.656836] Lustre: Modifying parameter general.debug_raw_pointers=Y in log params [ 226.313091] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 226.536819] Lustre: lustre-OST0000: new disk, initializing [ 226.540364] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 226.547391] Lustre: 8439:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 226.623380] Lustre: lustre-OST0000: Imperative Recovery not enabled, recovery window 60-180 [ 233.753757] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 234.525611] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 234.534768] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 234.652172] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 248.693642] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 248.820890] Lustre: lustre-OST0001: new disk, initializing [ 248.830912] Lustre: srv-lustre-OST0001: No data found on store. Initialize space. [ 248.840643] Lustre: 9512:0:(osd_compat.c:1352:osd_object_spec_find()) UNKNOWN COMPAT FID [0x200000001:0x101e:0x0] [ 248.908671] Lustre: lustre-OST0001: Imperative Recovery not enabled, recovery window 60-180 [ 255.262441] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 256.710077] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x00000002c0000400-0x0000000300000400]:1:ost [ 256.728551] Lustre: cli-lustre-OST0001-super: Allocated super-sequence [0x00000002c0000400-0x0000000300000400]:1:ost] [ 256.889509] Lustre: lustre-OST0001-osc-MDT0000: update sequence from 0x100010000 to 0x2c0000401 [ 268.153740] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 276.199751] Lustre: Setting parameter general.lod.*.mdt_hash=crush in log params [ 283.320561] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing check_logdir /tmp/testlogs/ [ 289.104487] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing yml_node [ 293.627509] Lustre: DEBUG MARKER: Client: 2.17.54.107 [ 296.118957] Lustre: DEBUG MARKER: MDS: 2.17.54.107 [ 298.545451] Lustre: DEBUG MARKER: OSS: 2.17.54.107 [ 300.313260] Lustre: DEBUG MARKER: -----============= acceptance-small: replay-dual ============----- Mon Aug 10 17:58:46 EDT 2026 [ 317.078935] Lustre: DEBUG MARKER: excepting tests: 14b 21b [ 318.843700] Lustre: DEBUG MARKER: skipping tests SLOW=no: 21b [ 320.597696] Lustre: DEBUG MARKER: === replay-dual: start setup 17:59:07 (1786399147) === [ 327.465460] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing check_config_client /mnt/lustre [ 348.164873] Lustre: DEBUG MARKER: Using TIMEOUT=20 [ 351.818893] Lustre: 13371:0:(mgs_llog.c:1450:mgs_modify_param()) MGS: modify general/lod.*.mdt_hash=crush (mode = 0) failed: rc = -17 [ 356.438934] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 363.107175] Lustre: DEBUG MARKER: === replay-dual: finish setup 17:59:49 (1786399189) === [ 365.562735] Lustre: DEBUG MARKER: == replay-dual test 0a: expired recovery with lost client ========================================================== 17:59:51 (1786399191) [ 373.412867] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 377.041320] Lustre: Failing over lustre-MDT0000 [ 377.509945] Lustre: server umount lustre-MDT0000 complete [ 379.878761] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 379.899560] Lustre: Skipped 3 previous similar messages [ 383.119614] LustreError: 6508:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 383.163075] LustreError: 6508:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 8 previous similar messages [ 384.994812] LustreError: 7845:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 385.016313] LustreError: 7845:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 3 previous similar messages [ 388.244971] LustreError: 6508:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 393.393294] LustreError: 6509:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 393.436162] LustreError: 6509:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 6 previous similar messages [ 396.255678] Lustre: 3635:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399208/real 1786399208] req@ffff97bdc4e15500 x1873175059221248/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399224 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 396.324750] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 398.481442] LustreError: 6507:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 398.497563] LustreError: 6507:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 11 previous similar messages [ 402.541443] LDISKFS-fs (dm-0): 10 truncates cleaned up [ 402.550043] LDISKFS-fs (dm-0): recovery complete [ 402.582679] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 406.003268] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2ec50b [ 406.025242] Lustre: MGC192.168.204.136@tcp: Connection restored to 0@lo (at 0@lo) [ 406.475349] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 408.173121] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 411.126232] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 411.639388] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 513.500193] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 513.505557] Lustre: 14917:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client eda05ca4-fc39-43e1-8668-3008ff425e0c@192.168.204.36@tcp [ 513.512355] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 513.535940] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 513.544744] Lustre: 14917:0:(ldlm_lib.c:2937:target_recovery_thread()) too long recovery - read logs [ 513.563100] Lustre: Skipped 2 previous similar messages [ 513.588153] LustreError: dumping log to /tmp/lustre-log.1786399341.14917 [ 513.814852] Lustre: lustre-MDT0000: Recovery over after 1:45, of 3 clients 2 recovered and 1 was evicted. [ 513.887447] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:65) [ 513.902204] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:65) [ 535.240786] Lustre: DEBUG MARKER: == replay-dual test 0b: lost client during waiting for next transno ========================================================== 18:02:42 (1786399362) [ 542.392717] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 544.473491] Lustre: Failing over lustre-MDT0000 [ 544.733070] Lustre: server umount lustre-MDT0000 complete [ 544.736843] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 544.747880] LustreError: 14901:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 544.761847] Lustre: Skipped 1 previous similar message [ 544.849266] LustreError: 14901:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 14 previous similar messages [ 549.346321] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 549.357787] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 549.373492] Lustre: Skipped 1 previous similar message [ 562.332465] LustreError: 14901:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 562.366477] LustreError: 14901:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 15 previous similar messages [ 566.240365] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399378/real 1786399378] req@ffff97bdc30b0a80 x1873175059302784/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399394 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 566.303891] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 566.902365] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 566.905171] LDISKFS-fs (dm-0): recovery complete [ 566.911774] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 576.491600] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2ee0c5 [ 576.504814] Lustre: MGC192.168.204.136@tcp: Connection restored to 0@lo (at 0@lo) [ 576.935579] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 577.687246] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 581.431783] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 582.153905] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 594.344946] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:53 [ 599.698798] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:47 [ 604.814989] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:42 [ 609.934812] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:37 [ 612.839380] Lustre: lustre-MDT0001: haven't heard from client eda05ca4-fc39-43e1-8668-3008ff425e0c (at 192.168.204.36@tcp) in 103 seconds. I think it's dead, and I am evicting it. exp ffff97beeebc6800, cur 1786399441 deadline 1786399438 last 1786399338 [ 615.062530] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:32 [ 625.301878] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:22 [ 625.326875] Lustre: Skipped 1 previous similar message [ 645.778309] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 0 evicted) to recover in 0:01 [ 645.792292] Lustre: Skipped 3 previous similar messages [ 647.500839] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 647.507233] Lustre: 16669:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client e69aface-25ba-43fe-bc1a-bf5314cfd028@ [ 647.522086] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 681.617819] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 1:06 [ 681.630414] Lustre: Skipped 6 previous similar messages [ 684.526481] Lustre: lustre-MDT0001: haven't heard from client 2b27f910-dca2-46c5-9f67-7344b69b64e8 (at 192.168.204.36@tcp) in 102 seconds. I think it's dead, and I am evicting it. exp ffff97bdc3373800, cur 1786399512 deadline 1786399510 last 1786399410 [ 748.181983] Lustre: lustre-MDT0000: Denying connection for new client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp), waiting for 3 known clients (1 recovered, 1 in progress, and 1 evicted) to recover in 0:00 [ 748.204533] Lustre: Skipped 12 previous similar messages [ 748.500152] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 748.507439] Lustre: 16669:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 2b27f910-dca2-46c5-9f67-7344b69b64e8@192.168.204.36@tcp [ 748.520418] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 748.530164] Lustre: 16669:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 748.550738] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 748.550797] Lustre: 16669:0:(ldlm_lib.c:2937:target_recovery_thread()) too long recovery - read logs [ 748.560806] Lustre: Skipped 2 previous similar messages [ 748.580250] LustreError: dumping log to /tmp/lustre-log.1786399576.16669 [ 748.651547] Lustre: lustre-MDT0000: Recovery over after 2:51, of 3 clients 1 recovered and 2 were evicted. [ 748.697670] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:28 to 0x280000401:97) [ 748.698846] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:28 to 0x2c0000401:97) [ 763.118858] Lustre: DEBUG MARKER: == replay-dual test 1: |X| simple create ================= 18:06:30 (1786399590) [ 772.020686] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 774.231978] Lustre: Failing over lustre-MDT0000 [ 774.493090] Lustre: server umount lustre-MDT0000 complete [ 774.852037] LustreError: 8104:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 774.901386] LustreError: 8104:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 21 previous similar messages [ 776.684024] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 776.709691] Lustre: Skipped 3 previous similar messages [ 793.131328] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399604/real 1786399604] req@ffff97becafac000 x1873175059403904/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399620 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 793.165566] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 796.720330] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 796.722622] LDISKFS-fs (dm-0): recovery complete [ 796.728486] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 802.802733] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2ee691 [ 802.816896] Lustre: MGC192.168.204.136@tcp: Connection restored to 0@lo (at 0@lo) [ 803.289517] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 803.358373] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 804.846983] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 807.892658] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 808.659417] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 808.704584] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:129) [ 808.705614] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:129) [ 817.131473] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 818.964935] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 828.141700] Lustre: DEBUG MARKER: == replay-dual test 2: |X| mkdir adir ==================== 18:07:34 (1786399654) [ 835.348415] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 837.212048] Lustre: Failing over lustre-MDT0000 [ 837.563973] Lustre: server umount lustre-MDT0000 complete [ 839.146518] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 839.162476] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 839.187664] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 41 previous similar messages [ 855.528882] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399667/real 1786399667] req@ffff97bdc52cdc00 x1873175059441024/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399683 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 855.582760] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 860.873651] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 860.878923] LDISKFS-fs (dm-0): recovery complete [ 860.895904] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 864.803675] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97becafaea00 x1873175059449472/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 865.358768] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 865.448423] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 867.997574] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 870.392337] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 870.404654] Lustre: Skipped 4 previous similar messages [ 870.500091] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 870.586612] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:161) [ 870.587056] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:161) [ 871.100653] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 881.084721] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 883.290667] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 893.495368] Lustre: DEBUG MARKER: == replay-dual test 3: |X| mkdir adir, mkdir adir/bdir === 18:08:40 (1786399720) [ 903.133165] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 905.410923] Lustre: Failing over lustre-MDT0000 [ 905.845291] Lustre: server umount lustre-MDT0000 complete [ 906.211504] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 906.223584] Lustre: Skipped 3 previous similar messages [ 922.529180] Lustre: 3638:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399734/real 1786399734] req@ffff97bdc4e14700 x1873175059485312/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399750 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 922.585134] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 929.901357] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 929.903489] LDISKFS-fs (dm-0): recovery complete [ 929.914657] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 931.810250] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc5a0c000 x1873175059493504/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 932.253659] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 932.330620] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 935.578125] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 937.455153] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 937.465494] Lustre: Skipped 3 previous similar messages [ 937.675436] Lustre: lustre-MDT0000: Recovery over after 0:02, of 3 clients 3 recovered and 0 were evicted. [ 937.765962] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:193) [ 937.768539] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:193) [ 937.964718] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 948.258800] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 950.003711] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 960.235572] Lustre: DEBUG MARKER: == replay-dual test 4: |X| mkdir adir (-EEXIST), mkdir adir/bdir ========================================================== 18:09:46 (1786399786) [ 969.694909] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 972.115783] Lustre: Failing over lustre-MDT0000 [ 972.378220] Lustre: server umount lustre-MDT0000 complete [ 973.289306] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 973.314256] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 973.356092] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 80 previous similar messages [ 989.665300] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399801/real 1786399801] req@ffff97beeeb71880 x1873175059530880/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399817 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 989.687831] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 997.118535] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 997.124554] LDISKFS-fs (dm-0): recovery complete [ 997.136984] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 998.886344] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc3e96300 x1873175059539200/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 999.555841] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1001.106152] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1004.819379] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 1004.860681] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1004.864311] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:225) [ 1004.864588] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:225) [ 1016.804441] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1018.956632] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1028.515992] Lustre: DEBUG MARKER: == replay-dual test 5: open, unlink |X| close ============ 18:10:55 (1786399855) [ 1037.204296] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1040.110917] Lustre: Failing over lustre-MDT0000 [ 1040.562897] Lustre: server umount lustre-MDT0000 complete [ 1040.873385] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1040.890131] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1040.927031] Lustre: Skipped 3 previous similar messages [ 1059.799167] Lustre: 3635:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399871/real 1786399871] req@ffff97beedd93480 x1873175059575168/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399887 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1059.835375] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1065.210261] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1065.213365] LDISKFS-fs (dm-0): recovery complete [ 1065.221968] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1070.063479] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bef95db100 x1873175059583488/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1070.540826] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1070.549491] Lustre: Skipped 1 previous similar message [ 1070.622708] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1071.631303] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1075.746205] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1075.763656] Lustre: Skipped 7 previous similar messages [ 1076.030810] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1076.123296] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:257) [ 1076.124686] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:257) [ 1076.639313] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1089.000146] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1090.747578] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1100.934829] Lustre: DEBUG MARKER: == replay-dual test 6: open1, open2, unlink |X| close1 [fail mds1] close2 ========================================================== 18:12:07 (1786399927) [ 1109.156612] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1111.341856] Lustre: Failing over lustre-MDT0000 [ 1111.525297] Lustre: lustre-MDT0000-osp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1111.535902] Lustre: Skipped 3 previous similar messages [ 1111.604934] Lustre: server umount lustre-MDT0000 complete [ 1133.030778] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786399944/real 1786399944] req@ffff97bec2bdfb80 x1873175059617024/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786399960 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1133.082206] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1134.952724] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1134.956671] LDISKFS-fs (dm-0): recovery complete [ 1134.964514] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1143.272492] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2f075a [ 1143.461307] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (not set up) [ 1143.485301] Lustre: Skipped 1 previous similar message [ 1143.821733] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1144.381285] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1149.100807] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 1149.166957] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:289) [ 1149.168085] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:289) [ 1149.379908] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1159.372420] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1161.547174] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1171.682584] Lustre: DEBUG MARKER: == replay-dual test 8: replay of resent request ========== 18:13:18 (1786399998) [ 1180.715435] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1181.833086] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1181.837987] LustreError: 6509:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bef9923100 x1873175039469440/t38654705670(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:561/0 lens 512/448 e 0 to 0 dl 1786400021 ref 1 fl Interpret:/200/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1198.311256] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnecting [ 1198.346095] Lustre: 6508:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bef0719c00 x1873175039469440/t38654705670(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:577/0 lens 512/2880 e 0 to 0 dl 1786400037 ref 1 fl Interpret:/202/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1202.334293] Lustre: Failing over lustre-MDT0000 [ 1202.608892] Lustre: server umount lustre-MDT0000 complete [ 1221.601673] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786400033/real 1786400033] req@ffff97beee9f1f80 x1873175059665920/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786400049 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1221.642051] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1227.789444] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1227.791450] LDISKFS-fs (dm-0): recovery complete [ 1227.800542] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1230.882088] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bef9921880 x1873175059674496/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1230.940174] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1230.967911] LustreError: 6513:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 160 previous similar messages [ 1231.540680] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1232.017318] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1236.570195] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 1236.626210] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:321) [ 1236.630488] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:321) [ 1237.653414] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1248.426360] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1250.093708] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1259.305982] Lustre: DEBUG MARKER: == replay-dual test 9: resending a replayed create ======= 18:14:45 (1786400085) [ 1267.159825] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1269.832100] Lustre: Failing over lustre-MDT0000 [ 1270.052796] Lustre: server umount lustre-MDT0000 complete [ 1272.294131] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1272.311048] Lustre: Skipped 7 previous similar messages [ 1292.255503] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1292.259169] LDISKFS-fs (dm-0): recovery complete [ 1292.269313] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1297.895643] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97beee9f1880 x1873175059715328/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1298.498110] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1303.594370] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1303.611926] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1303.615264] LustreError: 32171:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bef7a28e00 x1873175039486464/t42949672962(42949672962) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:678/0 lens 528/448 e 0 to 0 dl 1786400138 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1314.980280] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnected, waiting for 3 clients in recovery for 1:29 [ 1315.100629] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:353) [ 1315.114221] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:353) [ 1322.034133] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1323.731536] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1333.300743] Lustre: DEBUG MARKER: == replay-dual test 10: resending a replayed unlink ====== 18:16:00 (1786400160) [ 1341.022767] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1344.053035] Lustre: Failing over lustre-MDT0000 [ 1344.264467] Lustre: server umount lustre-MDT0000 complete [ 1346.027365] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1362.912754] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786400175/real 1786400175] req@ffff97beedd93100 x1873175059750912/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786400191 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1362.948254] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 1 previous similar message [ 1362.953305] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1362.961359] LustreError: Skipped 1 previous similar message [ 1367.307470] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1367.314652] LDISKFS-fs (dm-0): recovery complete [ 1367.330559] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1373.163582] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2f192e [ 1373.183086] Lustre: MGC192.168.204.136@tcp: Connection restored to 0@lo (at 0@lo) [ 1373.189569] Lustre: Skipped 16 previous similar messages [ 1373.549648] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1373.565313] Lustre: Skipped 3 previous similar messages [ 1373.637518] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1374.266624] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1374.277344] Lustre: Skipped 1 previous similar message [ 1378.786895] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1378.928161] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1378.930787] LustreError: 34226:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bf00a48000 x1873175039506048/t47244640260(47244640260) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:753/0 lens 528/448 e 0 to 0 dl 1786400213 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1390.764755] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnected, waiting for 3 clients in recovery for 1:28 [ 1390.852922] Lustre: lustre-MDT0000: Recovery over after 0:16, of 3 clients 3 recovered and 0 were evicted. [ 1390.880908] Lustre: Skipped 1 previous similar message [ 1390.936089] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:385) [ 1390.939856] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:385) [ 1397.348763] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1398.835419] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1407.687922] Lustre: DEBUG MARKER: == replay-dual test 11: both clients timeout during replay ========================================================== 18:17:14 (1786400234) [ 1415.490408] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1418.651433] Lustre: Failing over lustre-MDT0000 [ 1419.090043] Lustre: server umount lustre-MDT0000 complete [ 1442.089639] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1442.096457] LDISKFS-fs (dm-0): recovery complete [ 1442.106778] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1445.368807] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97beedd93480 x1873175059798400/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1450.515400] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1451.199864] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 1451.209735] LustreError: 36286:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bf00a47100 x1873175039525504/t51539607554(51539607554) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:70/0 lens 528/448 e 0 to 0 dl 1786400285 ref 1 fl Complete:/204/0 rc 0/0 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 1458.652451] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1462.451701] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnected, waiting for 3 clients in recovery for 1:29 [ 1462.639656] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:417) [ 1462.655822] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:417) [ 1465.441720] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 4 sec [ 1475.127732] Lustre: DEBUG MARKER: == replay-dual test 12: open resend timeout ============== 18:18:21 (1786400301) [ 1484.572332] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1487.916860] Lustre: Failing over lustre-MDT0000 [ 1488.173899] Lustre: server umount lustre-MDT0000 complete [ 1488.370298] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1511.450148] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1511.455554] LDISKFS-fs (dm-0): recovery complete [ 1511.466569] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1518.576535] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf2f2663 [ 1518.742232] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (not set up) [ 1519.060754] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1519.082317] Lustre: Skipped 1 previous similar message [ 1523.998631] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1524.447286] Lustre: *** cfs_fail_loc=302, val=2147483648*** [ 1540.775819] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnected, waiting for 3 clients in recovery for 1:23 [ 1540.891812] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:449) [ 1540.912355] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:449) [ 1547.323347] Lustre: DEBUG MARKER: == replay-dual test 13: close resend timeout ============= 18:19:34 (1786400374) [ 1553.687518] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1556.796431] Lustre: Failing over lustre-MDT0000 [ 1557.074817] Lustre: server umount lustre-MDT0000 complete [ 1560.039687] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1560.055825] Lustre: Skipped 15 previous similar messages [ 1577.549890] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1577.555282] LDISKFS-fs (dm-0): recovery complete [ 1577.574044] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1586.151661] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97befd7d3100 x1873175059876736/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1591.055828] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1591.939394] Lustre: *** cfs_fail_loc=115, val=2147483648*** [ 1608.370932] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnected, waiting for 3 clients in recovery for 1:24 [ 1608.493318] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:99 to 0x280000401:481) [ 1608.493653] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:99 to 0x2c0000401:481) [ 1615.417620] Lustre: DEBUG MARKER: SKIP: replay-dual test_14b skipping ALWAYS excluded test 14b [ 1617.002599] Lustre: DEBUG MARKER: == replay-dual test 15a: timeout waiting for lost client during replay, 1 client completes ========================================================== 18:20:43 (1786400443) [ 1624.697402] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1627.266166] Lustre: Failing over lustre-MDT0000 [ 1627.554124] Lustre: server umount lustre-MDT0000 complete [ 1649.120244] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786400460/real 1786400460] req@ffff97bed930c000 x1873175059906944/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786400476 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1649.183336] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 1649.196598] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1649.211609] LustreError: Skipped 3 previous similar messages [ 1652.547054] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1652.549941] LDISKFS-fs (dm-0): recovery complete [ 1652.559586] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1659.366869] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc37c7480 x1873175059915776/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 1661.025885] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1661.039541] Lustre: Skipped 3 previous similar messages [ 1665.443577] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1731.500368] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1731.504360] Lustre: 41956:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client bea03574-ad0f-4362-8518-13471a7d242b@ [ 1731.513098] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1732.033129] Lustre: lustre-MDT0000: Recovery over after 1:11, of 3 clients 2 recovered and 1 was evicted. [ 1732.040553] Lustre: Skipped 3 previous similar messages [ 1732.063207] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:513) [ 1732.063568] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:513) [ 1738.577369] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1740.286287] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1750.952025] Lustre: DEBUG MARKER: == replay-dual test 15c: remove multiple OST orphans ===== 18:22:57 (1786400577) [ 1759.431449] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1859.607252] Lustre: Failing over lustre-MDT0000 [ 1860.032666] Lustre: server umount lustre-MDT0000 complete [ 1860.070380] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 1860.080934] LustreError: 7845:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1860.096389] LustreError: 7845:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 228 previous similar messages [ 1881.213954] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 1881.216380] LDISKFS-fs (dm-0): recovery complete [ 1881.222268] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1890.421668] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1890.438186] Lustre: Skipped 4 previous similar messages [ 1890.573473] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1890.580220] Lustre: Skipped 2 previous similar messages [ 1895.407837] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 1895.414929] Lustre: Skipped 21 previous similar messages [ 1895.928594] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 1960.500317] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 1960.510951] Lustre: 43977:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client ce36d17d-8883-46f8-be99-ca552aa654e5@ [ 1960.522665] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 1960.733061] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:494 to 0x280000401:1537) [ 1960.733246] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:495 to 0x2c0000401:1537) [ 1967.894589] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1969.383793] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1977.835977] Lustre: DEBUG MARKER: == replay-dual test 16: fail MDS during recovery (3571) == 18:26:44 (1786400804) [ 1984.933815] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 1987.951709] Lustre: Failing over lustre-MDT0000 [ 1988.248915] Lustre: server umount lustre-MDT0000 complete [ 1991.649142] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2010.306625] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2010.311211] LDISKFS-fs (dm-0): recovery complete [ 2010.319143] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2023.916813] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2048.459243] Lustre: Failing over lustre-MDT0000 [ 2048.474913] LustreError: 46423:0:(ldlm_lib.c:2990:target_stop_recovery_thread()) lustre-MDT0000: Aborting recovery [ 2048.482484] Lustre: 45948:0:(ldlm_lib.c:2390:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2048.487504] Lustre: 45948:0:(ldlm_lib.c:1900:abort_req_replay_queue()) @@@ aborted: req@ffff97bdc82fb800 x1873175041983744/t0(73014444033) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:668/0 lens 528/0 e 2 to 0 dl 1786400883 ref 1 fl Complete:/204/ffffffff rc 0/-1 job:'mcreate.0' uid:0 gid:0 projid:4294967295 [ 2048.499868] Lustre: lustre-MDT0000-osd: cancel update llog [0x200000400:0x1:0x0] [ 2048.520112] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (stopping) [ 2048.528454] Lustre: lustre-MDT0001-osp-MDT0000: cancel update llog [0x240000401:0x1:0x0] [ 2048.528715] LustreError: 45948:0:(client.c:1381:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff97bdc94ef800 x1873175060118144/t0(0) o1000->lustre-MDT0001-osp-MDT0000@0@lo:24/4 lens 336/33016 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'tgt_recover_0.0' uid:0 gid:0 projid:4294967295 [ 2048.528851] LustreError: 45948:0:(llog_osd.c:1178:llog_osd_next_block()) lustre-MDT0001-osp-MDT0000: can't read llog block from log [0x240000401:0x1:0x0] offset 32768: rc = -5 [ 2048.528870] LustreError: 45948:0:(llog.c:875:llog_process_thread()) lustre-MDT0001-osp-MDT0000 retry remote llog process [ 2048.530983] LustreError: 45948:0:(fid_request.c:217:seq_client_alloc_seq()) cli-cli-lustre-MDT0001-osp-MDT0000: Cannot allocate new meta-sequence: rc = -5 [ 2048.534787] Lustre: Skipped 1 previous similar message [ 2048.579542] LustreError: 45948:0:(fid_request.c:321:seq_client_alloc_fid()) cli-cli-lustre-MDT0001-osp-MDT0000: Can't allocate new sequence: rc = -5 [ 2048.790977] Lustre: server umount lustre-MDT0000 complete [ 2065.258672] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2069.099487] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2139.500173] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2139.504305] Lustre: 46877:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client a4aeb4ef-1510-41df-a72a-a743f2377928@ [ 2139.517399] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2140.381832] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1551 to 0x2c0000401:1569) [ 2140.382770] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1550 to 0x280000401:1569) [ 2147.243613] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2149.157856] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2159.469453] Lustre: DEBUG MARKER: == replay-dual test 17: fail OST during recovery (3571) == 18:29:46 (1786400986) [ 2168.160990] Lustre: DEBUG MARKER: ost1 REPLAY BARRIER on lustre-OST0000 [ 2170.051117] Lustre: Failing over lustre-OST0000 [ 2170.142715] Lustre: server umount lustre-OST0000 complete [ 2170.344111] LustreError: lustre-OST0000-osc-MDT0001: operation ost_statfs to node 0@lo failed: rc = -107 [ 2170.361377] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2170.383874] Lustre: Skipped 17 previous similar messages [ 2192.565845] LDISKFS-fs (dm-2): 3 truncates cleaned up [ 2192.570270] LDISKFS-fs (dm-2): recovery complete [ 2192.580205] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2193.565495] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 2193.580740] Lustre: Skipped 3 previous similar messages [ 2199.916902] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2224.494152] Lustre: Failing over lustre-OST0000 [ 2224.501797] LustreError: 49389:0:(ldlm_lib.c:2990:target_stop_recovery_thread()) lustre-OST0000: Aborting recovery [ 2224.507803] Lustre: 48826:0:(ldlm_lib.c:2390:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 2224.515134] Lustre: 48826:0:(ldlm_lib.c:2390:target_recovery_overseer()) Skipped 2 previous similar messages [ 2224.520163] LustreError: 48826:0:(ofd_obd.c:1324:ofd_iocontrol()) lustre-OST0000: iocontrol from 'tgt_recover_0' cmd=c00866c1 _IOWR('f', 193, 8) unrecognized: rc = -25 [ 2224.645133] Lustre: server umount lustre-OST0000 complete [ 2234.860144] Lustre: 3634:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786401023/real 1786401023] req@ffff97beedd92d80 x1873175060195200/t0(0) o400->lustre-OST0000-osc-MDT0001@0@lo:28/4 lens 224/224 e 2 to 1 dl 1786401063 ref 1 fl Rpc:XQr/2c0/ffffffff rc 0/-1 job:'ldlm_lock_repla.0' uid:0 gid:0 projid:4294967295 [ 2234.884040] Lustre: 3634:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 3 previous similar messages [ 2244.087193] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 2250.421994] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2314.500150] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 2314.504402] Lustre: 49829:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client b3e4e6ab-1d1c-414b-9f15-ab00dc598f62@ [ 2314.515829] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 2314.539472] Lustre: lustre-OST0000: Recovery over after 1:10, of 4 clients 3 recovered and 1 was evicted. [ 2314.555178] Lustre: Skipped 4 previous similar messages [ 2320.771722] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 2322.560863] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 2333.278908] Lustre: DEBUG MARKER: == replay-dual test 18: ldlm_handle_enqueue succeeds on evicted export (3822) ========================================================== 18:32:40 (1786401160) [ 2337.754510] LustreError: 6508:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b sleeping for 40000ms [ 2377.839123] LustreError: 6508:0:(ldlm_lockd.c:1361:ldlm_handle_enqueue()) cfs_fail_timeout id 30b awake [ 2394.434437] Lustre: DEBUG MARKER: == replay-dual test 19: resend of open request =========== 18:33:41 (1786401221) [ 2402.974915] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2404.680336] Lustre: *** cfs_fail_loc=157, val=2147483648*** [ 2404.684248] LustreError: 6508:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bef9aac380 x1873175042109568/t0(0) o101->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:345/0 lens 576/688 e 0 to 0 dl 1786401315 ref 1 fl Interpret:/600/0 rc 0/0 job:'createmany.0' uid:0 gid:0 projid:0 [ 2492.051823] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnecting [ 2495.426401] Lustre: Failing over lustre-MDT0000 [ 2497.177436] LustreError: 14134:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 2497.212814] LustreError: 14134:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 108 previous similar messages [ 2497.693323] Lustre: server umount lustre-MDT0000 complete [ 2499.050326] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2499.059025] LustreError: Skipped 1 previous similar message [ 2515.407282] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 2515.414556] LustreError: Skipped 3 previous similar messages [ 2520.740915] LDISKFS-fs (dm-0): 4 truncates cleaned up [ 2520.746413] LDISKFS-fs (dm-0): recovery complete [ 2520.762766] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2525.667152] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97befd7d5880 x1873175060338176/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2526.052496] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 2526.056519] Lustre: Skipped 4 previous similar messages [ 2526.129095] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 2526.145320] Lustre: Skipped 4 previous similar messages [ 2531.154199] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2531.324795] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 2531.338028] Lustre: Skipped 12 previous similar messages [ 2531.369080] Lustre: 52483:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2531.520486] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1601) [ 2531.521051] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1584 to 0x2c0000401:1601) [ 2540.776933] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2542.341775] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2549.794325] Lustre: DEBUG MARKER: == replay-dual test 20: recovery time is not increasing == 18:36:16 (1786401376) [ 2556.960910] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2559.000869] Lustre: Failing over lustre-MDT0000 [ 2559.227281] Lustre: server umount lustre-MDT0000 complete [ 2581.629324] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2581.632291] LDISKFS-fs (dm-0): recovery complete [ 2581.638614] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2588.644729] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97beee9bc700 x1873175060377088/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2593.464520] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2731.500520] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2731.506270] Lustre: 54419:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client dcebbba4-8e83-4a4b-a720-da359eae4521@ [ 2731.518839] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2731.669237] Lustre: 54419:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2731.686996] Lustre: 54419:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 6 previous similar messages [ 2731.882042] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1603 to 0x2c0000401:1633) [ 2731.882939] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1584 to 0x280000401:1633) [ 2737.559563] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2739.384754] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2749.583838] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2752.042742] Lustre: Failing over lustre-MDT0000 [ 2752.169096] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (stopping) [ 2752.461068] Lustre: server umount lustre-MDT0000 complete [ 2752.486937] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_statfs to node 0@lo failed: rc = -107 [ 2774.829252] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2774.832963] LDISKFS-fs (dm-0): recovery complete [ 2774.846222] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2782.689007] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc896ca80 x1873175060462336/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2782.804893] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (not set up) [ 2788.113785] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 2926.500449] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 2926.509229] Lustre: 56206:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client e0ca3c14-96b5-4415-9429-b2a00c4c538e@ [ 2926.523584] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 2926.689118] Lustre: 56206:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 2926.718513] Lustre: 56206:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 2926.886711] Lustre: lustre-MDT0000: Recovery over after 2:22, of 3 clients 2 recovered and 1 was evicted. [ 2926.905704] Lustre: Skipped 2 previous similar messages [ 2926.957993] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1665) [ 2926.965595] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1603 to 0x2c0000401:1665) [ 2933.489628] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 2935.908967] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 2947.273318] Lustre: DEBUG MARKER: == replay-dual test 21a: commit on sharing =============== 18:42:53 (1786401773) [ 2957.077146] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 2959.077318] Lustre: Failing over lustre-MDT0000 [ 2959.498727] Lustre: server umount lustre-MDT0000 complete [ 2962.406747] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 2962.418389] Lustre: Skipped 15 previous similar messages [ 2978.784540] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786401790/real 1786401790] req@ffff97bdc896ce00 x1873175060544128/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786401806 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2978.815402] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 2982.479791] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 2982.483540] LDISKFS-fs (dm-0): recovery complete [ 2982.497064] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 2989.026487] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bef00c0e00 x1873175060552704/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 2990.228396] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 2990.243864] Lustre: Skipped 4 previous similar messages [ 2994.833843] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3131.500165] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 3131.506327] Lustre: 58235:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client d373b887-ac23-4287-b9eb-d2bd15c91df9@ [ 3131.515711] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 3131.611769] Lustre: 58235:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3131.626368] Lustre: 58235:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 3131.652548] Lustre: lustre-MDT0000-osp-MDT0001: Connection restored to 0@lo (at 0@lo) [ 3131.657578] Lustre: Skipped 14 previous similar messages [ 3131.742909] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1697) [ 3131.751108] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1697) [ 3143.706540] Lustre: DEBUG MARKER: SKIP: replay-dual test_21b skipping SLOW test 21b [ 3146.534760] Lustre: DEBUG MARKER: == replay-dual test 22a: c1 lfs mkdir -i 1 dir1, M1 drop reply [ 3147.872896] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3147.877362] LustreError: 14901:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97beeddd1f80 x1873175042233216/t4294967341(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:333/0 lens 560/448 e 0 to 0 dl 1786402058 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3150.890323] Lustre: Failing over lustre-MDT0001 [ 3151.257660] Lustre: server umount lustre-MDT0001 complete [ 3153.078878] LustreError: 6509:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0001: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3153.107597] LustreError: 6509:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 147 previous similar messages [ 3153.386867] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 3172.868166] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3173.214360] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3173.226276] Lustre: Skipped 3 previous similar messages [ 3173.267949] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3173.277418] Lustre: Skipped 3 previous similar messages [ 3178.562591] Lustre: 14134:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97beeea59c00 x1873175042233216/t4294967341(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:363/0 lens 560/2880 e 0 to 0 dl 1786402088 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3179.158449] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3189.671992] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3192.092707] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3202.677488] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3205.441073] Lustre: Failing over lustre-MDT0000 [ 3205.859557] Lustre: server umount lustre-MDT0000 complete [ 3225.569218] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3225.593191] LustreError: Skipped 3 previous similar messages [ 3230.390454] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3230.400385] LDISKFS-fs (dm-0): recovery complete [ 3230.413675] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3234.822380] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf317470 [ 3234.835474] LustreError: 61280:0:(ldlm_resource.c:1207:ldlm_resource_complain()) MGC192.168.204.136@tcp: namespace resource [0x65727473756c:0x5:0x0].0x0 (ffff97bdc23b1000) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3240.154964] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3240.616432] Lustre: 61305:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3240.745816] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1729) [ 3240.746662] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1729) [ 3251.835771] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3254.523951] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3265.197408] Lustre: DEBUG MARKER: == replay-dual test 22b: c1 lfs mkdir -i 1 d1, M1 drop reply [ 3267.066575] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3267.071161] LustreError: 6507:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97beeea58a80 x1873175042275328/t8589934617(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:452/0 lens 560/448 e 0 to 0 dl 1786402177 ref 1 fl Interpret:/200/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3269.854210] Lustre: Failing over lustre-MDT0000 [ 3270.281233] Lustre: server umount lustre-MDT0000 complete [ 3274.991893] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786402103 with bad export cookie 1859884051132937328 [ 3274.999809] Lustre: Failing over lustre-MDT0001 [ 3275.012906] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 3275.059522] LustreError: 62448:0:(client.c:1381:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff97bdc5a33100 x1873175060706432/t0(0) o1000->lustre-MDT0000-osp-MDT0001@0@lo:24/4 lens 304/4320 e 0 to 0 dl 0 ref 2 fl Rpc:QU/200/ffffffff rc 0/-1 job:'umount.0' uid:0 gid:0 projid:4294967295 [ 3275.087758] LustreError: 62448:0:(client.c:1381:ptlrpc_import_delay_req()) Skipped 1 previous similar message [ 3275.096411] LustreError: 62448:0:(osp_object.c:618:osp_attr_get()) lustre-MDT0000-osp-MDT0001: osp_attr_get update error [0x200000401:0x1:0x0]: rc = -5 [ 3275.530234] Lustre: server umount lustre-MDT0001 complete [ 3296.998424] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3297.102256] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3297.212075] LustreError: 63140:0:(llog.c:1655:llog_backup()) MGC192.168.204.136@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3297.220495] Lustre: 63140:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.136@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3300.781287] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf317c1f [ 3305.380508] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3305.597541] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3307.475574] Lustre: 63163:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bef7a29880 x1873175042275328/t8589934617(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:492/0 lens 560/2880 e 0 to 0 dl 1786402217 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3307.490432] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:36 to 0x2c0000400:65) [ 3307.492584] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:36 to 0x280000400:65) [ 3312.380890] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1761) [ 3312.386757] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1761) [ 3317.545284] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3319.612363] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3321.367516] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3330.696151] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3333.241941] Lustre: Failing over lustre-MDT0000 [ 3333.750634] Lustre: server umount lustre-MDT0000 complete [ 3356.758495] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3356.762291] LDISKFS-fs (dm-0): recovery complete [ 3356.773288] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3364.332409] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf3184a7 [ 3364.527640] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (not set up) [ 3364.544063] Lustre: Skipped 1 previous similar message [ 3370.053697] Lustre: 65379:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3370.086070] Lustre: 65379:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 3370.242153] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1793) [ 3370.244031] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1793) [ 3370.850566] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3380.605636] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3383.091922] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3393.312636] Lustre: DEBUG MARKER: == replay-dual test 22c: c1 lfs mkdir -i 1 d1, M1 drop update [ 3394.921961] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3394.925380] LustreError: 8434:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bdc9329500 x1873175060786944/t107374182411(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:509/0 lens 2520/4320 e 0 to 0 dl 1786402234 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3398.331411] Lustre: Failing over lustre-MDT0000 [ 3398.622559] Lustre: server umount lustre-MDT0000 complete [ 3417.922943] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3426.284484] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bf009f4a80 x1873175060799232/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3431.767316] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3432.154717] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1825) [ 3432.158522] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1825) [ 3441.373365] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3443.088806] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3452.775167] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3455.379598] Lustre: Failing over lustre-MDT0000 [ 3455.635589] Lustre: server umount lustre-MDT0000 complete [ 3478.879072] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3478.882845] LDISKFS-fs (dm-0): recovery complete [ 3478.905795] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3484.346975] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (not set up) [ 3484.367400] Lustre: Skipped 1 previous similar message [ 3489.905192] Lustre: 68569:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3489.910213] Lustre: 68569:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 3490.030779] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1857) [ 3490.034475] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1857) [ 3491.003196] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3501.493425] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3503.610917] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3513.174845] Lustre: DEBUG MARKER: == replay-dual test 22d: c1 lfs mkdir -i 1 d1, M1 drop update [ 3518.388551] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3518.391670] LustreError: 63932:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bedfa96680 x1873175060867328/t115964117001(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:632/0 lens 2520/4320 e 0 to 0 dl 1786402357 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3522.577978] Lustre: Failing over lustre-MDT0000 [ 3523.019547] Lustre: server umount lustre-MDT0000 complete [ 3527.360203] LustreError: 6494:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786402355 with bad export cookie 1859884051132945000 [ 3527.368594] Lustre: Failing over lustre-MDT0001 [ 3527.378583] LustreError: 6494:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 2 previous similar messages [ 3527.405552] LustreError: 69807:0:(ldlm_resource.c:1207:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x79:0x0].0xf7117594 (ffff97bdc8ac2000) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 3527.465060] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.36@tcp (stopping) [ 3533.136872] Lustre: server umount lustre-MDT0001 complete [ 3554.546590] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3554.678720] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3573.681948] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_connect to node 0@lo failed: rc = -114 [ 3580.173307] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3580.963382] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3582.568712] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 3582.573156] Lustre: Skipped 8 previous similar messages [ 3582.649245] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1889) [ 3582.652946] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1889) [ 3597.078964] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:70 to 0x2c0000400:97) [ 3597.083794] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:70 to 0x280000400:97) [ 3597.154776] Lustre: 71215:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bdc896c000 x1873175042364800/t12884901939(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:26/0 lens 560/2880 e 0 to 0 dl 1786402506 ref 1 fl Interpret:/202/0 rc 0/0 job:'lfs.0' uid:0 gid:0 projid:4294967295 [ 3604.685526] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3607.017336] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3609.592812] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3620.607725] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3623.346836] Lustre: Failing over lustre-MDT0000 [ 3623.680340] Lustre: server umount lustre-MDT0000 complete [ 3623.911978] Lustre: lustre-MDT0000-lwp-OST0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3623.924692] Lustre: Skipped 34 previous similar messages [ 3645.407753] Lustre: 3635:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786402457/real 1786402457] req@ffff97bdc75d0e00 x1873175060918400/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786402473 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3645.426635] Lustre: 3635:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 19 previous similar messages [ 3646.475765] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3646.483256] LDISKFS-fs (dm-0): recovery complete [ 3646.509631] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3656.076674] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3656.101309] Lustre: Skipped 9 previous similar messages [ 3661.345109] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3661.428680] Lustre: 72767:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3661.454364] Lustre: 72767:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 3661.663541] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1635 to 0x280000401:1921) [ 3661.667349] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1667 to 0x2c0000401:1921) [ 3671.306370] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3673.681881] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3683.331440] Lustre: DEBUG MARKER: == replay-dual test 23a: c1 rmdir d1, M1 drop reply and fail, client2 mkdir d1 ========================================================== 18:55:10 (1786402510) [ 3685.123473] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3685.133444] LustreError: 71215:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97beeecf8700 x1873175042413056/t17179869210(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:113/0 lens 496/456 e 0 to 0 dl 1786402593 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3688.656563] Lustre: Failing over lustre-MDT0001 [ 3688.966287] Lustre: server umount lustre-MDT0001 complete [ 3709.881368] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3715.731769] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:129) [ 3715.732705] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:129) [ 3715.810038] Lustre: 72227:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bedfa97100 x1873175042413056/t17179869210(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:143/0 lens 496/2888 e 0 to 0 dl 1786402623 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3716.133598] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3726.658594] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3728.768965] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3739.827342] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3742.706903] Lustre: Failing over lustre-MDT0000 [ 3742.976106] Lustre: server umount lustre-MDT0000 complete [ 3756.518252] LustreError: 70536:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3756.546592] LustreError: 70536:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 277 previous similar messages [ 3766.049879] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3766.057332] LDISKFS-fs (dm-0): recovery complete [ 3766.077661] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3777.512070] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3777.520768] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to 0@lo (at 0@lo) [ 3777.530256] Lustre: Skipped 41 previous similar messages [ 3777.631955] Lustre: 75938:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 3777.659916] Lustre: 75938:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 4 previous similar messages [ 3777.928975] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1953) [ 3777.931646] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1953) [ 3788.487954] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3790.416886] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3800.807908] Lustre: DEBUG MARKER: == replay-dual test 23b: c1 rmdir d1, M1 drop reply and fail M0/M1, c2 mkdir d1 ========================================================== 18:57:07 (1786402627) [ 3802.556149] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 3802.567731] LustreError: 72227:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bef2e85c00 x1873175042448896/t21474836483(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:230/0 lens 496/456 e 0 to 0 dl 1786402710 ref 1 fl Interpret:/200/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3806.555752] Lustre: Failing over lustre-MDT0000 [ 3806.612561] LustreError: 6492:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786402634 with bad export cookie 1859884051132952322 [ 3806.624539] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 3806.627656] Lustre: Skipped 6 previous similar messages [ 3806.962510] Lustre: server umount lustre-MDT0000 complete [ 3811.375789] LustreError: 10326:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786402639 with bad export cookie 1859884051132952217 [ 3811.379319] Lustre: Failing over lustre-MDT0001 [ 3811.388600] LustreError: 10326:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 3811.716445] Lustre: server umount lustre-MDT0001 complete [ 3834.841610] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3834.926180] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3835.331263] LustreError: 77833:0:(llog.c:1655:llog_backup()) MGC192.168.204.136@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 3835.339894] Lustre: 77833:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.136@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 3835.837753] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf31b5fb [ 3836.205483] Lustre: lustre-MDT0001: Imperative Recovery not enabled, recovery window 60-180 [ 3836.212843] Lustre: Skipped 11 previous similar messages [ 3836.250763] Lustre: lustre-MDT0001: in recovery but waiting for the first client to connect [ 3836.256295] Lustre: Skipped 11 previous similar messages [ 3842.709349] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:161) [ 3842.714768] Lustre: 77863:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bef00c3100 x1873175042448896/t21474836483(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:270/0 lens 496/2888 e 0 to 0 dl 1786402750 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 3842.715131] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:161) [ 3843.184447] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3843.599943] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3849.385237] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1923 to 0x2c0000401:1985) [ 3849.390820] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1923 to 0x280000401:1985) [ 3856.135951] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 3858.301802] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3859.921352] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3870.356820] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3872.912703] Lustre: Failing over lustre-MDT0000 [ 3873.325292] Lustre: server umount lustre-MDT0000 complete [ 3889.056454] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 3889.072893] LustreError: Skipped 8 previous similar messages [ 3898.189850] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 3898.194152] LDISKFS-fs (dm-0): recovery complete [ 3898.203216] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3899.364932] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc9241180 x1873175061087104/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 3899.422445] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) Skipped 14 previous similar messages [ 3905.368091] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2017) [ 3905.368644] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2017) [ 3906.935496] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3916.921426] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3918.794972] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3931.453170] Lustre: DEBUG MARKER: == replay-dual test 23c: c1 rmdir d1, M0 drop update reply and fail M0, c2 mkdir d1 ========================================================== 18:59:18 (1786402758) [ 3933.506519] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 3933.511803] LustreError: 63932:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bf009f7100 x1873175061117184/t137438953491(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:292/0 lens 1984/4320 e 0 to 0 dl 1786402772 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 3936.774911] Lustre: Failing over lustre-MDT0000 [ 3937.130291] Lustre: server umount lustre-MDT0000 complete [ 3956.715362] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 3962.118626] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 3962.532354] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:1987 to 0x2c0000401:2049) [ 3962.556072] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:1987 to 0x280000401:2049) [ 3973.823557] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 3975.452597] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 3986.910678] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 3989.561586] Lustre: Failing over lustre-MDT0000 [ 3989.790406] Lustre: server umount lustre-MDT0000 complete [ 4013.903682] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4013.906126] LDISKFS-fs (dm-0): recovery complete [ 4013.922516] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4025.304855] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4025.412656] Lustre: 83244:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4025.431738] Lustre: 83244:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 17 previous similar messages [ 4025.662384] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2081) [ 4025.665352] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2081) [ 4036.208871] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4038.285264] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4048.794157] Lustre: DEBUG MARKER: == replay-dual test 23d: c1 rmdir d1, M0 drop update reply and fail M0/M1, c2 mkdir d1 ========================================================== 19:01:15 (1786402875) [ 4054.259556] Lustre: *** cfs_fail_loc=1701, val=2147483648*** [ 4054.263437] LustreError: 63932:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bdcd85aa00 x1873175061196928/t146028888081(0) o1000->lustre-MDT0001-mdtlov_UUID@0@lo:413/0 lens 1984/4320 e 0 to 0 dl 1786402893 ref 1 fl Interpret:/200/0 rc 0/0 job:'osp_up0-1.0' uid:0 gid:0 projid:4294967295 [ 4058.053595] Lustre: Failing over lustre-MDT0000 [ 4060.376700] Lustre: server umount lustre-MDT0000 complete [ 4065.390432] LustreError: 33574:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786402893 with bad export cookie 1859884051132959504 [ 4065.391486] Lustre: Failing over lustre-MDT0001 [ 4065.409020] LustreError: 33574:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 4065.437596] LustreError: 84493:0:(ldlm_resource.c:1207:ldlm_resource_complain()) lustre-MDT0000-osp-MDT0001: namespace resource [0x2000013a1:0x81:0x0].0x0 (ffff97bdc1b48300) refcount nonzero (1) after lock cleanup; forcing cleanup. [ 4065.510682] Lustre: lustre-MDT0001: Not available for connect from 192.168.204.36@tcp (stopping) [ 4071.492965] Lustre: server umount lustre-MDT0001 complete [ 4095.642223] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4095.748360] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4096.059765] LustreError: 85199:0:(llog.c:1655:llog_backup()) MGC192.168.204.136@tcp: failed to open log lustre-sptlrpc: rc = -108 [ 4096.066475] Lustre: 85199:0:(mgc_request_server.c:770:mgc_llog_local_copy()) MGC192.168.204.136@tcp: failed to copy new config lustre-sptlrpc: rc = -108 [ 4110.329315] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf31d29c [ 4117.785215] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2051 to 0x280000401:2113) [ 4117.791295] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2051 to 0x2c0000401:2113) [ 4117.987033] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:100 to 0x2c0000400:193) [ 4117.988739] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:100 to 0x280000400:193) [ 4118.121699] Lustre: 85213:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bec8754700 x1873175042528768/t25769803783(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:546/0 lens 496/2888 e 0 to 0 dl 1786403026 ref 1 fl Interpret:/202/0 rc 0/0 job:'rmdir.0' uid:0 gid:0 projid:4294967295 [ 4119.872690] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4121.247922] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4133.141572] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid,mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4136.312690] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4139.819876] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4152.908558] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4156.323226] Lustre: Failing over lustre-MDT0000 [ 4156.439920] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 4156.462523] Lustre: Skipped 6 previous similar messages [ 4156.889911] Lustre: server umount lustre-MDT0000 complete [ 4182.782158] LDISKFS-fs (dm-0): 2 truncates cleaned up [ 4182.786245] LDISKFS-fs (dm-0): recovery complete [ 4182.793811] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4193.081066] Lustre: lustre-MDT0000: Recovery over after 0:05, of 3 clients 3 recovered and 0 were evicted. [ 4193.099378] Lustre: Skipped 11 previous similar messages [ 4193.130572] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2115 to 0x2c0000401:2145) [ 4193.131732] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2115 to 0x280000401:2145) [ 4194.464863] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4206.736647] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4208.965509] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4220.860532] Lustre: DEBUG MARKER: == replay-dual test 24: reconstruct on non-existing object ========================================================== 19:04:07 (1786403047) [ 4222.424434] Lustre: *** cfs_fail_loc=119, val=2147483648*** [ 4222.433774] LustreError: 85756:0:(ldlm_lib.c:3332:target_send_reply_msg()) @@@ dropping reply req@ffff97bdc99cfb80 x1873175042573568/t154618822673(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:650/0 lens 488/456 e 0 to 0 dl 1786403130 ref 1 fl Interpret:/200/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4307.656211] Lustre: lustre-MDT0000: Client a258cce8-f80f-4edd-89d3-3746f39e6fe3 (at 192.168.204.36@tcp) reconnecting [ 4307.683953] Lustre: 85214:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97bdc7bcc700 x1873175042573568/t154618822673(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:735/0 lens 488/3152 e 0 to 0 dl 1786403215 ref 1 fl Interpret:/202/0 rc 0/0 job:'truncate.0' uid:0 gid:0 projid:4294967295 [ 4321.532887] Lustre: DEBUG MARKER: == replay-dual test 25: replay|resend ==================== 19:05:46 (1786403146) [ 4327.435091] Lustre: Failing over lustre-OST0000 [ 4327.647722] Lustre: server umount lustre-OST0000 complete [ 4328.422837] LustreError: lustre-OST0000-osc-MDT0001: operation ost_statfs to node 0@lo failed: rc = -107 [ 4328.431462] LustreError: Skipped 3 previous similar messages [ 4328.439687] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4328.468633] Lustre: Skipped 35 previous similar messages [ 4349.619179] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4351.672611] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 4 clients reconnect [ 4351.694132] Lustre: Skipped 10 previous similar messages [ 4359.094934] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4372.542705] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4375.415378] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4387.229995] Lustre: DEBUG MARKER: == replay-dual test 26: dbench and tar with mds failover ========================================================== 19:06:53 (1786403213) [ 4405.775396] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4411.265165] Lustre: DEBUG MARKER: test_26 fail mds1 1 times [ 4414.107185] Lustre: Failing over lustre-MDT0000 [ 4414.200564] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (stopping) [ 4414.220188] Lustre: Skipped 1 previous similar message [ 4414.533326] LustreError: 85213:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0000: not available for connect from 192.168.204.36@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 4414.543027] LustreError: 85213:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 217 previous similar messages [ 4414.605097] Lustre: server umount lustre-MDT0000 complete [ 4433.375141] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786403245/real 1786403245] req@ffff97bdc7bcea00 x1873175061420544/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786403261 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4433.411693] Lustre: 3636:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 19 previous similar messages [ 4440.831765] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4440.837602] LDISKFS-fs (dm-0): recovery complete [ 4440.849546] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4443.039157] LustreError: 91327:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 4443.052983] LustreError: 91327:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff97bdc94e1500 x1873175061426560/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1786403271 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4443.096473] LustreError: 91327:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 4444.659915] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf321f4f [ 4444.683192] Lustre: MGC192.168.204.136@tcp: Connection restored to 0@lo (at 0@lo) [ 4444.701956] Lustre: Skipped 35 previous similar messages [ 4445.245325] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 4445.255255] Lustre: Skipped 8 previous similar messages [ 4445.393733] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 4445.403021] Lustre: Skipped 8 previous similar messages [ 4450.339351] Lustre: 91364:0:(ldlm_lib.c:2071:extend_recovery_timer()) lustre-MDT0000: extended recovery timer reached hard limit: 180, extend: 1 [ 4450.358214] Lustre: 91364:0:(ldlm_lib.c:2071:extend_recovery_timer()) Skipped 13 previous similar messages [ 4451.301738] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4452.878840] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2171 to 0x280000401:2209) [ 4452.879748] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2170 to 0x2c0000401:2209) [ 4469.356393] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4472.757554] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4494.050292] Lustre: DEBUG MARKER: mds2 REPLAY BARRIER on lustre-MDT0001 [ 4499.891337] Lustre: DEBUG MARKER: test_26 fail mds2 2 times [ 4504.477391] Lustre: Failing over lustre-MDT0001 [ 4505.569932] Lustre: server umount lustre-MDT0001 complete [ 4535.389358] LDISKFS-fs (dm-1): 6 truncates cleaned up [ 4535.397597] LDISKFS-fs (dm-1): recovery complete [ 4535.412336] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4542.138378] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4544.981034] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:269 to 0x2c0000400:289) [ 4545.010858] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:270 to 0x280000400:289) [ 4559.941156] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 4563.220253] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4579.836375] Lustre: DEBUG MARKER: mds1 REPLAY BARRIER on lustre-MDT0000 [ 4584.233675] Lustre: DEBUG MARKER: test_26 fail mds1 3 times [ 4588.430614] Lustre: Failing over lustre-MDT0000 [ 4588.512502] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (stopping) [ 4588.523873] Lustre: Skipped 2 previous similar messages [ 4594.380277] Lustre: server umount lustre-MDT0000 complete [ 4613.093456] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 4613.122973] LustreError: Skipped 5 previous similar messages [ 4621.586754] LDISKFS-fs (dm-0): 3 truncates cleaned up [ 4621.592064] LDISKFS-fs (dm-0): recovery complete [ 4621.604917] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4622.879751] LustreError: 94913:0:(import.c:339:ptlrpc_invalidate_import()) MGS: timeout waiting for callback (1 != 0) [ 4622.892727] LustreError: 94913:0:(import.c:363:ptlrpc_invalidate_import()) @@@ still on sending list req@ffff97beedd91c00 x1873175061812480/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 1786403451 ref 1 fl Rpc:NQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 4622.916320] LustreError: 94913:0:(import.c:373:ptlrpc_invalidate_import()) MGS: Unregistering RPCs found (0). Network is sluggish? Waiting for them to error out. [ 4623.362479] Lustre: Evicted from MGS (at 0@lo) after server handle changed from 0x0 to 0x19cfa2b0bf33373a [ 4630.173860] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4631.588853] Lustre: 85849:0:(mdt_recovery.c:102:mdt_req_from_lrd()) @@@ restoring transno req@ffff97beedd90e00 x1873175044024832/t158913791517(0) o36->a258cce8-f80f-4edd-89d3-3746f39e6fe3@192.168.204.36@tcp:304/0 lens 488/3152 e 0 to 0 dl 1786403539 ref 1 fl Interpret:/202/0 rc 0/0 job:'tar.0' uid:0 gid:0 projid:4294967295 [ 4631.625879] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2302 to 0x2c0000401:2337) [ 4631.665296] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2301 to 0x280000401:2337) [ 4645.216699] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4648.646636] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4723.331020] Lustre: DEBUG MARKER: == replay-dual test 28: lock replay should be ordered: waiting after granted ========================================================== 19:12:29 (1786403549) [ 4734.926241] Lustre: Failing over lustre-OST0000 [ 4735.112434] Lustre: server umount lustre-OST0000 complete [ 4755.983574] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4758.038748] Lustre: *** cfs_fail_loc=32a, val=0*** [ 4764.775953] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4776.477520] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4778.967866] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL state after 0 sec [ 4793.690185] Lustre: DEBUG MARKER: == replay-dual test 29: replay vs update with the same xid ========================================================== 19:13:39 (1786403619) [ 4795.612565] Lustre: DEBUG MARKER: SKIP: replay-dual test_29 needs >= 2 clients [ 4797.834103] Lustre: DEBUG MARKER: == replay-dual test 30: layout lock replay is not blocked on IO ========================================================== 19:13:44 (1786403624) [ 4801.309675] Lustre: Failing over lustre-MDT0000 [ 4801.702881] Lustre: server umount lustre-MDT0000 complete [ 4822.419234] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4834.710068] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4836.618226] Lustre: lustre-MDT0000: Recovery over after 0:04, of 3 clients 3 recovered and 0 were evicted. [ 4836.627847] Lustre: Skipped 5 previous similar messages [ 4836.742066] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2370 to 0x280000401:2401) [ 4836.742602] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2370 to 0x2c0000401:2401) [ 4844.947503] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 4847.515751] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 4857.778661] Lustre: DEBUG MARKER: == replay-dual test 31: deadlock on file_remove_privs and occupied mod rpc slots ========================================================== 19:14:44 (1786403684) [ 4861.891412] Lustre: Failing over lustre-OST0000 [ 4862.119588] Lustre: server umount lustre-OST0000 complete [ 4882.432107] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: errors=remount-ro,no_mbcache,nodelalloc [ 4890.177697] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4900.947917] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 4902.926682] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid in FULL [ 4915.260572] Lustre: DEBUG MARKER: == replay-dual test 32: gap in update llog shouldn't break recovery ========================================================== 19:15:41 (1786403741) [ 4916.649127] Lustre: *** cfs_fail_loc=131d, val=10*** [ 4917.228817] Lustre: *** cfs_fail_loc=131d, val=4294967294*** [ 4917.231654] Lustre: Skipped 11 previous similar messages [ 4918.246845] Lustre: *** cfs_fail_loc=131d, val=4294967276*** [ 4918.250578] Lustre: Skipped 17 previous similar messages [ 4920.444506] Lustre: Failing over lustre-MDT0001 [ 4920.782461] Lustre: server umount lustre-MDT0001 complete [ 4924.420767] Lustre: Failing over lustre-MDT0000 [ 4924.622211] Lustre: lustre-MDT0000: Not available for connect from 192.168.204.36@tcp (stopping) [ 4924.636183] Lustre: Skipped 8 previous similar messages [ 4924.847438] Lustre: server umount lustre-MDT0000 complete [ 4933.029347] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4933.356276] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 4933.368565] Lustre: Skipped 23 previous similar messages [ 4933.369977] Lustre: *** cfs_fail_loc=131d, val=4294967266*** [ 4933.377611] Lustre: Skipped 9 previous similar messages [ 4938.189710] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4946.863930] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4947.064344] Lustre: *** cfs_fail_loc=131d, val=4294967262*** [ 4947.071614] Lustre: Skipped 3 previous similar messages [ 4952.636807] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:353) [ 4952.636919] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:353) [ 4952.800466] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2370 to 0x2c0000401:2433) [ 4952.802090] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2497) [ 4953.219407] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 4965.889765] Lustre: DEBUG MARKER: == replay-dual test 33: Check for OBD_INCOMPAT_MULTI_RPCS in last_rcvd after abort_recovery ========================================================== 19:16:32 (1786403792) [ 4973.302824] Lustre: Failing over lustre-MDT0001 [ 4973.571132] Lustre: server umount lustre-MDT0001 complete [ 4978.161083] LustreError: lustre-MDT0001-osp-MDT0000: operation mds_statfs to node 0@lo failed: rc = -107 [ 4978.184646] LustreError: Skipped 5 previous similar messages [ 4994.615664] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 4996.837921] Lustre: lustre-MDT0001: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 4996.854186] Lustre: Skipped 8 previous similar messages [ 5000.355453] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 5010.253130] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount REPLAY_WAIT mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5012.370638] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in REPLAY_WAIT state after 0 sec [ 5013.435135] Lustre: lustre-MDT0001: Aborting client recovery [ 5013.444994] LustreError: 103963:0:(ldlm_lib.c:2990:target_stop_recovery_thread()) lustre-MDT0001: Aborting recovery [ 5013.462751] Lustre: 103329:0:(ldlm_lib.c:2390:target_recovery_overseer()) recovery is aborted, evict exports in recovery [ 5013.476435] Lustre: 103329:0:(ldlm_lib.c:2390:target_recovery_overseer()) Skipped 2 previous similar messages [ 5013.488743] Lustre: 103329:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0001: disconnect stale client 3ebefe4e-5ad3-4dc8-b224-1b7eb2b8da5e@ [ 5013.502852] Lustre: lustre-MDT0001: disconnecting 1 stale clients [ 5013.523089] Lustre: lustre-MDT0001-osd: cancel update llog [0x240000400:0x1:0x0] [ 5013.541919] Lustre: lustre-MDT0000-osp-MDT0001: cancel update llog [0x200000401:0x1:0x0] [ 5013.631872] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:385) [ 5013.666899] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:385) [ 5020.121682] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5022.770688] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5026.478626] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 5033.651586] Lustre: Failing over lustre-MDT0001 [ 5034.101892] Lustre: server umount lustre-MDT0001 complete [ 5034.468295] LustreError: 8434:0:(ldlm_lib.c:1179:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 5034.496752] LustreError: 8434:0:(ldlm_lib.c:1179:target_handle_connect()) Skipped 214 previous similar messages [ 5042.415936] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5047.595418] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug -1 all [ 5048.314364] Lustre: lustre-MDT0001-lwp-OST0001: Connection restored to 0@lo (at 0@lo) [ 5048.330774] Lustre: Skipped 29 previous similar messages [ 5048.448752] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:328 to 0x2c0000400:417) [ 5048.451825] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:328 to 0x280000400:417) [ 5055.877974] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 5057.407323] Lustre: DEBUG MARKER: mdc.lustre-MDT0001-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5060.891854] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing _wait_recovery_complete *.lustre-MDT0001.recovery_status 1475 [ 5070.174790] Lustre: DEBUG MARKER: == replay-dual test complete, duration 4768 sec ========== 19:18:16 (1786403896) [ 5072.024132] Lustre: DEBUG MARKER: === replay-dual: start cleanup 19:18:18 (1786403898) === [ 5082.637181] Lustre: DEBUG MARKER: === replay-dual: finish cleanup 19:18:29 (1786403909) === [ 5085.104859] Lustre: Failing over lustre-MDT0000 [ 5085.462454] Lustre: server umount lustre-MDT0000 complete [ 5105.619688] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1786403917/real 1786403917] req@ffff97bdc734b480 x1873175062621568/t0(0) o400->MGC192.168.204.136@tcp@0@lo:26/25 lens 224/224 e 0 to 1 dl 1786403933 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 5105.658702] Lustre: 3637:0:(client.c:2490:ptlrpc_expire_one_request()) Skipped 4 previous similar messages [ 5114.358543] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 5114.912736] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff97bdc9329500 x1873175062629504/t0(0) o250->MGC192.168.204.136@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 5114.966986] LustreError: 3634:0:(client.c:1391:ptlrpc_import_delay_req()) Skipped 18 previous similar messages [ 5115.405863] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 5115.416082] Lustre: Skipped 9 previous similar messages [ 5115.466750] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 5115.477625] Lustre: Skipped 9 previous similar messages [ 5120.365559] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 5257.500568] Lustre: lustre-MDT0000: recovery is timed out, evict stale exports [ 5257.508764] Lustre: 107110:0:(genops.c:1622:class_disconnect_stale_exports()) lustre-MDT0000: disconnect stale client 3ebefe4e-5ad3-4dc8-b224-1b7eb2b8da5e@ [ 5257.525475] Lustre: lustre-MDT0000: disconnecting 1 stale clients [ 5257.575176] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:2370 to 0x2c0000401:2465) [ 5257.575239] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:2441 to 0x280000401:2529) [ 5264.436925] Lustre: DEBUG MARKER: oleg436-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 5266.531276] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 5278.960169] Lustre: server umount lustre-MDT0000 complete [ 5288.392108] LustreError: 33574:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1786404116 with bad export cookie 1859884051133151115 [ 5288.396520] LustreError: MGC192.168.204.136@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 5288.401704] LustreError: 33574:0:(ldlm_lockd.c:2564:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 5288.437164] LustreError: Skipped 3 previous similar messages [ 5288.844741] Lustre: server umount lustre-MDT0001 complete [ 5307.273680] Lustre: server umount lustre-OST0000 complete [ 5326.484921] Lustre: server umount lustre-OST0001 complete [ 5342.924524] Lustre: DEBUG MARKER: oleg436-server.virtnet: executing unload_modules_local [ 5346.792982] Key type lgssc unregistered [ 5347.111510] LNet: 110052:0:(lib-ptl.c:967:lnet_clear_lazy_portal()) Active lazy portal 0 on exit [ 5347.117585] LNetError: 110052:0:(acceptor.c:246:lnet_acceptor_remove_socket()) Interface ens2 not found [ 5347.131176] LNet: Removed LNI 192.168.204.136@tcp [ 5348.262216] Key type .llcrypt unregistered [ 5348.268301] Key type ._llcrypt unregistered