************************ crashinfo ************************* /exports/testreports/42582/testresults/conf-sanity3-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg139-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.mK99P/vmlinux [TAINTED] DUMPFILE: /exports/testreports/42582/testresults/conf-sanity3-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg139-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Fri May 10 10:03:19 EDT 2024 UPTIME: 01:56:56 LOAD AVERAGE: 0.11, 0.40, 0.49 TASKS: 158 NODENAME: oleg139-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 17 last 5s 25 last 60s 30 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 154 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 28 watchdog/3 0 ms (no user stack) 14 watchdog/1 0 ms (no user stack) 13 watchdog/0 0 ms (no user stack) 21 watchdog/2 0 ms (no user stack) 898 in:imjournal 53 ms /usr/sbin/rsyslogd -n +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 816214 3.1 GB 85% of TOTAL MEM USED 138853 542.4 MB 14% of TOTAL MEM SHARED 8837 34.5 MB 0% of TOTAL MEM BUFFERS 6631 25.9 MB 0% of TOTAL MEM CACHED 70984 277.3 MB 7% of TOTAL MEM SLAB 15671 61.2 MB 1% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 62205 243 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 6 LISTEN 3 NAGLE disabled (TCP_NODELAY): 5 user_data set (NFS etc.): 5 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 7013.5 eth0 n/a 1.5 RSS_TOTAL=52868 pages, %mem= 0.8 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff880137668380 ffff88013767f800 sysfs sysfs /sys ffff880137668540 ffff880139944000 proc proc /proc ffff880137668700 ffff880137678000 devtmpfs devtmpfs /dev ffff8801376688c0 ffff88013717a800 securityfs securityfs /sys/kernel/security ffff880137668a80 ffff88012a330000 tmpfs tmpfs /dev/shm ffff880137668c40 ffff88013771f000 devpts devpts /dev/pts ffff880137668e00 ffff88012a330800 tmpfs tmpfs /run ffff880137668fc0 ffff88012a331000 tmpfs tmpfs /sys/fs/cgroup ffff880137669180 ffff88012a331800 cgroup cgroup /sys/fs/cgroup/systemd ffff880137669340 ffff88012a332000 pstore pstore /sys/fs/pstore ffff880137669500 ffff88012a334000 cgroup cgroup /sys/fs/cgroup/devices ffff8801376696c0 ffff88012a333800 cgroup cgroup /sys/fs/cgroup/perf_event ffff880137669880 ffff88012a333000 cgroup cgroup /sys/fs/cgroup/blkio ffff880137669a40 ffff88012a332800 cgroup cgroup /sys/fs/cgroup/memory ffff880137669c00 ffff88012a334800 cgroup cgroup /sys/fs/cgroup/freezer ffff880137669dc0 ffff88012a335000 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a36e000 ffff88012a335800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a36e1c0 ffff88012a336000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a36e380 ffff88012a336800 cgroup cgroup /sys/fs/cgroup/pids ffff88012a36e540 ffff88012a337000 cgroup cgroup /sys/fs/cgroup/cpuset ffff8800b52fe1c0 ffff88013730b000 configfs configfs /sys/kernel/config ffff8800b52fe700 ffff88013730f000 ext4 /dev/nbd0 / ffff8800b52fe8c0 ffff88013730b800 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff8801377f6700 ffff8800b41b4000 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff88012a36e700 ffff88012b2c0800 mqueue mqueue /dev/mqueue ffff8800b52fee00 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff8801377f68c0 ffff8800b41b6800 hugetlbfs hugetlbfs /dev/hugepages ffff8801377f6a80 ffff8800b41b5000 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff8801377f6e00 ffff8800b41b7800 ramfs none /mnt ffff8800b52fefc0 ffff8800b538e800 tmpfs none /var/lib/stateless/writable ffff8800b52ff180 ffff8800b5388800 squashfs /dev/vda /home/green/git/lustre-release ffff8800b52ff340 ffff8800b538e800 tmpfs none /var/cache/man ffff8800b52ff500 ffff8800b538e800 tmpfs none /var/log ffff8800b52ff6c0 ffff8800b538e800 tmpfs none /var/lib/dbus ffff880129c7e8c0 ffff8800b538e800 tmpfs none /tmp ffff88012a36e8c0 ffff8800b538e800 tmpfs none /var/lib/dhclient ffff880129c7ea80 ffff8800b538e800 tmpfs none /var/tmp ffff88012a36ea80 ffff8800b538e800 tmpfs none /var/lib/NetworkManager ffff880129c7ec40 ffff8800b538e800 tmpfs none /var/lib/systemd/random-seed ffff880129c7ee00 ffff8800b538e800 tmpfs none /var/spool ffff8800b52ff880 ffff8800b538e800 tmpfs none /var/lib/nfs ffff8800b52ffa40 ffff8800b538e800 tmpfs none /var/lib/gssproxy ffff88012a36ec40 ffff8800b538e800 tmpfs none /var/lib/logrotate ffff88012a36ee00 ffff8800b538e800 tmpfs none /etc ffff88012a36efc0 ffff8800b538e800 tmpfs none /var/lib/rsyslog ffff88012a36f180 ffff8800b538e800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b52ffc00 ffff8800b538e000 nfs4 192.168.200.253:/exports/state/oleg139-server.virtnet /var/lib/stateless/state ffff8800b52fe000 ffff8800b538e000 nfs4 192.168.200.253:/exports/state/oleg139-server.virtnet /boot ffff880129c7efc0 ffff8800b538e000 nfs4 192.168.200.253:/exports/state/oleg139-server.virtnet /etc/etc/kdump.conf ffff8800b52fea80 ffff88013730b800 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff88012a36fdc0 ffff8800b5388800 squashfs /dev/vda /usr/sbin/mount.lustre +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 6850.205469] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 6850.767767] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing set_default_debug -1 all [ 6853.619777] Lustre: DEBUG MARKER: oleg139-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 6854.803134] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing wait_import_state FULL os[cp].lustre-OST0000-osc-MDT0000.ost_server_uuid 50 [ 6854.880013] Lustre: DEBUG MARKER: os[cp].lustre-OST0000-osc-MDT0000.ost_server_uuid in FULL state after 0 sec [ 6856.024974] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing wait_import_state FULL os[cp].lustre-OST0000-osc-MDT0001.ost_server_uuid 50 [ 6856.101214] Lustre: DEBUG MARKER: os[cp].lustre-OST0000-osc-MDT0001.ost_server_uuid in FULL state after 0 sec [ 6860.205214] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 6860.212736] LustreError: 11-0: lustre-OST0000-osc-MDT0001: operation ost_statfs to node 0@lo failed: rc = -107 [ 6860.212743] Lustre: lustre-OST0000: Not available for connect from 0@lo (stopping) [ 6860.212745] Lustre: Skipped 3 previous similar messages [ 6862.872380] Lustre: server umount lustre-OST0000 complete [ 6870.316265] Lustre: server umount lustre-MDT0000 complete [ 6871.687909] LustreError: 16540:0:(ldlm_lockd.c:2594:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1715349654 with bad export cookie 9281587237678100775 [ 6871.689860] LustreError: 166-1: MGC192.168.201.139@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 6871.699980] LustreError: 16540:0:(ldlm_lockd.c:2594:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 6871.825360] Lustre: server umount lustre-MDT0001 complete [ 6874.853804] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 6875.033298] LustreError: 137-5: lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 6875.037495] LustreError: Skipped 5 previous similar messages [ 6876.144466] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing set_default_debug -1 all [ 6879.188840] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 6880.434505] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing set_default_debug -1 all [ 6881.943628] Lustre: DEBUG MARKER: oleg139-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 6882.809886] Lustre: DEBUG MARKER: oleg139-client.virtnet: executing wait_import_state_mount FULL mdc.lustre-MDT0001-mdc-*.mds_server_uuid [ 6885.216575] LDISKFS-fs (dm-2): file extents enabled, maximum tree depth=5 [ 6885.222234] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 6887.091382] Lustre: DEBUG MARKER: oleg139-server.virtnet: executing set_default_debug -1 all [ 6888.597600] Lustre: DEBUG MARKER: oleg139-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 6896.924604] LustreError: 11-0: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 6896.929931] Lustre: lustre-OST0000: Not available for connect from 0@lo (stopping) [ 6896.933572] Lustre: Skipped 5 previous similar messages [ 6901.342758] Lustre: server umount lustre-OST0000 complete [ 6908.784646] Lustre: server umount lustre-MDT0000 complete [ 6908.940533] LustreError: 137-5: lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 6908.947781] LustreError: Skipped 1 previous similar message [ 6910.162074] LustreError: 20999:0:(ldlm_lockd.c:2594:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1715349692 with bad export cookie 9281587237678101629 [ 6910.163632] LustreError: 166-1: MGC192.168.201.139@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 6910.174235] LustreError: 20999:0:(ldlm_lockd.c:2594:ldlm_cancel_handler()) Skipped 3 previous similar messages [ 6910.301303] Lustre: server umount lustre-MDT0001 complete ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 10.59s (real) 5.47s (CPU), Child processes: 5.12s