************************ crashinfo ************************* /exports/testreports/43549/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg132-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.MNejx/vmlinux [TAINTED] DUMPFILE: /exports/testreports/43549/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg132-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed Jun 19 23:04:14 EDT 2024 UPTIME: 00:16:59 LOAD AVERAGE: 0.00, 0.32, 0.64 TASKS: 332 NODENAME: oleg132-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 17 last 5s 71 last 60s 84 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 328 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 901 in:imjournal 0 ms /usr/sbin/rsyslogd -n 20 rcuos/1 0 ms (no user stack) 1 systemd 0 ms /usr/lib/systemd/systemd --switched-root --system --deserialize 22 4430 zthr_procedure 138 ms (no user stack) 4432 dbuf_evict 141 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 658451 2.5 GB 68% of TOTAL MEM USED 296616 1.1 GB 31% of TOTAL MEM SHARED 46451 181.4 MB 4% of TOTAL MEM BUFFERS 44141 172.4 MB 4% of TOTAL MEM CACHED 94554 369.4 MB 9% of TOTAL MEM SLAB 27235 106.4 MB 2% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 58983 230.4 MB 7% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 7 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 1017.1 eth0 n/a 3.3 RSS_TOTAL=55584 pages, %mem= 0.9 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff88012a2d4000 ffff88012a2d8000 sysfs sysfs /sys ffff88012a2d41c0 ffff880139944000 proc proc /proc ffff88012a2d4380 ffff880137678000 devtmpfs devtmpfs /dev ffff88012a2d4540 ffff8800b5239800 securityfs securityfs /sys/kernel/security ffff88012a2d4700 ffff88012a2d8800 tmpfs tmpfs /dev/shm ffff88012a2d48c0 ffff88012b230000 devpts devpts /dev/pts ffff88012a2d4a80 ffff88012a2d9000 tmpfs tmpfs /run ffff88012a2d4c40 ffff88012a2d9800 tmpfs tmpfs /sys/fs/cgroup ffff88012a2d4e00 ffff88012a2da000 cgroup cgroup /sys/fs/cgroup/systemd ffff88012a2d4fc0 ffff88012a2da800 pstore pstore /sys/fs/pstore ffff88012a2d5180 ffff88012a2dc800 cgroup cgroup /sys/fs/cgroup/pids ffff88012a2d5340 ffff88012a2dc000 cgroup cgroup /sys/fs/cgroup/freezer ffff88012a2d5500 ffff88012a2db800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a2d56c0 ffff88012a2db000 cgroup cgroup /sys/fs/cgroup/memory ffff88012a2d5880 ffff88012a2dd000 cgroup cgroup /sys/fs/cgroup/cpuset ffff88012a2d5a40 ffff88012a2dd800 cgroup cgroup /sys/fs/cgroup/blkio ffff88012a2d5c00 ffff88012a2de000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a2d5dc0 ffff88012a2de800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a294000 ffff88012a2df000 cgroup cgroup /sys/fs/cgroup/devices ffff88012a2941c0 ffff88012a2df800 cgroup cgroup /sys/fs/cgroup/perf_event ffff8800b4ec6000 ffff8800b53bf800 configfs configfs /sys/kernel/config ffff880137668540 ffff8800b523b000 ext4 /dev/nbd0 / ffff88012a38ae00 ffff8800b52b8800 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff880137668700 ffff880129f34800 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff8800b4ec68c0 ffff880129ce6000 hugetlbfs hugetlbfs /dev/hugepages ffff88012a294540 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff88012a294700 ffff88012b231800 mqueue mqueue /dev/mqueue ffff8800b4ec6a80 ffff8800b523a000 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880137668a80 ffff880129f37000 ramfs none /mnt ffff88012a2948c0 ffff880129dda000 tmpfs none /var/lib/stateless/writable ffff88012a38afc0 ffff880129ee6800 squashfs /dev/vda /home/green/git/lustre-release ffff880137668c40 ffff880129dda000 tmpfs none /var/cache/man ffff8800b4ec6c40 ffff880129dda000 tmpfs none /var/log ffff88012a294a80 ffff880129dda000 tmpfs none /var/lib/dbus ffff88012a38b340 ffff880129dda000 tmpfs none /tmp ffff88012a294c40 ffff880129dda000 tmpfs none /var/lib/dhclient ffff88012a294e00 ffff880129dda000 tmpfs none /var/tmp ffff880137668e00 ffff880129dda000 tmpfs none /var/lib/NetworkManager ffff88012a38b500 ffff880129dda000 tmpfs none /var/lib/systemd/random-seed ffff88012a294fc0 ffff880129dda000 tmpfs none /var/spool ffff88012a295180 ffff880129dda000 tmpfs none /var/lib/nfs ffff880137668fc0 ffff880129dda000 tmpfs none /var/lib/gssproxy ffff88012a38b6c0 ffff880129dda000 tmpfs none /var/lib/logrotate ffff880137669180 ffff880129dda000 tmpfs none /etc ffff88012a295340 ffff880129dda000 tmpfs none /var/lib/rsyslog ffff880137669340 ffff880129dda000 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b4ec6e00 ffff8800b41ad000 nfs4 192.168.200.253:/exports/state/oleg132-server.virtnet /var/lib/stateless/state ffff8800b4ec7180 ffff8800b41ad000 nfs4 192.168.200.253:/exports/state/oleg132-server.virtnet /boot ffff88012a295880 ffff8800b41ad000 nfs4 192.168.200.253:/exports/state/oleg132-server.virtnet /etc/etc/kdump.conf ffff88012a295a40 ffff8800b52b8800 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b4330700 ffff880129ee6800 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800b197efc0 ffff8800a7d2c800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 ffff8800b197fc00 ffff8800b1955000 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 ffff8800b197e380 ffff8800b1956800 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff8800b197e8c0 ffff88009a18b000 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 392.587200] [<0>] mdt_object_lock_internal+0x1a9/0x420 [mdt] [ 392.589330] [<0>] mdt_rename_lock+0xc3/0x2d0 [mdt] [ 392.590494] [<0>] mdt_reint_migrate+0x891/0x2420 [mdt] [ 392.591718] [<0>] mdt_reint_rec+0x87/0x240 [mdt] [ 392.593499] [<0>] mdt_reint_internal+0x74c/0xbc0 [mdt] [ 392.595110] [<0>] mdt_reint+0x67/0x150 [mdt] [ 392.596697] [<0>] tgt_request_handle+0x74e/0x1a50 [ptlrpc] [ 392.598318] [<0>] ptlrpc_server_handle_request+0x26c/0xcb0 [ptlrpc] [ 392.600470] [<0>] ptlrpc_main+0xc7e/0x1690 [ptlrpc] [ 392.601702] [<0>] kthread+0xe4/0xf0 [ 392.602500] [<0>] ret_from_fork_nospec_begin+0x7/0x21 [ 392.604339] [<0>] 0xfffffffffffffffe [ 392.964198] Lustre: mdt_io00_000: service thread pid 7130 was inactive for 40.037 seconds. Watchdog stack traces are limited to 3 per 300 seconds, skipping this one. [ 392.968170] Lustre: Skipped 16 previous similar messages [ 452.612268] LustreError: 7109:0:(ldlm_lockd.c:261:expired_lock_main()) ### lock callback timer expired after 101s: evicting client at 192.168.201.32@tcp ns: mdt-lustre-MDT0001_UUID lock: ffff88007ff70d80/0x7c0fcbfcf89b5b8 lrc: 3/0,0 mode: PR/PR res: [0x240000403:0x51db:0x0].0x0 bits 0x1b/0x0 rrc: 4 type: IBT gid 0 flags: 0x60200400000020 nid: 192.168.201.32@tcp remote: 0xcebfbb99101622d6 expref: 2549 pid: 7145 timeout: 451 lvb_type: 0 [ 452.647619] LustreError: 15406:0:(ldlm_lockd.c:1499:ldlm_handle_enqueue()) ### lock on destroyed export ffff880099c86800 ns: mdt-lustre-MDT0001_UUID lock: ffff880098cb8fc0/0x7c0fcbfcf89deed lrc: 3/0,0 mode: PR/PR res: [0x240000402:0x1:0x0].0x0 bits 0x12/0x0 rrc: 14 type: IBT gid 0 flags: 0x50200000000000 nid: 192.168.201.32@tcp remote: 0xcebfbb9910162ede expref: 1377 pid: 15406 timeout: 0 lvb_type: 0 [ 452.647713] LustreError: 15430:0:(mdt_reint.c:2519:mdt_reint_migrate()) lustre-MDT0001: migrate [0x240000402:0x1:0x0]/2 failed: rc = -114 [ 452.647717] LustreError: 15430:0:(mdt_reint.c:2519:mdt_reint_migrate()) Skipped 692 previous similar messages [ 452.647819] Lustre: mdt_io00_003: service thread pid 15430 completed after 100.346s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676690] Lustre: mdt00_012: service thread pid 15412 completed after 100.062s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676692] Lustre: mdt00_014: service thread pid 15415 completed after 100.050s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676697] Lustre: mdt00_016: service thread pid 15417 completed after 100.060s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676775] Lustre: mdt00_004: service thread pid 10340 completed after 100.029s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676841] Lustre: mdt00_003: service thread pid 7145 completed after 100.069s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.676875] Lustre: mdt00_013: service thread pid 15414 completed after 100.029s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.677265] Lustre: mdt00_027: service thread pid 15601 completed after 100.056s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.677475] Lustre: mdt00_006: service thread pid 15406 completed after 100.054s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.680366] Lustre: mdt00_021: service thread pid 15422 completed after 100.056s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.681170] Lustre: mdt_io00_004: service thread pid 15432 completed after 100.351s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.697939] Lustre: mdt_io00_002: service thread pid 7132 completed after 100.278s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.698830] Lustre: mdt_io00_011: service thread pid 15524 completed after 100.234s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.702096] Lustre: mdt_io00_008: service thread pid 15452 completed after 100.201s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.710578] Lustre: mdt_io00_010: service thread pid 15523 completed after 100.209s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.718415] Lustre: mdt_io00_007: service thread pid 15448 completed after 100.176s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.725486] Lustre: mdt_io00_009: service thread pid 15453 completed after 100.173s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.733997] Lustre: mdt_io00_006: service thread pid 15444 completed after 100.120s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.740034] Lustre: mdt_io00_012: service thread pid 15659 completed after 100.087s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.744963] Lustre: mdt_io00_013: service thread pid 15686 completed after 99.990s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.753925] Lustre: mdt_io00_005: service thread pid 15443 completed after 99.976s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 452.756714] Lustre: mdt_io00_000: service thread pid 7130 completed after 99.829s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.47s (real) 6.29s (CPU), Child processes: 5.15s