************************ crashinfo ************************* /exports/testreports/44421/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg429-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.nqLj2/vmlinux [TAINTED] DUMPFILE: /exports/testreports/44421/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg429-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed Jul 24 23:52:34 EDT 2024 UPTIME: 00:16:59 LOAD AVERAGE: 0.00, 0.39, 0.65 TASKS: 323 NODENAME: oleg429-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 42 last 5s 73 last 60s 80 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 319 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 488 kworker/3:2 0 ms (no user stack) 915 in:imjournal 0 ms /usr/sbin/rsyslogd -n 7111 ldlm_bl_01 0 ms (no user stack) 17 kworker/1:0 0 ms (no user stack) 46 kworker/0:1 0 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 661836 2.5 GB 69% of TOTAL MEM USED 293231 1.1 GB 30% of TOTAL MEM SHARED 53220 207.9 MB 5% of TOTAL MEM BUFFERS 50877 198.7 MB 5% of TOTAL MEM CACHED 84259 329.1 MB 8% of TOTAL MEM SLAB 29473 115.1 MB 3% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 59007 230.5 MB 7% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 7 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 1016.7 eth0 n/a 0.9 RSS_TOTAL=52012 pages, %mem= 0.8 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff88012a2d8000 ffff88012a2a8000 sysfs sysfs /sys ffff88012a2d81c0 ffff880139944000 proc proc /proc ffff88012a2d8380 ffff880137678000 devtmpfs devtmpfs /dev ffff88012a2d8540 ffff880137678800 securityfs securityfs /sys/kernel/security ffff88012a2d8700 ffff88012a2a8800 tmpfs tmpfs /dev/shm ffff88012a2d88c0 ffff880137329000 devpts devpts /dev/pts ffff88012a2d8a80 ffff88012a2a9000 tmpfs tmpfs /run ffff88012a2d8c40 ffff88012a2a9800 tmpfs tmpfs /sys/fs/cgroup ffff88012a2d8e00 ffff88012a2aa000 cgroup cgroup /sys/fs/cgroup/systemd ffff88012a2d8fc0 ffff88012a2aa800 pstore pstore /sys/fs/pstore ffff88012a2d9180 ffff88012a2ac800 cgroup cgroup /sys/fs/cgroup/memory ffff88012a2d9340 ffff88012a2ac000 cgroup cgroup /sys/fs/cgroup/freezer ffff88012a2d9500 ffff88012a2ab800 cgroup cgroup /sys/fs/cgroup/blkio ffff88012a2d96c0 ffff88012a2ab000 cgroup cgroup /sys/fs/cgroup/cpuset ffff88012a2d9880 ffff88012a2ad000 cgroup cgroup /sys/fs/cgroup/pids ffff88012a2d9a40 ffff88012a2ad800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a2d9c00 ffff88012a2ae000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a2d9dc0 ffff88012a2ae800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a260000 ffff88012a2af000 cgroup cgroup /sys/fs/cgroup/devices ffff88012a2601c0 ffff88012a2af800 cgroup cgroup /sys/fs/cgroup/perf_event ffff880138ccbc00 ffff880129cc9000 configfs configfs /sys/kernel/config ffff8801376688c0 ffff88013767d000 ext4 /dev/nbd0 / ffff880137668a80 ffff88013767b800 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff88012a2608c0 ffff8800b51dd800 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff88012a260a80 ffff8800b51dc800 hugetlbfs hugetlbfs /dev/hugepages ffff880137668c40 ffff88013732a800 mqueue mqueue /dev/mqueue ffff8800b401e000 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff8800b401e1c0 ffff8800b4beb000 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880137668fc0 ffff8800b5231800 ramfs none /mnt ffff880137669180 ffff8800b417a800 tmpfs none /var/lib/stateless/writable ffff8800b401e380 ffff8800b4bcb800 squashfs /dev/vda /home/green/git/lustre-release ffff880137669340 ffff8800b417a800 tmpfs none /var/cache/man ffff880137669500 ffff8800b417a800 tmpfs none /var/log ffff8800b401e540 ffff8800b417a800 tmpfs none /var/lib/dbus ffff88012a260c40 ffff8800b417a800 tmpfs none /tmp ffff88012a260e00 ffff8800b417a800 tmpfs none /var/lib/dhclient ffff88012a260fc0 ffff8800b417a800 tmpfs none /var/tmp ffff88012a261180 ffff8800b417a800 tmpfs none /var/lib/NetworkManager ffff88012a261340 ffff8800b417a800 tmpfs none /var/lib/systemd/random-seed ffff8801376696c0 ffff8800b417a800 tmpfs none /var/spool ffff88012a261500 ffff8800b417a800 tmpfs none /var/lib/nfs ffff880137669880 ffff8800b417a800 tmpfs none /var/lib/gssproxy ffff880137669a40 ffff8800b417a800 tmpfs none /var/lib/logrotate ffff880137669c00 ffff8800b417a800 tmpfs none /etc ffff880137669dc0 ffff8800b417a800 tmpfs none /var/lib/rsyslog ffff880137668540 ffff8800b417a800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff88012a2616c0 ffff8800b4be9800 nfs4 192.168.200.253:/exports/state/oleg429-server.virtnet /var/lib/stateless/state ffff88012a261dc0 ffff8800b4be9800 nfs4 192.168.200.253:/exports/state/oleg429-server.virtnet /boot ffff88012a261c00 ffff8800b4be9800 nfs4 192.168.200.253:/exports/state/oleg429-server.virtnet /etc/etc/kdump.conf ffff88012a261a40 ffff88013767b800 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800ad401dc0 ffff8800b4bcb800 squashfs /dev/vda /usr/sbin/mount.lustre ffff880129e76700 ffff8800b417d800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 ffff880129e76e00 ffff88009814c000 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 ffff880129e77dc0 ffff880093b92000 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff880129e76a80 ffff880092df5000 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 384.012338] [] ? __llog_ctxt_put+0xdb/0x140 [obdclass] [ 384.016197] [] top_trans_start+0x351/0xab0 [ptlrpc] [ 384.020434] [] lod_trans_start+0x7f/0x2e0 [lod] [ 384.023073] [] mdd_trans_start+0x14/0x20 [mdd] [ 384.026329] [] mdd_migrate_object+0x141f/0x1dc0 [mdd] [ 384.030731] [] ? __wake_up_common_lock+0x91/0xc0 [ 384.034538] [] mdd_migrate+0x28/0x30 [mdd] [ 384.038382] [] mdt_reint_migrate+0x1b63/0x23f0 [mdt] [ 384.042737] [] mdt_reint_rec+0x87/0x240 [mdt] [ 384.045225] [] mdt_reint_internal+0x74c/0xbc0 [mdt] [ 384.049593] [] ? mdt_thread_info_init+0xae/0xd0 [mdt] [ 384.051991] [] mdt_reint+0x67/0x150 [mdt] [ 384.054199] [] tgt_request_handle+0x74e/0x1a50 [ptlrpc] [ 384.057474] [] ptlrpc_server_handle_request+0x273/0xcc0 [ptlrpc] [ 384.061279] [] ptlrpc_main+0xc7e/0x1690 [ptlrpc] [ 384.064043] [] ? put_prev_entity+0x31/0x400 [ 384.066566] [] ? do_raw_spin_unlock+0x49/0x90 [ 384.070736] [] ? ptlrpc_wait_event+0x630/0x630 [ptlrpc] [ 384.074680] [] kthread+0xe4/0xf0 [ 384.078225] [] ? kthread_create_on_node+0x140/0x140 [ 384.080689] [] ret_from_fork_nospec_begin+0x7/0x21 [ 384.083316] [] ? kthread_create_on_node+0x140/0x140 [ 386.282556] LustreError: 7133:0:(mdd_dir.c:4472:mdd_migrate_cmd_check()) lustre-MDD0000: '10' migration was interrupted, run 'lfs migrate -m 1 -c 1 -H crush 10' to finish migration: rc = -1 [ 386.291429] LustreError: 7133:0:(mdd_dir.c:4472:mdd_migrate_cmd_check()) Skipped 75 previous similar messages [ 387.399002] LustreError: 7121:0:(mdt_xattr.c:415:mdt_dir_layout_update()) lustre-MDT0000: [0x200000403:0x68e9:0x0] migrate mdt count mismatch 2 != 1 [ 399.546140] LustreError: 15427:0:(mdt_xattr.c:415:mdt_dir_layout_update()) lustre-MDT0001: [0x240000402:0x5c89:0x0] migrate mdt count mismatch 2 != 1 [ 399.549772] LustreError: 15427:0:(mdt_xattr.c:415:mdt_dir_layout_update()) Skipped 1 previous similar message [ 406.343927] Lustre: 15412:0:(lod_lov.c:1438:lod_parse_striping()) lustre-MDT0001-mdtlov: EXTENSION flags=40 set on component[2]=1 of non-SEL file [0x240000403:0x575f:0x0] with magic=0xbd60bd0 [ 406.349767] Lustre: 15412:0:(lod_lov.c:1438:lod_parse_striping()) Skipped 325 previous similar messages [ 433.671076] LustreError: 7113:0:(ldlm_lockd.c:261:expired_lock_main()) ### lock callback timer expired after 101s: evicting client at 192.168.204.29@tcp ns: mdt-lustre-MDT0001_UUID lock: ffff880090d20000/0xb3dd5d0bd64bed38 lrc: 3/0,0 mode: PR/PR res: [0x240000403:0x4a82:0x0].0x0 bits 0x1b/0x0 rrc: 8 type: IBT gid 0 flags: 0x60200400000020 nid: 192.168.204.29@tcp remote: 0xd176ee0f21c67a16 expref: 2870 pid: 7120 timeout: 432 lvb_type: 0 [ 433.705587] Lustre: mdt00_020: service thread pid 15419 completed after 100.186s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 433.705606] LustreError: 15404:0:(ldlm_lockd.c:1535:ldlm_handle_enqueue()) ### lock on destroyed export ffff8800b5231000 ns: mdt-lustre-MDT0001_UUID lock: ffff88008f4f2f40/0xb3dd5d0bd64bf0e9 lrc: 3/0,0 mode: PR/PR res: [0x240000403:0x4a82:0x0].0x0 bits 0x20/0x0 rrc: 5 type: IBT gid 0 flags: 0x50200000000000 nid: 192.168.204.29@tcp remote: 0xd176ee0f21c67aef expref: 1082 pid: 15404 timeout: 0 lvb_type: 0 [ 433.706648] Lustre: mdt00_005: service thread pid 15404 completed after 100.181s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 433.706836] Lustre: mdt00_010: service thread pid 15409 completed after 99.126s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 433.707842] Lustre: mdt00_011: service thread pid 15410 completed after 99.479s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 433.708234] Lustre: mdt00_012: service thread pid 15411 completed after 99.979s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 467.719149] LustreError: 7113:0:(ldlm_lockd.c:261:expired_lock_main()) ### lock callback timer expired after 100s: evicting client at 192.168.204.29@tcp ns: mdt-lustre-MDT0001_UUID lock: ffff88009b2e3cc0/0xb3dd5d0bd6637043 lrc: 3/0,0 mode: PR/PR res: [0x240000402:0x528a:0x0].0x0 bits 0x1b/0x0 rrc: 6 type: IBT gid 0 flags: 0x60200400000020 nid: 192.168.204.29@tcp remote: 0xd176ee0f21cd3bfd expref: 2908 pid: 15418 timeout: 467 lvb_type: 0 [ 467.759802] LustreError: 15406:0:(ldlm_lockd.c:1535:ldlm_handle_enqueue()) ### lock on destroyed export ffff8800b5231000 ns: mdt-lustre-MDT0001_UUID lock: ffff880099cf0b40/0xb3dd5d0bd6637cbb lrc: 3/0,0 mode: PR/PR res: [0x240000402:0x528a:0x0].0x0 bits 0x12/0x0 rrc: 2 type: IBT gid 0 flags: 0x50200000000000 nid: 192.168.204.29@tcp remote: 0xd176ee0f21cd3fca expref: 2 pid: 15406 timeout: 0 lvb_type: 0 [ 467.774330] LustreError: 15406:0:(ldlm_lockd.c:1535:ldlm_handle_enqueue()) Skipped 2 previous similar messages [ 467.778894] Lustre: 15406:0:(service.c:2359:ptlrpc_server_handle_request()) @@@ Request took longer than estimated (81/19s); client may timeout req@ffff880092906d80 x1805520724189568/t0(0) o101->9fdbbb53-553b-4e87-95c6-cc924bf59d32@192.168.204.29@tcp:313/0 lens 576/1032 e 0 to 0 dl 1721878983 ref 1 fl Complete:/200/0 rc -107/-107 job:'ln.0' uid:0 gid:0 ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.52s (real) 6.32s (CPU), Child processes: 5.13s