************************ crashinfo ************************* /exports/testreports/51305/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg435-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.0QaMf/vmlinux [TAINTED] DUMPFILE: /exports/testreports/51305/testresults/racer-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg435-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Thu May 1 17:04:26 EDT 2025 UPTIME: 00:33:38 LOAD AVERAGE: 0.00, 0.02, 0.32 TASKS: 312 NODENAME: oleg435-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2400 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 21 last 5s 59 last 60s 83 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 308 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 16470 kworker/0:0 0 ms (no user stack) 50 kworker/1:1 0 ms (no user stack) 3758 socknal_reaper 0 ms (no user stack) 3768 ptlrpcd_00_02 2 ms (no user stack) 16271 mdt_out00_005 2 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 681737 2.6 GB 71% of TOTAL MEM USED 273330 1 GB 28% of TOTAL MEM SHARED 27977 109.3 MB 2% of TOTAL MEM BUFFERS 25632 100.1 MB 2% of TOTAL MEM CACHED 97414 380.5 MB 10% of TOTAL MEM SLAB 21702 84.8 MB 2% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 59005 230.5 MB 7% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 7 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 2016.0 eth0 n/a 2.1 RSS_TOTAL=55672 pages, %mem= 0.9 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff8801376688c0 ffff88013767d800 sysfs sysfs /sys ffff880137668a80 ffff880139944000 proc proc /proc ffff880137668c40 ffff880137678000 devtmpfs devtmpfs /dev ffff880137668e00 ffff88013767d000 securityfs securityfs /sys/kernel/security ffff880137668fc0 ffff88013767e000 tmpfs tmpfs /dev/shm ffff880137669180 ffff88013771f800 devpts devpts /dev/pts ffff880137669340 ffff88013767e800 tmpfs tmpfs /run ffff880137669500 ffff88013767f000 tmpfs tmpfs /sys/fs/cgroup ffff8801376696c0 ffff88013767f800 cgroup cgroup /sys/fs/cgroup/systemd ffff880137669880 ffff88012a2d8000 pstore pstore /sys/fs/pstore ffff880137669a40 ffff88012a2da000 cgroup cgroup /sys/fs/cgroup/blkio ffff880137669c00 ffff88012a2d9800 cgroup cgroup /sys/fs/cgroup/pids ffff880137669dc0 ffff88012a2d9000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a328000 ffff88012a2d8800 cgroup cgroup /sys/fs/cgroup/devices ffff88012a3281c0 ffff88012a2da800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a328380 ffff88012a2db000 cgroup cgroup /sys/fs/cgroup/perf_event ffff88012a328540 ffff88012a2db800 cgroup cgroup /sys/fs/cgroup/memory ffff88012a328700 ffff88012a2dc000 cgroup cgroup /sys/fs/cgroup/cpuset ffff88012a3288c0 ffff88012a2dc800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a328a80 ffff88012a2dd000 cgroup cgroup /sys/fs/cgroup/freezer ffff8801377f6540 ffff8800b78b3800 configfs configfs /sys/kernel/config ffff8800b4fc61c0 ffff880137302000 ext4 /dev/nbd0 / ffff8800b4fc6380 ffff8800b5217000 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff88012a329c00 ffff8800b4d78800 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff8800b4fc6a80 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff8801377f68c0 ffff880129c02800 hugetlbfs hugetlbfs /dev/hugepages ffff88012a329dc0 ffff880137678800 mqueue mqueue /dev/mqueue ffff880137668540 ffff880129f39000 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff8801377f6a80 ffff8800b4d7c800 ramfs none /mnt ffff8800b4fc6c40 ffff8800b4073800 squashfs /dev/vda /home/green/git/lustre-release ffff8800b52641c0 ffff8800b4dbc800 tmpfs none /var/lib/stateless/writable ffff8800b5264540 ffff8800b4dbc800 tmpfs none /var/cache/man ffff8801377f6c40 ffff8800b4dbc800 tmpfs none /var/log ffff8800b4fc6e00 ffff8800b4dbc800 tmpfs none /var/lib/dbus ffff8800b5264700 ffff8800b4dbc800 tmpfs none /tmp ffff8800b4fc6fc0 ffff8800b4dbc800 tmpfs none /var/lib/dhclient ffff8801377f6e00 ffff8800b4dbc800 tmpfs none /var/tmp ffff8800b4fc7180 ffff8800b4dbc800 tmpfs none /var/lib/NetworkManager ffff8801377f6fc0 ffff8800b4dbc800 tmpfs none /var/lib/systemd/random-seed ffff8801377f7180 ffff8800b4dbc800 tmpfs none /var/spool ffff8800b52648c0 ffff8800b4dbc800 tmpfs none /var/lib/nfs ffff8800b5264a80 ffff8800b4dbc800 tmpfs none /var/lib/gssproxy ffff8800b4fc7340 ffff8800b4dbc800 tmpfs none /var/lib/logrotate ffff8800b5264c40 ffff8800b4dbc800 tmpfs none /etc ffff8800b4fc7500 ffff8800b4dbc800 tmpfs none /var/lib/rsyslog ffff8801377f7340 ffff8800b4dbc800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b4fc76c0 ffff8800b4fda800 nfs4 192.168.200.253:/exports/state/oleg435-server.virtnet /var/lib/stateless/state ffff8800b5264e00 ffff8800b4fda800 nfs4 192.168.200.253:/exports/state/oleg435-server.virtnet /boot ffff8800b4fc7880 ffff8800b4fda800 nfs4 192.168.200.253:/exports/state/oleg435-server.virtnet /etc/etc/kdump.conf ffff8800b4fc7a40 ffff8800b5217000 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b4fc7dc0 ffff8800b4073800 squashfs /dev/vda /usr/sbin/mount.lustre ffff880138ccbdc0 ffff880129c00800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 ffff8800a76441c0 ffff88007274a800 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 ffff8800a7645340 ffff880136a36000 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff8800a7644000 ffff88009e32b800 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 372.758160] Lustre: mdt_io00_008: service thread pid 16201 was inactive for 40.104 seconds. Watchdog stack traces are limited to 3 per 300 seconds, skipping this one. [ 372.763660] Lustre: Skipped 4 previous similar messages [ 374.294243] Lustre: mdt_io00_007: service thread pid 16200 was inactive for 40.091 seconds. Watchdog stack traces are limited to 3 per 300 seconds, skipping this one. [ 374.300845] Lustre: Skipped 10 previous similar messages [ 432.150097] LustreError: 7230:0:(ldlm_lockd.c:257:expired_lock_main()) ### lock callback timer expired after 100s: evicting client at 192.168.204.35@tcp ns: mdt-lustre-MDT0000_UUID lock: ffff8800933b4000/0xca7ac43800a6c321 lrc: 3/0,0 mode: PW/PW res: [0x200000403:0x2301:0x0].0x0 bits 0x4/0x0 rrc: 6 type: IBT gid 0 flags: 0x60200400000020 nid: 192.168.204.35@tcp remote: 0x1da02ae5202f5fe7 expref: 1112 pid: 16161 timeout: 431 lvb_type: 0 [ 432.175389] LustreError: 16165:0:(ldlm_lockd.c:1447:ldlm_handle_enqueue()) ### lock on destroyed export ffff88009740c800 ns: mdt-lustre-MDT0000_UUID lock: ffff88009185d200/0xca7ac43800a7279e lrc: 3/0,0 mode: PR/PR res: [0x200000403:0x2301:0x0].0x0 bits 0x1b/0x0 rrc: 4 type: IBT gid 0 flags: 0x50200000000000 nid: 192.168.204.35@tcp remote: 0x1da02ae5202f669a expref: 875 pid: 16165 timeout: 0 lvb_type: 0 [ 432.175693] Lustre: mdt_io00_009: service thread pid 16254 completed after 100.277s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.175695] Lustre: mdt00_004: service thread pid 14195 completed after 99.940s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.177102] Lustre: mdt00_001: service thread pid 7241 completed after 100.289s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.211693] Lustre: mdt00_007: service thread pid 16161 completed after 99.928s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.211834] Lustre: mdt00_008: service thread pid 16162 completed after 98.911s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.211989] Lustre: mdt00_006: service thread pid 16160 completed after 98.929s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.212139] Lustre: mdt00_010: service thread pid 16164 completed after 98.912s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.212273] Lustre: mdt00_012: service thread pid 16166 completed after 98.913s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.212389] Lustre: mdt00_013: service thread pid 16167 completed after 98.938s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.213492] Lustre: mdt00_020: service thread pid 16293 completed after 98.921s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.215494] Lustre: mdt00_015: service thread pid 16169 completed after 98.934s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.226067] LustreError: 16169:0:(client.c:1282:ptlrpc_import_delay_req()) @@@ IMP_CLOSED req@ffff880080ea0a80 x1830951587937024/t0(0) o104->lustre-MDT0000@192.168.204.35@tcp:15/16 lens 328/224 e 0 to 0 dl 0 ref 1 fl Rpc:QU/0/ffffffff rc 0/-1 job:'' uid:4294967295 gid:4294967295 [ 432.280581] LustreError: 16165:0:(ldlm_lockd.c:1447:ldlm_handle_enqueue()) Skipped 7 previous similar messages [ 432.284207] Lustre: mdt_io00_000: service thread pid 7253 completed after 100.315s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.298361] Lustre: mdt00_011: service thread pid 16165 completed after 100.163s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.314371] Lustre: mdt_io00_002: service thread pid 7255 completed after 100.339s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.323317] Lustre: mdt_io00_006: service thread pid 16191 completed after 100.316s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.330568] Lustre: mdt_io00_001: service thread pid 7254 completed after 99.686s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.354701] Lustre: mdt_io00_008: service thread pid 16201 completed after 99.700s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.359081] Lustre: mdt_io00_003: service thread pid 16181 completed after 99.499s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.379828] Lustre: mdt_io00_004: service thread pid 16183 completed after 99.123s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 432.382588] Lustre: mdt_io00_007: service thread pid 16200 completed after 98.179s. This likely indicates the system was overloaded (too many service threads, or not enough hardware resources). [ 433.324874] LustreError: 16293:0:(mdt_open.c:1703:mdt_reint_open()) lustre-MDT0000: name '5' present, but FID [0x200000402:0xdb:0x0] is invalid [ 433.335182] LustreError: 16293:0:(mdt_open.c:1703:mdt_reint_open()) Skipped 107 previous similar messages [ 434.045435] Lustre: 16183:0:(mdd_dir.c:4741:mdd_migrate_object()) lustre-MDD0000: [0x200000402:0x2:0x0]/9 is open, migrate only dentry [ 434.054331] Lustre: 16183:0:(mdd_dir.c:4741:mdd_migrate_object()) Skipped 14 previous similar messages [ 434.317203] LustreError: 16201:0:(mdt_reint.c:2523:mdt_reint_migrate()) lustre-MDT0000: migrate [0x200000402:0x1:0x0]/8 failed: rc = -2 [ 434.321263] LustreError: 16201:0:(mdt_reint.c:2523:mdt_reint_migrate()) Skipped 13 previous similar messages [ 437.629270] Lustre: 10932:0:(lod_lov.c:1417:lod_parse_striping()) lustre-MDT0000-mdtlov: EXTENSION flags=40 set on component[2]=1 of non-SEL file [0x200000402:0x1ea9:0x0] with magic=0xbd60bd0 [ 437.640407] Lustre: 10932:0:(lod_lov.c:1417:lod_parse_striping()) Skipped 101 previous similar messages [ 446.828491] LustreError: 16200:0:(mdd_dir.c:4662:mdd_migrate_cmd_check()) lustre-MDD0001: '11' migration was interrupted, run 'lfs migrate -m 0 -c 1 -H crush 11' to finish migration: rc = -1 [ 446.844108] LustreError: 16200:0:(mdd_dir.c:4662:mdd_migrate_cmd_check()) Skipped 7 previous similar messages [ 448.860702] LustreError: 16226:0:(mdt_xattr.c:402:mdt_dir_layout_update()) lustre-MDT0000: [0x200000402:0x2489:0x0] migrate mdt count mismatch 2 != 1 [ 448.876967] LustreError: 16226:0:(mdt_xattr.c:402:mdt_dir_layout_update()) Skipped 2 previous similar messages ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.74s (real) 6.39s (CPU), Child processes: 5.28s