************************ crashinfo ************************* /exports/testreports/47154/testresults/sanity-scrub-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg415-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.M6sWy/vmlinux [TAINTED] DUMPFILE: /exports/testreports/47154/testresults/sanity-scrub-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg415-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed Nov 13 13:14:13 EST 2024 UPTIME: 01:07:00 LOAD AVERAGE: 0.00, 0.01, 0.05 TASKS: 257 NODENAME: oleg415-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 19 last 5s 65 last 60s 77 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 253 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 34 rcuos/3 0 ms (no user stack) 906 in:imjournal 0 ms /usr/sbin/rsyslogd -n 1 systemd 0 ms /usr/lib/systemd/systemd --switched-root --system --deserialize 22 28 watchdog/3 54 ms (no user stack) 21 watchdog/2 61 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 735415 2.8 GB 77% of TOTAL MEM USED 219652 858 MB 22% of TOTAL MEM SHARED 10853 42.4 MB 1% of TOTAL MEM BUFFERS 9155 35.8 MB 0% of TOTAL MEM CACHED 72449 283 MB 7% of TOTAL MEM SLAB 16584 64.8 MB 1% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 60052 234.6 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 8 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- CLOSE 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 4017.7 eth0 n/a 3.7 RSS_TOTAL=51256 pages, %mem= 0.8 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff88012a2e4000 ffff880137123800 sysfs sysfs /sys ffff88012a2e41c0 ffff880139944000 proc proc /proc ffff88012a2e4380 ffff880137678000 devtmpfs devtmpfs /dev ffff88012a2e4540 ffff8800b5204800 securityfs securityfs /sys/kernel/security ffff88012a2e4700 ffff880137124000 tmpfs tmpfs /dev/shm ffff88012a2e48c0 ffff88013701e800 devpts devpts /dev/pts ffff88012a2e4a80 ffff880137124800 tmpfs tmpfs /run ffff88012a2e4c40 ffff880137125000 tmpfs tmpfs /sys/fs/cgroup ffff88012a2e4e00 ffff880137125800 cgroup cgroup /sys/fs/cgroup/systemd ffff88012a2e4fc0 ffff880137126000 pstore pstore /sys/fs/pstore ffff8801377f6fc0 ffff8800b526d800 cgroup cgroup /sys/fs/cgroup/blkio ffff8801377f7180 ffff8800b526d000 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff8801377f7340 ffff8800b526c800 cgroup cgroup /sys/fs/cgroup/freezer ffff8801377f7500 ffff8800b526c000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff8801377f76c0 ffff8800b526e000 cgroup cgroup /sys/fs/cgroup/memory ffff8801377f7880 ffff8800b526e800 cgroup cgroup /sys/fs/cgroup/perf_event ffff8801377f7a40 ffff8800b526f000 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff8801377f7c00 ffff8800b526f800 cgroup cgroup /sys/fs/cgroup/cpuset ffff8801377f7dc0 ffff88012a340000 cgroup cgroup /sys/fs/cgroup/pids ffff88012a32e000 ffff88012a340800 cgroup cgroup /sys/fs/cgroup/devices ffff88012a32e1c0 ffff88012a347000 configfs configfs /sys/kernel/config ffff880137668380 ffff88013767d000 ext4 /dev/nbd0 / ffff88012a32e540 ffff8800b5362000 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff880129c40540 ffff8800b5206800 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff880137668540 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff88012a32ec40 ffff8800b4e06000 hugetlbfs hugetlbfs /dev/hugepages ffff880129c40700 ffff88012b2c0000 mqueue mqueue /dev/mqueue ffff88012a32ee00 ffff8800b4e01800 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880129c408c0 ffff8800b5363000 ramfs none /mnt ffff880138ccae00 ffff88013734a800 tmpfs none /var/lib/stateless/writable ffff880137668700 ffff8800b41e3000 squashfs /dev/vda /home/green/git/lustre-release ffff880129c40c40 ffff88013734a800 tmpfs none /var/cache/man ffff88012a32f180 ffff88013734a800 tmpfs none /var/log ffff880129c40e00 ffff88013734a800 tmpfs none /var/lib/dbus ffff880129c40fc0 ffff88013734a800 tmpfs none /tmp ffff8801376688c0 ffff88013734a800 tmpfs none /var/lib/dhclient ffff880129c41180 ffff88013734a800 tmpfs none /var/tmp ffff880129c41340 ffff88013734a800 tmpfs none /var/lib/NetworkManager ffff88012a32f340 ffff88013734a800 tmpfs none /var/lib/systemd/random-seed ffff880137668a80 ffff88013734a800 tmpfs none /var/spool ffff880129c41500 ffff88013734a800 tmpfs none /var/lib/nfs ffff88012a32f500 ffff88013734a800 tmpfs none /var/lib/gssproxy ffff88012a32f6c0 ffff88013734a800 tmpfs none /var/lib/logrotate ffff880138ccafc0 ffff88013734a800 tmpfs none /etc ffff880137668c40 ffff88013734a800 tmpfs none /var/lib/rsyslog ffff880137668e00 ffff88013734a800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff880137668fc0 ffff8800b4283800 nfs4 192.168.200.253:/exports/state/oleg415-server.virtnet /var/lib/stateless/state ffff880137669340 ffff8800b4283800 nfs4 192.168.200.253:/exports/state/oleg415-server.virtnet /boot ffff880129c41880 ffff8800b4283800 nfs4 192.168.200.253:/exports/state/oleg415-server.virtnet /etc/etc/kdump.conf ffff880129c41a40 ffff8800b5362000 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff880138ccb880 ffff8800b41e3000 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800b1fcd880 ffff88006d5d3800 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff8800b1fcc000 ffff88012f6a1800 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 ffff8800b1ff2a80 ffff8800ace54800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 ffff8800b1fca380 ffff88012dd77000 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 680.912572] LustreError: 30787:0:(ldlm_lockd.c:2575:ldlm_cancel_handler()) Skipped 5 previous similar messages [ 681.002774] Lustre: server umount lustre-MDT0001 complete [ 682.385412] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: (null) [ 684.900414] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: errors=remount-ro [ 685.167957] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: (null) [ 687.618941] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: (null) [ 690.885335] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: errors=remount-ro [ 691.185124] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: (null) [ 694.477625] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,user_xattr,no_mbcache,nodelalloc [ 694.485649] Lustre: lustre-MDT0000: reset Object Index mappings [ 697.918917] Lustre: lustre-MDT0001-lwp-OST0001: Connection to lustre-MDT0001 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 697.927334] Lustre: Skipped 8 previous similar messages [ 705.930017] LustreError: 18866:0:(client.c:1310:ptlrpc_import_delay_req()) @@@ invalidate in flight req@ffff880130a0f100 x1815627990922496/t0(0) o250->MGC192.168.204.115@tcp@0@lo:26/25 lens 520/544 e 0 to 0 dl 0 ref 1 fl Rpc:NQU/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 [ 706.025290] LustreError: lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 706.032965] LustreError: Skipped 4 previous similar messages [ 706.042084] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 706.758589] Lustre: DEBUG MARKER: oleg415-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 708.035326] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to (at 0@lo) [ 708.038385] Lustre: Skipped 16 previous similar messages [ 709.182005] LDISKFS-fs (dm-1): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 709.247902] LustreError: lustre-MDT0000-osp-MDT0001: operation mds_connect to node 0@lo failed: rc = -114 [ 709.250993] LustreError: Skipped 1 previous similar message [ 709.270202] Lustre: lustre-OST0000: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x280000400:617 to 0x280000400:641) [ 709.270315] Lustre: lustre-OST0001: new connection from lustre-MDT0001-mdtlov (cleaning up unused objects from 0x2c0000400:618 to 0x2c0000400:641) [ 710.013296] Lustre: DEBUG MARKER: oleg415-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 710.995200] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 1 client reconnects [ 710.997318] Lustre: lustre-MDT0000: Denying connection for new client 9e47db43-c028-45a8-b3d9-991429024a94 (at 192.168.204.15@tcp), waiting for 1 known clients (0 recovered, 0 in progress, and 0 evicted) to recover in 0:59 [ 714.276634] Lustre: lustre-MDT0000: Recovery over after 0:03, of 1 clients 1 recovered and 0 were evicted. [ 714.290141] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:649 to 0x2c0000401:673) [ 714.293303] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:650 to 0x280000401:673) [ 717.306996] Lustre: lustre-MDT0000: trigger partial OI scrub for RPC inconsistency, checking FID [0x200005222:0x1:0x0]/32010: rc = 0 [ 719.331871] Lustre: lustre-MDT0001: trigger OI scrub by RPC for [0x240004a51:0x1:0x0]/32045 with flags 0x52: rc = 0 [ 778.882456] Lustre: DEBUG MARKER: == sanity-scrub test 4d: FID in LMA mismatch with object FID won't block create ========================================================== 12:20:10 (1731518410) [ 781.209229] Lustre: *** cfs_fail_loc=19b, val=0*** [ 781.211426] Lustre: Skipped 1 previous similar message [ 784.854278] Lustre: DEBUG MARKER: == sanity-scrub test 4e: FID reuse can be fixed ========== 12:20:16 (1731518416) [ 785.959668] Lustre: *** cfs_fail_loc=1a0, val=0*** [ 785.964138] Lustre: Skipped 455 previous similar messages [ 789.383315] Lustre: DEBUG MARKER: == sanity-scrub test 5: OI scrub state machine =========== 12:20:20 (1731518420) [ 819.439589] Lustre: lustre-MDT0001: haven't heard from client 9e47db43-c028-45a8-b3d9-991429024a94 (at 192.168.204.15@tcp) in 33 seconds. I think it's dead, and I am evicting it. exp ffff880086c2f800, cur 1731518451 expire 1731518421 last 1731518418 ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.49s (real) 6.23s (CPU), Child processes: 5.23s