************************ crashinfo ************************* /exports/testreports/49254/testresults/sanity-flr-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg425-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.ONXPz/vmlinux [TAINTED] DUMPFILE: /exports/testreports/49254/testresults/sanity-flr-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg425-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Fri Feb 14 07:45:58 EST 2025 UPTIME: 01:06:46 LOAD AVERAGE: 0.00, 0.02, 0.05 TASKS: 247 NODENAME: oleg425-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2400 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 32 last 5s 68 last 60s 82 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 243 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 830 kworker/3:2 0 ms (no user stack) 11 rcuos/0 0 ms (no user stack) 9 rcu_sched 0 ms (no user stack) 3778 ptlrpcd_00_01 0 ms (no user stack) 13229 ll_ost_create00 0 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 732798 2.8 GB 76% of TOTAL MEM USED 222269 868.2 MB 23% of TOTAL MEM SHARED 11184 43.7 MB 1% of TOTAL MEM BUFFERS 7647 29.9 MB 0% of TOTAL MEM CACHED 78358 306.1 MB 8% of TOTAL MEM SLAB 15458 60.4 MB 1% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 64998 253.9 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 10 LISTEN 3 NAGLE disabled (TCP_NODELAY): 8 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 18 LISTEN 8 Raw sockets info -------------------- CLOSE 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 4002.5 eth0 n/a 3.7 RSS_TOTAL=65644 pages, %mem= 1.0 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff880137668380 ffff8800b54d9800 sysfs sysfs /sys ffff880138ccaa80 ffff880139944000 proc proc /proc ffff880138ccac40 ffff880137678000 devtmpfs devtmpfs /dev ffff880138ccae00 ffff88012b3ac800 securityfs securityfs /sys/kernel/security ffff880138ccafc0 ffff8800b5203000 tmpfs tmpfs /dev/shm ffff880138ccb180 ffff88012b2c0000 devpts devpts /dev/pts ffff880138ccb340 ffff8800b5203800 tmpfs tmpfs /run ffff880138ccb500 ffff8800b5204000 tmpfs tmpfs /sys/fs/cgroup ffff880138ccb6c0 ffff8800b5204800 cgroup cgroup /sys/fs/cgroup/systemd ffff880138ccb880 ffff8800b5205000 pstore pstore /sys/fs/pstore ffff880138ccba40 ffff8800b5205800 cgroup cgroup /sys/fs/cgroup/pids ffff880138ccbc00 ffff8800b5206000 cgroup cgroup /sys/fs/cgroup/devices ffff880138ccbdc0 ffff8800b5206800 cgroup cgroup /sys/fs/cgroup/memory ffff88012a2b4000 ffff8800b5207000 cgroup cgroup /sys/fs/cgroup/cpuset ffff88012a2b41c0 ffff8800b5207800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a2b4380 ffff88012a368000 cgroup cgroup /sys/fs/cgroup/perf_event ffff88012a2b4540 ffff88012a368800 cgroup cgroup /sys/fs/cgroup/freezer ffff88012a2b4700 ffff88012a369000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a2b48c0 ffff88012a369800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff880137668540 ffff8800b54db800 cgroup cgroup /sys/fs/cgroup/blkio ffff8801376688c0 ffff8800b54df000 configfs configfs /sys/kernel/config ffff8801377f6c40 ffff8800b5233800 ext4 /dev/nbd0 / ffff8801377f6e00 ffff88012b3ae000 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff8801377f6fc0 ffff8800b4062800 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff880137669c00 ffff880137301000 mqueue mqueue /dev/mqueue ffff880137669dc0 ffff880129ee8000 hugetlbfs hugetlbfs /dev/hugepages ffff88012a2b4a80 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff8800b4164000 ffff880129eeb800 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff88012b2b6fc0 ffff8800b4d9d000 ramfs none /mnt ffff88012b2b7180 ffff8800b4d99800 tmpfs none /var/lib/stateless/writable ffff8800b41641c0 ffff8800b5237000 squashfs /dev/vda /home/green/git/lustre-release ffff8801377f7340 ffff8800b4d99800 tmpfs none /var/cache/man ffff88012b2b7340 ffff8800b4d99800 tmpfs none /var/log ffff8800b4164540 ffff8800b4d99800 tmpfs none /var/lib/dbus ffff8801377f7500 ffff8800b4d99800 tmpfs none /tmp ffff8801377f76c0 ffff8800b4d99800 tmpfs none /var/lib/dhclient ffff8800b4164700 ffff8800b4d99800 tmpfs none /var/tmp ffff88012a2b4c40 ffff8800b4d99800 tmpfs none /var/lib/NetworkManager ffff8800b41648c0 ffff8800b4d99800 tmpfs none /var/lib/systemd/random-seed ffff88012a2b4e00 ffff8800b4d99800 tmpfs none /var/spool ffff88012a2b4fc0 ffff8800b4d99800 tmpfs none /var/lib/nfs ffff88012a2b5180 ffff8800b4d99800 tmpfs none /var/lib/gssproxy ffff88012a2b5340 ffff8800b4d99800 tmpfs none /var/lib/logrotate ffff8800b4164a80 ffff8800b4d99800 tmpfs none /etc ffff8800b4164c40 ffff8800b4d99800 tmpfs none /var/lib/rsyslog ffff88012a2b5500 ffff8800b4d99800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff88012a2b56c0 ffff8800b4da9000 nfs4 192.168.200.253:/exports/state/oleg425-server.virtnet /var/lib/stateless/state ffff88012b2b7500 ffff8800b4da9000 nfs4 192.168.200.253:/exports/state/oleg425-server.virtnet /boot ffff8801377f7c00 ffff8800b4da9000 nfs4 192.168.200.253:/exports/state/oleg425-server.virtnet /etc/etc/kdump.conf ffff8800b4164e00 ffff88012b3ae000 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b1ec88c0 ffff8800b5237000 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800b1c71880 ffff88012b527800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 ffff8800b1ec8540 ffff88009e23d800 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 ffff8800b1c70fc0 ffff8800941ea000 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 ffff8800b1c701c0 ffff880095102800 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff8800b431b340 ffff88009f06c800 tmpfs tmpfs /run/user/0 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 1787.747472] Lustre: DEBUG MARKER: osc.lustre-OST0001-osc-ffff8800b599e000.ost_server_uuid in IDLE state after 0 sec [ 1787.790337] LustreError: 9604:0:(ldlm_lib.c:1094:target_handle_connect()) lustre-OST0001: not available for connect from 192.168.204.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1787.797385] LustreError: 9604:0:(ldlm_lib.c:1094:target_handle_connect()) Skipped 1 previous similar message [ 1790.770282] LustreError: 24250:0:(ldlm_lib.c:1094:target_handle_connect()) lustre-OST0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 1790.776759] LustreError: 24250:0:(ldlm_lib.c:1094:target_handle_connect()) Skipped 1 previous similar message [ 1791.005658] LDISKFS-fs (dm-3): file extents enabled, maximum tree depth=5 [ 1791.012356] LDISKFS-fs (dm-3): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 1791.047428] Lustre: 25223:0:(mgc_request_server.c:553:mgc_llog_local_copy()) MGC192.168.204.125@tcp: no remote llog for lustre-sptlrpc, check MGS config [ 1791.082706] Lustre: lustre-OST0001: Imperative Recovery enabled, recovery window shrunk from 60-180 down to 60-180 [ 1791.090492] Lustre: lustre-OST0001: in recovery but waiting for the first client to connect [ 1792.633667] Lustre: DEBUG MARKER: oleg425-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1792.725248] Lustre: lustre-OST0001: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 1792.731592] Lustre: lustre-OST0001: Recovery over after 0:01, of 2 clients 2 recovered and 0 were evicted. [ 1792.731705] Lustre: lustre-OST0001-osc-MDT0001: Connection restored to 192.168.204.125@tcp (at 0@lo) [ 1792.731707] Lustre: Skipped 1 previous similar message [ 1794.007446] Lustre: DEBUG MARKER: oleg425-server.virtnet: executing _wait_recovery_complete *.lustre-OST0001.recovery_status 1475 [ 3272.473154] Lustre: DEBUG MARKER: sanity-flr test_31: @@@@@@ FAIL: test_31 failed with 1 [ 3274.736884] Lustre: DEBUG MARKER: == sanity-flr test 32: data should be mirrored to newly created mirror ========================================================== 07:33:46 (1739536426) [ 3276.028393] Lustre: Failing over lustre-OST0000 [ 3276.059521] Lustre: server umount lustre-OST0000 complete [ 3277.090152] LustreError: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 3277.090162] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 3277.090390] LustreError: 9604:0:(ldlm_lib.c:1094:target_handle_connect()) lustre-OST0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3277.107665] LustreError: Skipped 1 previous similar message [ 3277.583310] Lustre: DEBUG MARKER: oleg425-client.virtnet: executing wait_import_state (DISCONN|IDLE) osc.lustre-OST0000-osc-ffff8800b599e000.ost_server_uuid 50 [ 3278.450243] LustreError: 9602:0:(ldlm_lib.c:1094:target_handle_connect()) lustre-OST0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3278.458457] LustreError: 9602:0:(ldlm_lib.c:1094:target_handle_connect()) Skipped 2 previous similar messages [ 3280.200612] LustreError: 9602:0:(ldlm_lib.c:1094:target_handle_connect()) lustre-OST0000: not available for connect from 192.168.204.25@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 3281.056058] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-ffff8800b599e000.ost_server_uuid in DISCONN state after 3 sec [ 3282.947786] LDISKFS-fs (dm-2): file extents enabled, maximum tree depth=5 [ 3282.952021] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 3282.976038] Lustre: 29533:0:(mgc_request_server.c:553:mgc_llog_local_copy()) MGC192.168.204.125@tcp: no remote llog for lustre-sptlrpc, check MGS config [ 3283.003318] Lustre: lustre-OST0000: Imperative Recovery enabled, recovery window shrunk from 60-180 down to 60-180 [ 3283.007803] Lustre: lustre-OST0000: in recovery but waiting for the first client to connect [ 3284.314294] Lustre: DEBUG MARKER: oleg425-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 3284.745041] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 3284.891708] Lustre: lustre-OST0000: Recovery over after 0:01, of 3 clients 3 recovered and 0 were evicted. [ 3284.891793] Lustre: lustre-OST0000-osc-MDT0001: Connection restored to 192.168.204.125@tcp (at 0@lo) [ 3284.891795] Lustre: Skipped 1 previous similar message [ 3285.429708] Lustre: DEBUG MARKER: oleg425-server.virtnet: executing _wait_recovery_complete *.lustre-OST0000.recovery_status 1475 ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 12.64s (real) 6.69s (CPU), Child processes: 5.91s