************************ crashinfo ************************* /exports/testreports/42551/testresults/sanity1-zfs-centos7_x86_64-centos7_x86_64/oleg303-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.fpBgt/vmlinux [TAINTED] DUMPFILE: /exports/testreports/42551/testresults/sanity1-zfs-centos7_x86_64-centos7_x86_64/oleg303-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed May 8 19:13:14 EDT 2024 UPTIME: 01:12:24 LOAD AVERAGE: 0.00, 0.01, 0.05 TASKS: 406 NODENAME: oleg303-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 23 last 5s 63 last 60s 74 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 402 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 9 rcu_sched 0 ms (no user stack) 909 in:imjournal 0 ms /usr/sbin/rsyslogd -n 3406 zthr_procedure 0 ms (no user stack) 3408 dbuf_evict 27 ms (no user stack) 20 rcuos/1 84 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 703357 2.7 GB 73% of TOTAL MEM USED 251710 983.2 MB 26% of TOTAL MEM SHARED 8821 34.5 MB 0% of TOTAL MEM BUFFERS 5096 19.9 MB 0% of TOTAL MEM CACHED 71550 279.5 MB 7% of TOTAL MEM SLAB 26655 104.1 MB 2% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 64477 251.9 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 7 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 4341.3 eth0 n/a 1.6 RSS_TOTAL=54540 pages, %mem= 0.9 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff88012a2d0000 ffff8801372f8800 sysfs sysfs /sys ffff88012a2d01c0 ffff880139944000 proc proc /proc ffff88012a2d0380 ffff880137678000 devtmpfs devtmpfs /dev ffff88012a2d0540 ffff8800b5211000 securityfs securityfs /sys/kernel/security ffff88012a2d0700 ffff8801372f9000 tmpfs tmpfs /dev/shm ffff88012a2d08c0 ffff88013771e800 devpts devpts /dev/pts ffff88012a2d0a80 ffff8801372f9800 tmpfs tmpfs /run ffff88012a2d0c40 ffff8801372fa000 tmpfs tmpfs /sys/fs/cgroup ffff88012a2d0e00 ffff8801372fa800 cgroup cgroup /sys/fs/cgroup/systemd ffff88012a2d0fc0 ffff8801372fb000 pstore pstore /sys/fs/pstore ffff8801377f6540 ffff88012a341800 cgroup cgroup /sys/fs/cgroup/freezer ffff8801377f6700 ffff88012a341000 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff8801377f68c0 ffff88012a340800 cgroup cgroup /sys/fs/cgroup/cpuset ffff8801377f6a80 ffff88012a340000 cgroup cgroup /sys/fs/cgroup/blkio ffff8801377f6c40 ffff88012a342000 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff8801377f6e00 ffff88012a342800 cgroup cgroup /sys/fs/cgroup/hugetlb ffff8801377f6fc0 ffff88012a343000 cgroup cgroup /sys/fs/cgroup/pids ffff8801377f7180 ffff88012a343800 cgroup cgroup /sys/fs/cgroup/devices ffff8801377f7340 ffff88012a344000 cgroup cgroup /sys/fs/cgroup/perf_event ffff8801377f7500 ffff88012a344800 cgroup cgroup /sys/fs/cgroup/memory ffff8801377f7880 ffff88012a346800 configfs configfs /sys/kernel/config ffff8801377f7dc0 ffff8800b4b8c800 ext4 /dev/nbd0 / ffff880137668380 ffff88012a346000 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff880129eb01c0 ffff8800b407c000 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff8800b53ac380 ffff8800b4980000 hugetlbfs hugetlbfs /dev/hugepages ffff880137668540 ffff88012b2b8000 mqueue mqueue /dev/mqueue ffff880137668700 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff880138ccb340 ffff88012a3c0800 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880138ccba40 ffff88012a3c0000 ramfs none /mnt ffff880137668a80 ffff8800b4079000 tmpfs none /var/lib/stateless/writable ffff8801376688c0 ffff8800b407e800 squashfs /dev/vda /home/green/git/lustre-release ffff880137668c40 ffff8800b4079000 tmpfs none /var/cache/man ffff880137668e00 ffff8800b4079000 tmpfs none /var/log ffff880137668fc0 ffff8800b4079000 tmpfs none /var/lib/dbus ffff880137669180 ffff8800b4079000 tmpfs none /tmp ffff880129eb0540 ffff8800b4079000 tmpfs none /var/lib/dhclient ffff880137669340 ffff8800b4079000 tmpfs none /var/tmp ffff8800b53ac540 ffff8800b4079000 tmpfs none /var/lib/NetworkManager ffff880137669500 ffff8800b4079000 tmpfs none /var/lib/systemd/random-seed ffff8800b53ac700 ffff8800b4079000 tmpfs none /var/spool ffff8801376696c0 ffff8800b4079000 tmpfs none /var/lib/nfs ffff880129eb0700 ffff8800b4079000 tmpfs none /var/lib/gssproxy ffff8800b53ac8c0 ffff8800b4079000 tmpfs none /var/lib/logrotate ffff880129eb08c0 ffff8800b4079000 tmpfs none /etc ffff880129eb0a80 ffff8800b4079000 tmpfs none /var/lib/rsyslog ffff880129eb0c40 ffff8800b4079000 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff880137669880 ffff8800b4001000 nfs4 192.168.200.253:/exports/state/oleg303-server.virtnet /var/lib/stateless/state ffff880129eb1180 ffff8800b4001000 nfs4 192.168.200.253:/exports/state/oleg303-server.virtnet /boot ffff880129eb1340 ffff8800b4001000 nfs4 192.168.200.253:/exports/state/oleg303-server.virtnet /etc/etc/kdump.conf ffff880137669a40 ffff88012a346000 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b1f81340 ffff8800b407e800 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800b1f801c0 ffff880094301000 lustre lustre-ost2/ost2 /mnt/lustre-ost2 ffff8800b4875c00 ffff88009d5af000 lustre lustre-mdt1/mdt1 /mnt/lustre-mds1 ffff8800b1d0c540 ffff8800905e6000 lustre lustre-ost1/ost1 /mnt/lustre-ost1 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 624.946273] Lustre: DEBUG MARKER: == sanity test 27g: /home/green/git/lustre-release/lustre/utils/lfs getstripe with no objects ========================================================== 18:11:14 (1715206274) [ 627.944028] Lustre: DEBUG MARKER: == sanity test 27ga: /home/green/git/lustre-release/lustre/utils/lfs getstripe with missing file (should return error) ========================================================== 18:11:17 (1715206277) [ 631.020360] Lustre: DEBUG MARKER: == sanity test 27i: /home/green/git/lustre-release/lustre/utils/lfs getstripe with some objects ========================================================== 18:11:20 (1715206280) [ 634.058237] Lustre: DEBUG MARKER: == sanity test 27j: setstripe with bad stripe offset (should return error) ========================================================== 18:11:23 (1715206283) [ 637.020189] Lustre: DEBUG MARKER: == sanity test 27k: limit i_blksize for broken user apps ========================================================== 18:11:26 (1715206286) [ 640.110968] Lustre: DEBUG MARKER: == sanity test 27l: check setstripe permissions (should return error) ========================================================== 18:11:29 (1715206289) [ 642.124987] Lustre: DEBUG MARKER: SKIP: sanity test_27m skipping SLOW test 27m [ 643.479460] Lustre: DEBUG MARKER: == sanity test 27n: create file with some full OSTs ====== 18:11:33 (1715206293) [ 654.460630] Lustre: *** cfs_fail_loc=215, val=0*** [ 671.191476] Lustre: DEBUG MARKER: == sanity test 27o: create file with all full OSTs (should error) ========================================================== 18:12:00 (1715206320) [ 680.123311] Lustre: *** cfs_fail_loc=215, val=0*** [ 681.403216] Lustre: *** cfs_fail_loc=215, val=1*** [ 685.131196] Lustre: *** cfs_fail_loc=215, val=0*** [ 690.139213] Lustre: *** cfs_fail_loc=215, val=0*** [ 690.140492] Lustre: Skipped 1 previous similar message [ 700.802191] Lustre: DEBUG MARKER: == sanity test 27oo: don't let few threads to reserve too many objects ========================================================== 18:12:30 (1715206350) [ 728.399498] Lustre: Failing over lustre-OST0000 [ 728.412442] Lustre: server umount lustre-OST0000 complete [ 728.544809] LustreError: 137-5: lustre-OST0000: not available for connect from 192.168.203.3@tcp (no target). If you are running an HA pair check that the target is mounted on the other server. [ 730.203132] LustreError: 11-0: lustre-OST0000-osc-MDT0000: operation ost_statfs to node 0@lo failed: rc = -107 [ 730.205606] Lustre: lustre-OST0000-osc-MDT0000: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 730.209682] LustreError: 137-5: lustre-OST0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 731.499216] LustreError: 137-5: lustre-OST0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 731.570925] Lustre: lustre-OST0000: Imperative Recovery enabled, recovery window shrunk from 60-180 down to 60-180 [ 731.576400] Lustre: lustre-OST0000: in recovery but waiting for the first client to connect [ 731.580677] mount.lustre (6392) used greatest stack depth: 10168 bytes left [ 732.840332] Lustre: DEBUG MARKER: oleg303-server.virtnet: executing set_default_debug all all [ 733.148090] Lustre: lustre-OST0000: Will be in recovery for at least 1:00, or until 2 clients reconnect [ 735.622019] Lustre: DEBUG MARKER: == sanity test 27p: append to a truncated file with some full OSTs ========================================================== 18:13:05 (1715206385) [ 802.555758] Lustre: lustre-OST0000: recovery is timed out, evict stale exports [ 802.558531] Lustre: 6421:0:(genops.c:1528:class_disconnect_stale_exports()) lustre-OST0000: disconnect stale client c18cb355-8184-45fe-b4f5-d4bf64b4fc23@192.168.203.3@tcp [ 802.562908] Lustre: lustre-OST0000: disconnecting 1 stale clients [ 802.575335] Lustre: 6421:0:(ldlm_lib.c:2874:target_recovery_thread()) too long recovery - read logs [ 802.575503] Lustre: lustre-OST0000-osc-MDT0000: Connection restored to 192.168.203.103@tcp (at 0@lo) [ 802.576727] Lustre: *** cfs_fail_loc=215, val=0*** [ 802.576729] Lustre: Skipped 1 previous similar message [ 802.584053] LustreError: dumping log to /tmp/lustre-log.1715206452.6421 [ 802.623656] Lustre: lustre-OST0000: Recovery over after 1:10, of 2 clients 1 recovered and 1 was evicted. [ 812.603203] Lustre: *** cfs_fail_loc=215, val=0*** [ 812.604679] Lustre: Skipped 4 previous similar messages ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 12.70s (real) 7.16s (CPU), Child processes: 5.53s