************************ crashinfo ************************* /exports/testreports/51653/testresults/sanity2-zfs-centos7_x86_64-centos7_x86_64/oleg224-client-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.GHHfU/vmlinux [TAINTED] DUMPFILE: /exports/testreports/51653/testresults/sanity2-zfs-centos7_x86_64-centos7_x86_64/oleg224-client-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed May 14 07:52:24 EDT 2025 UPTIME: 03:43:44 LOAD AVERAGE: 0.90, 1.04, 1.11 TASKS: 171 NODENAME: oleg224-client.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=30175 CPU=0 CMD=fio #-1 strrchr+0x10, 449 bytes of data #0 libcfs_debug_msg+0xc2 #1 cl_page_assume+0x164 #2 ll_read_ahead_page+0x423 #3 ll_read_ahead_pages+0x216 #4 ll_readahead+0x4b9 #5 ll_io_read_page+0x6c7 #6 ll_readpage+0xa7c #7 filemap_fault+0x205 #8 ll_filemap_fault+0x39 #9 vvp_io_fault_start+0x4de #10 cl_io_start+0x6d #11 cl_io_loop+0x9f #12 ll_fault+0x533 #13 __do_fault+0x84 #14 do_read_fault+0x50 #15 handle_pte_fault+0x2ef #16 __handle_mm_fault+0x31d #17 handle_mm_fault+0xc2 #18 __do_page_fault+0x1a0 #19 trace_do_page_fault+0x56 #20 do_async_page_fault+0x22 #21 async_page_fault+0x28, 477 bytes of data -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 29 last 5s 39 last 60s 46 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 166 TASK_RUNNING 2 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 7576 ptlrpcd_01_00 0 ms (no user stack) 9 rcu_sched 0 ms (no user stack) 11 rcuos/0 0 ms (no user stack) 392 kworker/1:1 0 ms (no user stack) 30175 fio 1 ms fio --name=read_test --ioengine=mmap --filename=/mnt/lustre/f835.sanity --rw=read --bs=1m +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955079 3.6 GB ---- FREE 307157 1.2 GB 32% of TOTAL MEM USED 647922 2.5 GB 67% of TOTAL MEM SHARED 524754 2 GB 54% of TOTAL MEM BUFFERS 1461 5.7 MB 0% of TOTAL MEM CACHED 547311 2.1 GB 57% of TOTAL MEM SLAB 37517 146.6 MB 3% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 194 776 KB 0% of TOTAL SWAP SWAP FREE 261949 1023.2 MB 99% of TOTAL SWAP COMMIT LIMIT 739682 2.8 GB ---- COMMITTED 364770 1.4 GB 49% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=fio ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 10 LISTEN 13 NAGLE disabled (TCP_NODELAY): 8 user_data set (NFS etc.): 12 UDP Connection Info ------------------- 15 UDP sockets, 0 in ESTABLISHED user_data set (NFS etc.): 4 Unix Connection Info ------------------------ ESTABLISHED 29 CLOSE 20 LISTEN 8 Raw sockets info -------------------- CLOSE 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 13421.6 eth0 n/a 0.0 RSS_TOTAL=336020 pages, %mem= 6.1 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff880138ccac40 ffff8800b6cc1000 sysfs sysfs /sys ffff880138ccae00 ffff880139944000 proc proc /proc ffff880138ccafc0 ffff880137678000 devtmpfs devtmpfs /dev ffff880138ccb180 ffff8800b6cc0800 securityfs securityfs /sys/kernel/security ffff880138ccb340 ffff8800b6cc1800 tmpfs tmpfs /dev/shm ffff880138ccb500 ffff880137301800 devpts devpts /dev/pts ffff880138ccb6c0 ffff8800b6cc2000 tmpfs tmpfs /run ffff880138ccb880 ffff8800b6cc2800 tmpfs tmpfs /sys/fs/cgroup ffff880137668380 ffff88013767c800 cgroup cgroup /sys/fs/cgroup/systemd ffff880137668540 ffff88013767d000 pstore pstore /sys/fs/pstore ffff880137668700 ffff88013767f000 cgroup cgroup /sys/fs/cgroup/pids ffff8801376688c0 ffff88013767e800 cgroup cgroup /sys/fs/cgroup/perf_event ffff880137668a80 ffff88013767e000 cgroup cgroup /sys/fs/cgroup/cpuset ffff880137668c40 ffff88013767d800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff880137668e00 ffff88013767f800 cgroup cgroup /sys/fs/cgroup/devices ffff880137668fc0 ffff88012aae8000 cgroup cgroup /sys/fs/cgroup/blkio ffff880137669180 ffff88012aae8800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff880137669340 ffff88012aae9000 cgroup cgroup /sys/fs/cgroup/freezer ffff880137669500 ffff88012aae9800 cgroup cgroup /sys/fs/cgroup/hugetlb ffff880138ccba40 ffff8800b6cc3000 cgroup cgroup /sys/fs/cgroup/memory ffff88012b2a0a80 ffff8800b60bd000 configfs configfs /sys/kernel/config ffff88012b2a1dc0 ffff8800b6295000 ext4 /dev/nbd0 / ffff880137669a40 ffff8800b6cc3800 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff8801377d6a80 ffff8800b6cc4000 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff8801377d6c40 ffff880137303000 mqueue mqueue /dev/mqueue ffff8801377d6e00 ffff8800b60bc000 hugetlbfs hugetlbfs /dev/hugepages ffff8801377d6fc0 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff88012b2a08c0 ffff880137306000 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880137669dc0 ffff8800b63c3000 ramfs none /mnt ffff8800b6e6c000 ffff8800b6f47800 tmpfs none /var/lib/stateless/writable ffff8801377d7180 ffff8800b60b9800 squashfs /dev/vda /home/green/git/lustre-release ffff8801377d7340 ffff8800b6f47800 tmpfs none /var/cache/man ffff8800b6e6c1c0 ffff8800b6f47800 tmpfs none /var/log ffff8800b3dd2380 ffff8800b6f47800 tmpfs none /var/lib/dbus ffff8801377d7500 ffff8800b6f47800 tmpfs none /tmp ffff8800b6e6c380 ffff8800b6f47800 tmpfs none /var/lib/dhclient ffff8800b6e6c540 ffff8800b6f47800 tmpfs none /var/tmp ffff8801377d76c0 ffff8800b6f47800 tmpfs none /var/lib/NetworkManager ffff880138ccbdc0 ffff8800b6f47800 tmpfs none /var/lib/systemd/random-seed ffff8800b6e6c700 ffff8800b6f47800 tmpfs none /var/spool ffff8801377d7880 ffff8800b6f47800 tmpfs none /var/lib/nfs ffff8800b6e6c8c0 ffff8800b6f47800 tmpfs none /var/lib/gssproxy ffff8800b6e6ca80 ffff8800b6f47800 tmpfs none /var/lib/logrotate ffff8800b6e6cc40 ffff8800b6f47800 tmpfs none /etc ffff880138ccbc00 ffff8800b6f47800 tmpfs none /var/lib/rsyslog ffff8800b3dd2540 ffff8800b6f47800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b3dd2700 ffff88012aaec000 nfs4 192.168.200.253:/exports/state/oleg224-client.virtnet /var/lib/stateless/state ffff8800b3dd2a80 ffff88012aaec000 nfs4 192.168.200.253:/exports/state/oleg224-client.virtnet /boot ffff8801377d7dc0 ffff88012aaec000 nfs4 192.168.200.253:/exports/state/oleg224-client.virtnet /etc/etc/kdump.conf ffff8800b3dd2c40 ffff8800b6cc3800 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b3dd3880 ffff8800ae407000 nfs4 192.168.200.253://exports/testreports/51653/testresults/sanity2-zfs-centos7_x86_64-centos7_x86_64 /tmp/tmp/testlogs ffff8800b3489a40 ffff8800ae581000 tmpfs tmpfs /run/user/0 ffff8801377d68c0 ffff8800b60b9800 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800b3dd21c0 ffff88013093d000 lustre 192.168.202.124@tcp:/lustre /mnt/lustre ffff8800b5803340 ffff88012a587800 nfsd nfsd /proc/fs/nfsd +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 9661.013066] Lustre: DEBUG MARKER: oleg224-client.virtnet: executing wait_import_state IDLE osc.lustre-OST0000-osc-ffff88013093d000.ost_server_uuid 50 [ 9675.672074] Lustre: DEBUG MARKER: osc.lustre-OST0000-osc-ffff88013093d000.ost_server_uuid in IDLE state after 14 sec [ 9677.491346] Lustre: DEBUG MARKER: == sanity test 817: nfsd won't cache write lock for exec file ========================================================== 06:49:56 (1747219796) [ 9677.601312] Installing knfsd (copyright (C) 1996 okir@monad.swb.de). [ 9677.734740] NFSD: starting 90-second grace period (net ffffffff81d41940) [ 9680.107963] Lustre: DEBUG MARKER: == sanity test 818: unlink with failed llog ============== 06:49:58 (1747219798) [ 9684.688581] Lustre: lustre-MDT0000-mdc-ffff88013093d000: Connection to lustre-MDT0000 (at 192.168.202.124@tcp) was lost; in progress operations using this service will wait for recovery to complete [ 9684.688972] LustreError: MGC192.168.202.124@tcp: Connection to MGS (at 192.168.202.124@tcp) was lost; in progress operations using this service will fail [ 9684.691970] Lustre: Evicted from MGS (at 192.168.202.124@tcp) after server handle changed from 0x532f1bda3f614ed4 to 0x532f1bda3f6ffb03 [ 9684.752712] LustreError: 7573:0:(client.c:3393:ptlrpc_replay_interpret()) @@@ status 301, old was 0 req@ffff880079c9fb80 x1832088102755584/t38654760518(38654760518) o101->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 576/608 e 0 to 0 dl 1747219819 ref 2 fl Interpret:RPQU/604/0 rc 301/301 job:'md5sum.0' uid:0 gid:0 projid:0 [ 9684.762229] LustreError: 7573:0:(client.c:3393:ptlrpc_replay_interpret()) Skipped 7 previous similar messages [ 9704.599960] LustreError: lustre-MDT0000-mdc-ffff88013093d000: operation ldlm_enqueue to node 192.168.202.124@tcp failed: rc = -107 [ 9704.607119] LustreError: Skipped 19 previous similar messages [ 9704.641266] LustreError: 7573:0:(client.c:3393:ptlrpc_replay_interpret()) @@@ status 301, old was 0 req@ffff880079c9fb80 x1832088102755584/t38654760518(38654760518) o101->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 576/608 e 0 to 0 dl 1747219839 ref 2 fl Interpret:RPQU/604/0 rc 301/301 job:'md5sum.0' uid:0 gid:0 projid:0 [ 9704.654725] LustreError: 7573:0:(client.c:3393:ptlrpc_replay_interpret()) Skipped 4 previous similar messages [ 9704.720083] LustreError: MGC192.168.202.124@tcp: Connection to MGS (at 192.168.202.124@tcp) was lost; in progress operations using this service will fail [ 9704.730066] Lustre: Evicted from MGS (at 192.168.202.124@tcp) after server handle changed from 0x532f1bda3f6ffb03 to 0x532f1bda3f700218 [ 9705.726605] Lustre: 7575:0:(client.c:2451:ptlrpc_expire_one_request()) @@@ Request sent has timed out for slow reply: [sent 1747219808/real 1747219808] req@ffff8801352fb480 x1832088102948992/t0(0) o400->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 224/224 e 0 to 1 dl 1747219824 ref 1 fl Rpc:XNQr/200/ffffffff rc 0/-1 job:'kworker.0' uid:0 gid:0 projid:4294967295 [ 9705.732048] Lustre: DEBUG MARKER: oleg224-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 9706.162397] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 9707.881957] Lustre: DEBUG MARKER: == sanity test 819a: too big niobuf in read ============== 06:50:26 (1747219826) [ 9710.147374] Lustre: DEBUG MARKER: == sanity test 819b: too big niobuf in write ============= 06:50:28 (1747219828) [ 9710.459463] LustreError: 7576:0:(osc_request.c:2439:osc_brw_redo_request()) @@@ redo for recoverable error -12 req@ffff880079c9c000 x1832088102957696/t0(0) o4->lustre-OST0001-osc-ffff88013093d000@192.168.202.124@tcp:6/4 lens 488/448 e 0 to 0 dl 1747219845 ref 2 fl Interpret:ReMQU/600/0 rc -12/-12 job:'dd.0' uid:0 gid:0 projid:0 [ 9710.469808] LustreError: 7576:0:(osc_request.c:2439:osc_brw_redo_request()) Skipped 9 previous similar messages [ 9714.699237] Lustre: DEBUG MARKER: == sanity test 820: update max EA from open intent ======= 06:50:33 (1747219833) [ 9715.128293] Lustre: DEBUG MARKER: SKIP: sanity test_820 needs >= 2 MDTs [ 9715.673085] Lustre: DEBUG MARKER: == sanity test 823: Setting create_count > OST_MAX_PRECREATE is lowered to maximum ========================================================== 06:50:34 (1747219834) [ 9717.483325] Lustre: DEBUG MARKER: setting create_count to 100200: [ 9717.884841] Lustre: DEBUG MARKER: -result- count: 9984 with max: 20000, expecting: 9984 [ 9720.760292] Lustre: DEBUG MARKER: == sanity test 831: throttling unlink/setattr queuing on OSP ========================================================== 06:50:39 (1747219839) [ 9734.778098] Lustre: 27963:0:(client.c:1609:after_reply()) @@@ resending request on EINPROGRESS req@ffff88007d056d80 x1832088103500928/t0(0) o36->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 488/456 e 0 to 0 dl 1747219869 ref 2 fl Rpc:RQU/202/0 rc 0/-115 job:'unlinkmany.0' uid:0 gid:0 projid:4294967295 [ 9741.031168] Lustre: 27963:0:(client.c:1609:after_reply()) @@@ resending request on EINPROGRESS req@ffff88013695d180 x1832088103527808/t0(0) o36->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 488/456 e 0 to 0 dl 1747219876 ref 2 fl Rpc:RQU/202/0 rc 0/-115 job:'unlinkmany.0' uid:0 gid:0 projid:4294967295 [ 9744.271620] Lustre: 27963:0:(client.c:1609:after_reply()) @@@ resending request on EINPROGRESS req@ffff88007d057480 x1832088103554176/t0(0) o36->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 488/456 e 0 to 0 dl 1747219879 ref 2 fl Rpc:RQU/202/0 rc 0/-115 job:'unlinkmany.0' uid:0 gid:0 projid:4294967295 [ 9750.452079] Lustre: 27963:0:(client.c:1609:after_reply()) @@@ resending request on EINPROGRESS req@ffff88013695dc00 x1832088103581056/t0(0) o36->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 488/456 e 0 to 0 dl 1747219885 ref 2 fl Rpc:RQU/202/0 rc 0/-115 job:'unlinkmany.0' uid:0 gid:0 projid:4294967295 [ 9759.700328] Lustre: 27963:0:(client.c:1609:after_reply()) @@@ resending request on EINPROGRESS req@ffff88013695ea00 x1832088103634304/t0(0) o36->lustre-MDT0000-mdc-ffff88013093d000@192.168.202.124@tcp:12/10 lens 488/456 e 0 to 0 dl 1747219894 ref 2 fl Rpc:RQU/202/0 rc 0/-115 job:'unlinkmany.0' uid:0 gid:0 projid:4294967295 [ 9759.708542] Lustre: 27963:0:(client.c:1609:after_reply()) Skipped 1 previous similar message [ 9783.841242] Lustre: DEBUG MARKER: == sanity test 832: lfs rm_entry ========================= 06:51:42 (1747219902) [ 9784.228579] Lustre: DEBUG MARKER: SKIP: sanity test_832 needs >= 2 MDTs [ 9784.681249] Lustre: DEBUG MARKER: == sanity test 833: Mixed buffered/direct read and write should not return -EIO ========================================================== 06:51:43 (1747219903) [ 9821.054248] Lustre: DEBUG MARKER: == sanity test 835: Setting read_ahead_kb explictly will be revised automatically ========================================================== 06:52:19 (1747219939) ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.34s (real) 5.88s (CPU), Child processes: 5.40s