************************ crashinfo ************************* /exports/testreports/51848/testresults/conf-sanity3-ldiskfs-DNE-centos7_x86_64-centos7_x86_64-retry3/oleg142-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.5td54/vmlinux [TAINTED] DUMPFILE: /exports/testreports/51848/testresults/conf-sanity3-ldiskfs-DNE-centos7_x86_64-centos7_x86_64-retry3/oleg142-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Wed May 21 15:48:04 EDT 2025 UPTIME: 02:30:18 LOAD AVERAGE: 1.00, 0.59, 0.59 TASKS: 206 NODENAME: oleg142-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 34 last 5s 96 last 60s 110 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 202 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 10154 socknal_reaper 0 ms (no user stack) 10159 monitor_thread 0 ms (no user stack) 907 in:imjournal 0 ms /usr/sbin/rsyslogd -n 4522 arc_reap 9 ms (no user stack) 4524 dbuf_evict 11 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 772103 2.9 GB 80% of TOTAL MEM USED 182964 714.7 MB 19% of TOTAL MEM SHARED 9729 38 MB 1% of TOTAL MEM BUFFERS 6980 27.3 MB 0% of TOTAL MEM CACHED 75583 295.2 MB 7% of TOTAL MEM SLAB 17122 66.9 MB 1% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 62988 246 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 10 LISTEN 3 NAGLE disabled (TCP_NODELAY): 8 user_data set (NFS etc.): 8 Unusual Situations: Doing Retransmission: 1 (run xportshow --retrans for details) UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 28 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 9015.6 eth0 n/a 0.1 RSS_TOTAL=58312 pages, %mem= 1.0 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff880137668380 ffff8800b5202000 sysfs sysfs /sys ffff880137668540 ffff880139944000 proc proc /proc ffff880137668700 ffff880137678000 devtmpfs devtmpfs /dev ffff8801376688c0 ffff8801372e5000 securityfs securityfs /sys/kernel/security ffff880137668a80 ffff8800b5202800 tmpfs tmpfs /dev/shm ffff880137668c40 ffff8801370f2800 devpts devpts /dev/pts ffff880137668e00 ffff8800b5203000 tmpfs tmpfs /run ffff880137668fc0 ffff8800b5203800 tmpfs tmpfs /sys/fs/cgroup ffff880137669180 ffff8800b5204000 cgroup cgroup /sys/fs/cgroup/systemd ffff880137669340 ffff8800b5204800 pstore pstore /sys/fs/pstore ffff880137669500 ffff8800b5206800 cgroup cgroup /sys/fs/cgroup/cpuset ffff8801376696c0 ffff8800b5206000 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff880137669880 ffff8800b5205800 cgroup cgroup /sys/fs/cgroup/hugetlb ffff880137669a40 ffff8800b5205000 cgroup cgroup /sys/fs/cgroup/blkio ffff880137669c00 ffff8800b5207000 cgroup cgroup /sys/fs/cgroup/pids ffff880137669dc0 ffff8800b5207800 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a30a000 ffff88012a378000 cgroup cgroup /sys/fs/cgroup/perf_event ffff88012a30a1c0 ffff88012a378800 cgroup cgroup /sys/fs/cgroup/memory ffff88012a30a380 ffff88012a379000 cgroup cgroup /sys/fs/cgroup/devices ffff88012a30a540 ffff88012a379800 cgroup cgroup /sys/fs/cgroup/freezer ffff880138ccba40 ffff8800b4d57000 configfs configfs /sys/kernel/config ffff88012a30a8c0 ffff88012a37b800 ext4 /dev/nbd0 / ffff880138ccbc00 ffff8800b5234000 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff88012a30aa80 ffff8800b53eb000 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff88012a30ac40 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff880138ccb880 ffff880129cb3800 hugetlbfs hugetlbfs /dev/hugepages ffff8800b52096c0 ffff8801370f4000 mqueue mqueue /dev/mqueue ffff88012a30e540 ffff8800b41f9800 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff880138ccbdc0 ffff880129cb4000 ramfs none /mnt ffff8800b32ca000 ffff8800b4073000 tmpfs none /var/lib/stateless/writable ffff88012a30e700 ffff88013734a000 squashfs /dev/vda /home/green/git/lustre-release ffff8800b5209a40 ffff8800b4073000 tmpfs none /var/cache/man ffff8800b32ca1c0 ffff8800b4073000 tmpfs none /var/log ffff88012a30afc0 ffff8800b4073000 tmpfs none /var/lib/dbus ffff8800b32ca380 ffff8800b4073000 tmpfs none /tmp ffff88012a30b180 ffff8800b4073000 tmpfs none /var/lib/dhclient ffff8800b32ca540 ffff8800b4073000 tmpfs none /var/tmp ffff88012a30b340 ffff8800b4073000 tmpfs none /var/lib/NetworkManager ffff88012a30b500 ffff8800b4073000 tmpfs none /var/lib/systemd/random-seed ffff8800b32ca700 ffff8800b4073000 tmpfs none /var/spool ffff88012a30b6c0 ffff8800b4073000 tmpfs none /var/lib/nfs ffff88012a30b880 ffff8800b4073000 tmpfs none /var/lib/gssproxy ffff8800b32ca8c0 ffff8800b4073000 tmpfs none /var/lib/logrotate ffff88012a30ba40 ffff8800b4073000 tmpfs none /etc ffff8800b5209c00 ffff8800b4073000 tmpfs none /var/lib/rsyslog ffff8800b5209dc0 ffff8800b4073000 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b5209500 ffff8800b4e2b800 nfs4 192.168.200.253:/exports/state/oleg142-server.virtnet /var/lib/stateless/state ffff8800b5209180 ffff8800b4e2b800 nfs4 192.168.200.253:/exports/state/oleg142-server.virtnet /boot ffff8800b5208000 ffff8800b4e2b800 nfs4 192.168.200.253:/exports/state/oleg142-server.virtnet /etc/etc/kdump.conf ffff8800b5209340 ffff8800b5234000 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff880136951880 ffff88013734a000 squashfs /dev/vda /usr/sbin/mount.lustre ffff880136950a80 ffff8800aa03e000 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 8983.122334] LDISKFS-fs (dm-2): file extents enabled, maximum tree depth=5 [ 8983.127653] LDISKFS-fs (dm-2): mounted filesystem with ordered data mode. Opts: user_xattr,acl,no_mbcache,nodelalloc [ 8983.250047] Lustre: lustre-OST0000: new disk, initializing [ 8983.252616] Lustre: Skipped 1 previous similar message [ 8983.255969] Lustre: srv-lustre-OST0000: No data found on store. Initialize space. [ 8983.259456] Lustre: Skipped 2 previous similar messages [ 8984.494350] Lustre: ctl-lustre-MDT0000: super-sequence allocation rc = 0 [0x0000000280000400-0x00000002c0000400]:0:ost [ 8984.500901] Lustre: Skipped 1 previous similar message [ 8984.504952] Lustre: cli-lustre-OST0000-super: Allocated super-sequence [0x0000000280000400-0x00000002c0000400]:0:ost] [ 8984.540679] Lustre: lustre-OST0000-osc-MDT0000: update sequence from 0x100000000 to 0x280000401 [ 8985.008106] Lustre: DEBUG MARKER: oleg142-server.virtnet: executing set_default_debug -1 all [ 8987.945274] Lustre: DEBUG MARKER: oleg142-client.virtnet: executing wait_import_state_mount (FULL|IDLE) osc.lustre-OST0000-osc-[-0-9a-f]*.ost_server_uuid [ 8989.065949] Lustre: DEBUG MARKER: oleg142-server.virtnet: executing wait_import_state FULL os[cp].lustre-OST0000-osc-MDT0000.ost_server_uuid 50 [ 8989.137059] Lustre: DEBUG MARKER: os[cp].lustre-OST0000-osc-MDT0000.ost_server_uuid in FULL state after 0 sec [ 8990.223144] Lustre: DEBUG MARKER: oleg142-server.virtnet: executing wait_import_state FULL os[cp].lustre-OST0000-osc-MDT0001.ost_server_uuid 50 [ 8990.300568] Lustre: DEBUG MARKER: os[cp].lustre-OST0000-osc-MDT0001.ost_server_uuid in FULL state after 0 sec [ 8994.530953] LustreError: lustre-OST0000-osc-MDT0001: operation ost_statfs to node 0@lo failed: rc = -107 [ 8994.535013] Lustre: lustre-OST0000-osc-MDT0001: Connection to lustre-OST0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 8994.541542] Lustre: Skipped 1 previous similar message [ 8994.544801] Lustre: lustre-OST0000: Not available for connect from 0@lo (stopping) [ 8994.548171] Lustre: Skipped 1 previous similar message [ 8996.992969] LustreError: 2715:0:(obd_class.h:479:obd_check_dev()) Device 19 not setup [ 8996.996457] LustreError: 2715:0:(obd_class.h:479:obd_check_dev()) Skipped 9 previous similar messages [ 8997.074248] Lustre: server umount lustre-OST0000 complete [ 8999.555096] Lustre: lustre-MDT0000-lwp-MDT0001: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 8999.555888] Lustre: lustre-MDT0000: Not available for connect from 0@lo (stopping) [ 8999.565519] Lustre: Skipped 2 previous similar messages [ 9004.515588] Lustre: server umount lustre-MDT0000 complete [ 9004.562633] LustreError: 32567:0:(ldlm_lib.c:1113:target_handle_connect()) lustre-MDT0000: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 9004.570284] LustreError: 32567:0:(ldlm_lib.c:1113:target_handle_connect()) Skipped 1 previous similar message [ 9005.890362] LustreError: 1340:0:(ldlm_lockd.c:2550:ldlm_cancel_handler()) ldlm_cancel from 0@lo arrived at 1747856871 with bad export cookie 16983359334760089897 [ 9005.892445] LustreError: MGC192.168.201.142@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 9005.901833] LustreError: 1340:0:(ldlm_lockd.c:2550:ldlm_cancel_handler()) Skipped 4 previous similar messages [ 9006.045003] Lustre: server umount lustre-MDT0001 complete [ 9011.322651] Lustre: DEBUG MARKER: oleg142-server.virtnet: executing load_modules_local [ 9014.452156] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 9014.643368] LustreError: 4663:0:(ldlm_lib.c:1113:target_handle_connect()) lustre-MDT0001: not available for connect from 0@lo (no target). If you are running an HA pair check that the target is mounted on the other server. [ 9014.669681] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 9014.673295] Lustre: Skipped 2 previous similar messages [ 9015.801221] Lustre: DEBUG MARKER: oleg142-server.virtnet: executing set_default_debug -1 all ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 10.85s (real) 5.71s (CPU), Child processes: 5.12s