************************ crashinfo ************************* /exports/testreports/49584/testresults/sanity-hsm-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg346-server-timeout-core (3.10.0-7.9-debug) +==========================+ | *** Crashinfo v1.3.7 *** | +==========================+ +++WARNING+++ PARTIAL DUMP with size(vmcore) < 25% size(RAM) KERNEL: /tmp/crash-anaysis.fs0Eg/vmlinux [TAINTED] DUMPFILE: /exports/testreports/49584/testresults/sanity-hsm-ldiskfs-DNE-centos7_x86_64-centos7_x86_64/oleg346-server-timeout-core [PARTIAL DUMP] CPUS: 4 DATE: Tue Feb 25 08:22:14 EST 2025 UPTIME: 01:23:37 LOAD AVERAGE: 0.00, 0.01, 0.05 TASKS: 274 NODENAME: oleg346-server.virtnet RELEASE: 3.10.0-7.9-debug VERSION: #1 SMP Sat Mar 26 23:28:42 EDT 2022 MACHINE: x86_64 (2399 Mhz) MEMORY: 4 GB PANIC: "" +--------------------------+ >------------------------| Per-cpu Stacks ('bt -a') |------------------------< +--------------------------+ -- CPU#0 -- PID=0 CPU=0 CMD=swapper/0 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 rest_init+0x8e #7 start_kernel+0x456 #8 x86_64_start_reservations+0x2a #9 x86_64_start_kernel+0x152 #10 start_cpu+0x5 -- CPU#1 -- PID=0 CPU=1 CMD=swapper/1 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#2 -- PID=0 CPU=2 CMD=swapper/2 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 -- CPU#3 -- PID=0 CPU=3 CMD=swapper/3 #-1 native_safe_halt+0xb, 449 bytes of data #0 default_idle+0x1e #1 default_enter_idle+0x45 #2 cpuidle_enter_state+0x40 #3 cpuidle_idle_call+0xd8 #4 arch_cpu_idle+0xe #5 cpu_startup_entry+0x14a #6 start_secondary+0x1eb #7 start_cpu+0x5 +--------------------------------+ >---------------------| How This Dump Has Been Created |---------------------< +--------------------------------+ Cannot identify the specific condition that triggered vmcore +---------------+ >------------------------------| Tasks Summary |------------------------------< +---------------+ Number of Threads That Ran Recently ----------------------------------- last second 26 last 5s 65 last 60s 79 ----- Total Numbers of Threads per State ------ TASK_INTERRUPTIBLE 270 TASK_RUNNING 1 +++WARNING+++ There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details +-----------------------+ >--------------------------| 5 Most Recent Threads |--------------------------< +-----------------------+ PID CMD Age ARGS ----- -------------- ------ ---------------------------- 1 systemd 0 ms /usr/lib/systemd/systemd --switched-root --system --deserialize 22 914 in:imjournal 0 ms /usr/sbin/rsyslogd -n 34 rcuos/3 0 ms (no user stack) 4511 dbuf_evict 102 ms (no user stack) 9 rcu_sched 194 ms (no user stack) +------------------------+ >-------------------------| Memory Usage (kmem -i) |-------------------------< +------------------------+ PAGES TOTAL PERCENTAGE TOTAL MEM 955067 3.6 GB ---- FREE 646646 2.5 GB 67% of TOTAL MEM USED 308421 1.2 GB 32% of TOTAL MEM SHARED 11633 45.4 MB 1% of TOTAL MEM BUFFERS 8871 34.7 MB 0% of TOTAL MEM CACHED 157392 614.8 MB 16% of TOTAL MEM SLAB 18083 70.6 MB 1% of TOTAL MEM TOTAL HUGE 0 0 ---- HUGE FREE 0 0 0% of TOTAL HUGE TOTAL SWAP 262143 1024 MB ---- SWAP USED 0 0 0% of TOTAL SWAP SWAP FREE 262143 1024 MB 100% of TOTAL SWAP COMMIT LIMIT 739676 2.8 GB ---- COMMITTED 61027 238.4 MB 8% of TOTAL LIMIT +-------------------------------+ >----------------------| Scheduler Runqueues (per CPU) |----------------------< +-------------------------------+ ---+ CPU=0 ---- | CURRENT TASK , CMD=swapper/0 ---+ CPU=1 ---- | CURRENT TASK , CMD=swapper/1 ---+ CPU=2 ---- | CURRENT TASK , CMD=swapper/2 ---+ CPU=3 ---- | CURRENT TASK , CMD=swapper/3 +------------------------+ >-------------------------| Network Status Summary |-------------------------< +------------------------+ TCP Connection Info ------------------- ESTABLISHED 9 LISTEN 3 NAGLE disabled (TCP_NODELAY): 7 user_data set (NFS etc.): 8 UDP Connection Info ------------------- 2 UDP sockets, 0 in ESTABLISHED Unix Connection Info ------------------------ ESTABLISHED 26 CLOSE 17 LISTEN 8 Raw sockets info -------------------- ESTABLISHED 1 Interfaces Info --------------- How long ago (in seconds) interfaces transmitted/received? Name RX TX ---- ---------- --------- lo n/a 5015.1 eth0 n/a 0.0 RSS_TOTAL=53264 pages, %mem= 0.9 +------------+ >-------------------------------| Mounted FS |-------------------------------< +------------+ MOUNT SUPERBLK TYPE DEVNAME DIRNAME ffff880138cca000 ffff880139940800 rootfs rootfs / ffff88012a330000 ffff88012a338000 sysfs sysfs /sys ffff88012a3301c0 ffff880139944000 proc proc /proc ffff88012a330380 ffff880137678000 devtmpfs devtmpfs /dev ffff88012a330540 ffff8800b5255000 securityfs securityfs /sys/kernel/security ffff88012a330700 ffff88012a338800 tmpfs tmpfs /dev/shm ffff88012a3308c0 ffff8801372f6000 devpts devpts /dev/pts ffff88012a330a80 ffff88012a339000 tmpfs tmpfs /run ffff88012a330c40 ffff88012a339800 tmpfs tmpfs /sys/fs/cgroup ffff88012a330e00 ffff88012a33a000 cgroup cgroup /sys/fs/cgroup/systemd ffff88012a330fc0 ffff88012a33a800 pstore pstore /sys/fs/pstore ffff88012a331180 ffff88012a33c800 cgroup cgroup /sys/fs/cgroup/perf_event ffff88012a331340 ffff88012a33c000 cgroup cgroup /sys/fs/cgroup/cpuset ffff88012a331500 ffff88012a33b800 cgroup cgroup /sys/fs/cgroup/devices ffff88012a3316c0 ffff88012a33b000 cgroup cgroup /sys/fs/cgroup/blkio ffff88012a331880 ffff88012a33d000 cgroup cgroup /sys/fs/cgroup/cpu,cpuacct ffff88012a331a40 ffff88012a33d800 cgroup cgroup /sys/fs/cgroup/net_cls,net_prio ffff88012a331c00 ffff88012a33e000 cgroup cgroup /sys/fs/cgroup/memory ffff88012a331dc0 ffff88012a33e800 cgroup cgroup /sys/fs/cgroup/pids ffff88012a2b8000 ffff88012a33f000 cgroup cgroup /sys/fs/cgroup/hugetlb ffff88012a2b81c0 ffff88012a33f800 cgroup cgroup /sys/fs/cgroup/freezer ffff88012a2b8380 ffff88012a3bd800 configfs configfs /sys/kernel/config ffff88012a2b9500 ffff88012a3b9000 ext4 /dev/nbd0 / ffff88012a2b96c0 ffff88013767e800 rpc_pipefs rpc_pipefs /var/lib/nfs/rpc_pipefs ffff88012a2b9dc0 ffff8800b4b80000 autofs systemd-1 /proc/sys/fs/binfmt_misc ffff88012b262e00 ffff880139947800 debugfs debugfs /sys/kernel/debug ffff88012a2b9c00 ffff8800b4b84000 hugetlbfs hugetlbfs /dev/hugepages ffff88012a2b9a40 ffff8801372f7800 mqueue mqueue /dev/mqueue ffff88012a2b9880 ffff8800b4b80800 binfmt_misc binfmt_misc /proc/sys/fs/binfmt_misc/ ffff8800b52421c0 ffff8800b412d000 ramfs none /mnt ffff8800b5242380 ffff8800b41b3800 squashfs /dev/vda /home/green/git/lustre-release ffff8800b5242540 ffff8800b41b0800 tmpfs none /var/lib/stateless/writable ffff880137668380 ffff8800b41b0800 tmpfs none /var/cache/man ffff880138ccb340 ffff8800b41b0800 tmpfs none /var/log ffff880137668540 ffff8800b41b0800 tmpfs none /var/lib/dbus ffff8800b52428c0 ffff8800b41b0800 tmpfs none /tmp ffff8800b5242a80 ffff8800b41b0800 tmpfs none /var/lib/dhclient ffff88012b262fc0 ffff8800b41b0800 tmpfs none /var/tmp ffff8800b5242c40 ffff8800b41b0800 tmpfs none /var/lib/NetworkManager ffff8800b5242e00 ffff8800b41b0800 tmpfs none /var/lib/systemd/random-seed ffff88012b263180 ffff8800b41b0800 tmpfs none /var/spool ffff8800b5242fc0 ffff8800b41b0800 tmpfs none /var/lib/nfs ffff8800b5243180 ffff8800b41b0800 tmpfs none /var/lib/gssproxy ffff88012b263340 ffff8800b41b0800 tmpfs none /var/lib/logrotate ffff8800b5243340 ffff8800b41b0800 tmpfs none /etc ffff880137668700 ffff8800b41b0800 tmpfs none /var/lib/rsyslog ffff880138ccb500 ffff8800b41b0800 tmpfs none /var/lib/dhclient/var/lib/dhclient ffff8800b5243500 ffff8800b4199000 nfs4 192.168.200.253:/exports/state/oleg346-server.virtnet /var/lib/stateless/state ffff8800b5243880 ffff8800b4199000 nfs4 192.168.200.253:/exports/state/oleg346-server.virtnet /boot ffff88012b263880 ffff8800b4199000 nfs4 192.168.200.253:/exports/state/oleg346-server.virtnet /etc/etc/kdump.conf ffff8800b5243a40 ffff88013767e800 rpc_pipefs sunrpc /var/lib/nfs/var/lib/nfs/rpc_pipefs ffff8800b1e03500 ffff8800b41b3800 squashfs /dev/vda /usr/sbin/mount.lustre ffff8800ac026a80 ffff88006edd2800 lustre /dev/mapper/ost1_flakey /mnt/lustre-ost1 ffff8800ac027340 ffff88007476c800 lustre /dev/mapper/ost2_flakey /mnt/lustre-ost2 ffff8800b2b86e00 ffff88013159d800 lustre /dev/mapper/mds2_flakey /mnt/lustre-mds2 ffff8801301c6c40 ffff88009aaf3800 lustre /dev/mapper/mds1_flakey /mnt/lustre-mds1 +-------------------------------+ >----------------------| Last 40 lines of dmesg buffer |----------------------< +-------------------------------+ [ 1638.602711] LustreError: MGC192.168.203.146@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1638.677621] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1638.687007] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1639.587616] Lustre: DEBUG MARKER: oleg346-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1641.026744] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1643.693383] Lustre: lustre-MDT0000-lwp-OST0000: Connection restored to (at 0@lo) [ 1643.696786] Lustre: Skipped 2 previous similar messages [ 1643.706670] Lustre: lustre-MDT0000: Recovery over after 0:03, of 3 clients 3 recovered and 0 were evicted. [ 1643.732467] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:487 to 0x280000401:513) [ 1643.732471] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:486 to 0x2c0000401:513) [ 1644.697799] Lustre: DEBUG MARKER: oleg346-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1645.139751] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1654.342041] Lustre: DEBUG MARKER: == sanity-hsm test 408: Verify fiemap on release file ==== 07:26:09 (1740486369) [ 1658.576046] Lustre: DEBUG MARKER: == sanity-hsm test 409a: Coordinator should not stop when in use ========================================================== 07:26:13 (1740486373) [ 1660.845781] LustreError: 8914:0:(mdt_hsm_cdt_client.c:366:mdt_hsm_register_hal()) cfs_fail_timeout id 164 sleeping for 5000ms [ 1665.851190] LustreError: 8914:0:(mdt_hsm_cdt_client.c:366:mdt_hsm_register_hal()) cfs_fail_timeout id 164 awake [ 1672.649181] Lustre: DEBUG MARKER: == sanity-hsm test 409b: getattr released file with CDT stopped after remount ========================================================== 07:26:27 (1740486387) [ 1673.724737] Lustre: Modifying parameter lustre.mdt.lustre-MDT0000.hsm_control in log params [ 1673.728797] Lustre: Skipped 1 previous similar message [ 1695.437702] Lustre: Failing over lustre-MDT0000 [ 1695.541338] Lustre: server umount lustre-MDT0000 complete [ 1698.779720] Lustre: lustre-MDT0000-lwp-OST0000: Connection to lustre-MDT0000 (at 0@lo) was lost; in progress operations using this service will wait for recovery to complete [ 1698.786925] Lustre: Skipped 5 previous similar messages [ 1708.436807] LDISKFS-fs (dm-0): mounted filesystem with ordered data mode. Opts: user_xattr,errors=remount-ro,no_mbcache,nodelalloc [ 1708.484560] LustreError: MGC192.168.203.146@tcp: Connection to MGS (at 0@lo) was lost; in progress operations using this service will fail [ 1708.563064] Lustre: lustre-MDT0000: Imperative Recovery not enabled, recovery window 60-180 [ 1708.572085] Lustre: lustre-MDT0000: in recovery but waiting for the first client to connect [ 1709.715343] Lustre: DEBUG MARKER: oleg346-server.virtnet: executing set_default_debug vfstrace rpctrace dlmtrace neterror ha config ioctl super lfsck all [ 1711.138443] Lustre: lustre-MDT0000: Will be in recovery for at least 1:00, or until 3 clients reconnect [ 1713.581102] Lustre: lustre-MDT0000-lwp-MDT0001: Connection restored to (at 0@lo) [ 1713.583995] Lustre: Skipped 3 previous similar messages [ 1713.595822] Lustre: lustre-MDT0000: Recovery over after 0:02, of 3 clients 3 recovered and 0 were evicted. [ 1713.621255] Lustre: lustre-OST0001: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x2c0000401:516 to 0x2c0000401:545) [ 1713.621273] Lustre: lustre-OST0000: new connection from lustre-MDT0000-mdtlov (cleaning up unused objects from 0x280000401:517 to 0x280000401:545) [ 1714.727598] Lustre: DEBUG MARKER: oleg346-client.virtnet: executing wait_import_state_mount (FULL|IDLE) mdc.lustre-MDT0000-mdc-*.mds_server_uuid [ 1715.317420] Lustre: DEBUG MARKER: mdc.lustre-MDT0000-mdc-*.mds_server_uuid in FULL state after 0 sec [ 1739.688667] Lustre: DEBUG MARKER: == sanity-hsm test 410: lfs data_version -s allows release of force-archived file ========================================================== 07:27:35 (1740486455) [ 1742.387198] Lustre: DEBUG MARKER: == sanity-hsm test 411: hsm_ops rbac role ================ 07:27:37 (1740486457) [ 1796.215718] Lustre: DEBUG MARKER: == sanity-hsm test 500: various LLAPI HSM tests ========== 07:28:31 (1740486511) [ 1817.924713] Lustre: HSM agent 7c3f5a47-5538-497e-acf8-2d8346fd2f85 already registered ****************************************************************************** ************************ A Summary Of Problems Found ************************* ****************************************************************************** -------------------- A list of all +++WARNING+++ messages -------------------- PARTIAL DUMP with size(vmcore) < 25% size(RAM) There are 3 threads running in their own namespaces Use 'taskinfo --ns' to get more details ------------------------------------------------------------------------------ ** Execution took 11.70s (real) 6.34s (CPU), Child processes: 5.33s