Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > linux.debian.kernel > #86088 > unrolled thread
| Started by | Paul DeKraker <pdekraker+debbug@gmail.com> |
|---|---|
| First post | 2025-02-22 22:40 +0100 |
| Last post | 2025-03-23 07:20 +0100 |
| Articles | 8 — 5 participants |
Back to article view | Back to linux.debian.kernel
Bug#1098698: linux: Segfault and system hang on larger network file transfers Paul DeKraker <pdekraker+debbug@gmail.com> - 2025-02-22 22:40 +0100
Bug#1098698: linux: Segfault and system hang on larger network file transfers Salvatore Bonaccorso <carnil@debian.org> - 2025-02-23 16:40 +0100
Bug#1098698: linux: Segfault and system hang on larger network file transfers Paul DeKraker <pdekraker@gmail.com> - 2025-02-25 03:20 +0100
Bug#1098698: linux: Segfault and system hang on larger network file transfers Salvatore Bonaccorso <carnil@debian.org> - 2025-02-27 17:50 +0100
Processed: Re: Bug#1098698: linux: Segfault and system hang on larger network file transfers "Debian Bug Tracking System" <owner@bugs.debian.org> - 2025-02-23 16:40 +0100
Bug#1098698: linux: Segfault and system hang on larger network file transfers Salvatore Bonaccorso <carnil@debian.org> - 2025-03-01 15:20 +0100
Bug#1098698: linux: Segfault and system hang on larger network file transfers Norbert Lange <nolange79@gmail.com> - 2025-03-07 17:10 +0100
Bug#1098698: marked as done (linux: Segfault and system hang on larger network file transfers) "Debian Bug Tracking System" <owner@bugs.debian.org> - 2025-03-23 07:20 +0100
| From | Paul DeKraker <pdekraker+debbug@gmail.com> |
|---|---|
| Date | 2025-02-22 22:40 +0100 |
| Subject | Bug#1098698: linux: Segfault and system hang on larger network file transfers |
| Message-ID | <Kj5o5-1dVd-7@gated-at.bofh.it> |
Source: linux Severity: important Tags: upstream X-Debbugs-Cc: pdekraker+debbug@gmail.com Dear Maintainer, I am experiencing an issue where my system completely locks up when attempting a large network file transfer from a mounted smb share. When copying a file above 1 GB I am consistienly experiencing this behavior. I tried going back to the 6.12.3 kernel which is the oldest I have on the system and the probelm is there as well. Looking at the dump below my guess is that it was introduced with 6.12 and netfs/read_collect.c. I have been unable to get a dump with 6.12.15, but the behavior is consistient. The transfer starts, but after a few seconds the whole system locks up. 2/22/25 8:38 AM ------------[ cut here ]------------ 2/22/25 8:38 AM WARNING CPU: 4 PID: 291 at fs/netfs/read_collect.c:110 netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs] 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4 nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer snd_seq snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT nf_reject_ipv4 xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6 nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash algif_skcipher af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek snd_hda_codec_generic snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii crct10dif_pclmul nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel snd_intel_dspcfg btusb snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep gf128mul r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer libata rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc ppdev lp parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4 mbcache jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq raid1 raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu video amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched drm_suballoc_helper drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme usbcore cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt gpio_generic 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G OE 6.12.3-amd64 #1 Debian 6.12.3-1 2/22/25 8:38 AM Tainted [O]=OOT_MODULE, [E]=UNSIGNED_MODULE 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs] 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs] 2/22/25 8:38 AM Code 43 28 48 39 c8 0f 84 04 02 00 00 4c 89 40 58 0f 1f 44 00 00 0f 1f 44 00 00 48 8b 43 78 48 89 43 68 48 89 43 70 e9 6e fe ff ff <0f> 0b 49 8b 47 70 48 8b 74 24 30 8b 7c 24 38 41 0f b7 97 96 00 00 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010246 2/22/25 8:38 AM RAX 0000000000000000 RBX: 0000000000000000 RCX: 000000003b200000 2/22/25 8:38 AM RDX 000000003b600000 RSI: 000000003b600000 RDI: ffffdd69cfb90000 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000002 R09: 0000000000400000 2/22/25 8:38 AM R10 0000000000000008 R11: 0000000000000008 R12: ffff9cbba2abdaa8 2/22/25 8:38 AM R13 0000000000200000 R14: 000000003b400000 R15: ffff9cbd00ce2280 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000) knlGS:0000000000000000 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033 2/22/25 8:38 AM CR2 00007f724be0412c CR3: 000000010a98e000 CR4: 0000000000f50ef0 2/22/25 8:38 AM PKRU 55555554 2/22/25 8:38 AM Call Trace 2/22/25 8:38 AM <TASK> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs] 2/22/25 8:38 AM ? __warn.cold+0x93/0xf6 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs] 2/22/25 8:38 AM ? report_bug+0xff/0x140 2/22/25 8:38 AM ? handle_bug+0x58/0x90 2/22/25 8:38 AM ? exc_invalid_op+0x17/0x70 2/22/25 8:38 AM ? asm_exc_invalid_op+0x1a/0x20 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs] 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x48b/0xb50 [netfs] 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs] 2/22/25 8:38 AM process_one_work+0x177/0x330 2/22/25 8:38 AM worker_thread+0x252/0x390 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10 2/22/25 8:38 AM kthread+0xd2/0x100 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10 2/22/25 8:38 AM ret_from_fork+0x34/0x50 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30 2/22/25 8:38 AM </TASK> 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]--- 2/22/25 8:38 AM netfs R=0000003e[2] s=3b200000-3b7fffff ctl=400000/600000/600000 sl=4 2/22/25 8:38 AM netfs folioq: orders=09090909 2/22/25 8:38 AM BUG kernel NULL pointer dereference, address: 0000000000000000 2/22/25 8:38 AM #PF supervisor write access in kernel mode 2/22/25 8:38 AM #PF error_code(0x0002) - not-present page 2/22/25 8:38 AM PGD 0 P4D 0 2/22/25 8:38 AM Oops Oops: 0002 [#1] PREEMPT SMP NOPTI 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G W OE 6.12.3-amd64 #1 Debian 6.12.3-1 2/22/25 8:38 AM Tainted [W]=WARN, [O]=OOT_MODULE, [E]=UNSIGNED_MODULE 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs] 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs] 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41 e8 8b 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08 <f0> 41 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX: 0000000000000027 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI: ffff9cc93ee21780 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09: ffffab8240b07c50 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12: ffff9cbba2abdaa8 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15: ffff9cbd00ce2280 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000) knlGS:0000000000000000 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4: 0000000000f50ef0 2/22/25 8:38 AM PKRU 55555554 2/22/25 8:38 AM Call Trace 2/22/25 8:38 AM <TASK> 2/22/25 8:38 AM ? __die_body.cold+0x19/0x27 2/22/25 8:38 AM ? page_fault_oops+0x15a/0x2d0 2/22/25 8:38 AM ? exc_page_fault+0x7e/0x180 2/22/25 8:38 AM ? asm_exc_page_fault+0x26/0x30 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs] 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs] 2/22/25 8:38 AM process_one_work+0x177/0x330 2/22/25 8:38 AM worker_thread+0x252/0x390 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10 2/22/25 8:38 AM kthread+0xd2/0x100 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10 2/22/25 8:38 AM ret_from_fork+0x34/0x50 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30 2/22/25 8:38 AM </TASK> 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4 nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer snd_seq snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT nf_reject_ipv4 xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6 nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash algif_skcipher af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek snd_hda_codec_generic snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii crct10dif_pclmul nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel snd_intel_dspcfg btusb snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep gf128mul r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer libata rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc ppdev lp parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4 mbcache jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq raid1 raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu video amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched drm_suballoc_helper drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme usbcore cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt gpio_generic 2/22/25 8:38 AM CR2 0000000000000000 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]--- 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs] 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41 e8 8b 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08 <f0> 41 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX: 0000000000000027 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI: ffff9cc93ee21780 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09: ffffab8240b07c50 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12: ffff9cbba2abdaa8 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15: ffff9cbd00ce2280 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000) knlGS:0000000000000000 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4: 0000000000f50ef0 2/22/25 8:38 AM PKRU 55555554 2/22/25 8:38 AM note kworker/4:2[291] exited with irqs disabled -- System Information: Debian Release: trixie/sid APT prefers unstable APT policy: (500, 'unstable') Architecture: amd64 (x86_64) Foreign Architectures: i386 Kernel: Linux 6.12.15-amd64 (SMP w/12 CPU threads; PREEMPT) Kernel taint flags: TAINT_OOT_MODULE, TAINT_UNSIGNED_MODULE Locale: LANG=en_US.UTF-8, LC_CTYPE=en_US.UTF-8 (charmap=UTF-8), LANGUAGE not set Shell: /bin/sh linked to /usr/bin/dash Init: systemd (via /run/systemd/system) LSM: AppArmor: enabled
[toc] | [next] | [standalone]
| From | Salvatore Bonaccorso <carnil@debian.org> |
|---|---|
| Date | 2025-02-23 16:40 +0100 |
| Message-ID | <Kjmfg-1oKE-7@gated-at.bofh.it> |
| In reply to | #86088 |
Control: tags -1 + moreinfo
Hi Paul,
On Sat, Feb 22, 2025 at 04:26:50PM -0500, Paul DeKraker wrote:
> Source: linux
> Severity: important
> Tags: upstream
> X-Debbugs-Cc: pdekraker+debbug@gmail.com
>
> Dear Maintainer,
>
> I am experiencing an issue where my system completely locks up when attempting
> a large network file transfer from a mounted smb share. When copying a file
> above 1 GB I am consistienly experiencing this behavior. I tried going back to
> the 6.12.3 kernel which is the oldest I have on the system and the probelm is
> there as well. Looking at the dump below my guess is that it was introduced
> with 6.12 and netfs/read_collect.c. I have been unable to get a dump with
> 6.12.15, but the behavior is consistient. The transfer starts, but after a few
> seconds the whole system locks up.
>
>
> 2/22/25 8:38 AM ------------[ cut here ]------------
> 2/22/25 8:38 AM WARNING CPU: 4 PID: 291 at fs/netfs/read_collect.c:110
> netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4
> nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer snd_seq
> snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT nf_reject_ipv4
> xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6
> nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash algif_skcipher
> af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc
> edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek snd_hda_codec_generic
> snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii crct10dif_pclmul
> nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel snd_intel_dspcfg btusb
> snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl
> sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep gf128mul
> r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer libata
> rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore
> i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp
> 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc ppdev lp
> parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4 mbcache
> jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic
> async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq raid1
> raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu video
> amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched drm_suballoc_helper
> drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme usbcore
> cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt gpio_generic
> 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G OE
> 6.12.3-amd64 #1 Debian 6.12.3-1
> 2/22/25 8:38 AM Tainted [O]=OOT_MODULE, [E]=UNSIGNED_MODULE
> 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS
> PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021
> 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs]
> 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> 2/22/25 8:38 AM Code 43 28 48 39 c8 0f 84 04 02 00 00 4c 89 40 58 0f 1f 44
> 00 00 0f 1f 44 00 00 48 8b 43 78 48 89 43 68 48 89 43 70 e9 6e fe ff ff <0f> 0b
> 49 8b 47 70 48 8b 74 24 30 8b 7c 24 38 41 0f b7 97 96 00 00
> 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010246
> 2/22/25 8:38 AM RAX 0000000000000000 RBX: 0000000000000000 RCX:
> 000000003b200000
> 2/22/25 8:38 AM RDX 000000003b600000 RSI: 000000003b600000 RDI:
> ffffdd69cfb90000
> 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000002 R09:
> 0000000000400000
> 2/22/25 8:38 AM R10 0000000000000008 R11: 0000000000000008 R12:
> ffff9cbba2abdaa8
> 2/22/25 8:38 AM R13 0000000000200000 R14: 000000003b400000 R15:
> ffff9cbd00ce2280
> 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> knlGS:0000000000000000
> 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> 2/22/25 8:38 AM CR2 00007f724be0412c CR3: 000000010a98e000 CR4:
> 0000000000f50ef0
> 2/22/25 8:38 AM PKRU 55555554
> 2/22/25 8:38 AM Call Trace
> 2/22/25 8:38 AM <TASK>
> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> 2/22/25 8:38 AM ? __warn.cold+0x93/0xf6
> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> 2/22/25 8:38 AM ? report_bug+0xff/0x140
> 2/22/25 8:38 AM ? handle_bug+0x58/0x90
> 2/22/25 8:38 AM ? exc_invalid_op+0x17/0x70
> 2/22/25 8:38 AM ? asm_exc_invalid_op+0x1a/0x20
> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x48b/0xb50 [netfs]
> 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0
> 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs]
> 2/22/25 8:38 AM process_one_work+0x177/0x330
> 2/22/25 8:38 AM worker_thread+0x252/0x390
> 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10
> 2/22/25 8:38 AM kthread+0xd2/0x100
> 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> 2/22/25 8:38 AM ret_from_fork+0x34/0x50
> 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30
> 2/22/25 8:38 AM </TASK>
> 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]---
> 2/22/25 8:38 AM netfs R=0000003e[2] s=3b200000-3b7fffff
> ctl=400000/600000/600000 sl=4
> 2/22/25 8:38 AM netfs folioq: orders=09090909
> 2/22/25 8:38 AM BUG kernel NULL pointer dereference, address:
> 0000000000000000
> 2/22/25 8:38 AM #PF supervisor write access in kernel mode
> 2/22/25 8:38 AM #PF error_code(0x0002) - not-present page
> 2/22/25 8:38 AM PGD 0 P4D 0
> 2/22/25 8:38 AM Oops Oops: 0002 [#1] PREEMPT SMP NOPTI
> 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G W OE
> 6.12.3-amd64 #1 Debian 6.12.3-1
> 2/22/25 8:38 AM Tainted [W]=WARN, [O]=OOT_MODULE, [E]=UNSIGNED_MODULE
> 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS
> PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021
> 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs]
> 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs]
> 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41 e8 8b
> 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08 <f0> 41
> 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00
> 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206
> 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX:
> 0000000000000027
> 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI:
> ffff9cc93ee21780
> 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09:
> ffffab8240b07c50
> 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12:
> ffff9cbba2abdaa8
> 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15:
> ffff9cbd00ce2280
> 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> knlGS:0000000000000000
> 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4:
> 0000000000f50ef0
> 2/22/25 8:38 AM PKRU 55555554
> 2/22/25 8:38 AM Call Trace
> 2/22/25 8:38 AM <TASK>
> 2/22/25 8:38 AM ? __die_body.cold+0x19/0x27
> 2/22/25 8:38 AM ? page_fault_oops+0x15a/0x2d0
> 2/22/25 8:38 AM ? exc_page_fault+0x7e/0x180
> 2/22/25 8:38 AM ? asm_exc_page_fault+0x26/0x30
> 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs]
> 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0
> 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs]
> 2/22/25 8:38 AM process_one_work+0x177/0x330
> 2/22/25 8:38 AM worker_thread+0x252/0x390
> 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10
> 2/22/25 8:38 AM kthread+0xd2/0x100
> 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> 2/22/25 8:38 AM ret_from_fork+0x34/0x50
> 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30
> 2/22/25 8:38 AM </TASK>
> 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4
> nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer snd_seq
> snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT nf_reject_ipv4
> xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6
> nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash algif_skcipher
> af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc
> edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek snd_hda_codec_generic
> snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii crct10dif_pclmul
> nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel snd_intel_dspcfg btusb
> snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl
> sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep gf128mul
> r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer libata
> rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore
> i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp
> 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc ppdev lp
> parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4 mbcache
> jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic
> async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq raid1
> raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu video
> amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched drm_suballoc_helper
> drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme usbcore
> cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt gpio_generic
> 2/22/25 8:38 AM CR2 0000000000000000
> 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]---
> 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50 [netfs]
> 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41 e8 8b
> 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08 <f0> 41
> 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00
> 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206
> 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX:
> 0000000000000027
> 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI:
> ffff9cc93ee21780
> 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09:
> ffffab8240b07c50
> 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12:
> ffff9cbba2abdaa8
> 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15:
> ffff9cbd00ce2280
> 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> knlGS:0000000000000000
> 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4:
> 0000000000f50ef0
> 2/22/25 8:38 AM PKRU 55555554
> 2/22/25 8:38 AM note kworker/4:2[291] exited with irqs disabled
Thanks for the report. This very much sounded at first like
https://lore.kernel.org/all/CANT5p=qBwjBm-D8soFVVtswGEfmMtQXVW83=TNfUtvyHeFQZBA@mail.gmail.com/
which has a fix c8b90d40d5bb ("netfs: Fix non-contiguous donation
between completed reads") which OTOH has landed in 6.13-rc7 and
6.12.11.
So can you confirm: With the most recent kernel in unstable you do not
get anymore above trace, but you observe stalls in transfering a large
file. In which case this might be orthogonal to the above.
Regards,
Salvatore
[toc] | [prev] | [next] | [standalone]
| From | Paul DeKraker <pdekraker@gmail.com> |
|---|---|
| Date | 2025-02-25 03:20 +0100 |
| Message-ID | <KjSI9-1Ma3-1@gated-at.bofh.it> |
| In reply to | #86097 |
[Multipart message — attachments visible in raw view] — view raw
I just tried 6.12.16 and as with 6.12.15 the system locks up / becomes
completely unresponsive and needs to be hard reset when attempting a large
transfer from a network share to the local computer, so this probably is a
separate issue. If there is something more I can do to capture this
behavior please let me know.
Paul
On Sun, Feb 23, 2025 at 10:36 AM Salvatore Bonaccorso <carnil@debian.org>
wrote:
> Control: tags -1 + moreinfo
>
> Hi Paul,
>
> On Sat, Feb 22, 2025 at 04:26:50PM -0500, Paul DeKraker wrote:
> > Source: linux
> > Severity: important
> > Tags: upstream
> > X-Debbugs-Cc: pdekraker+debbug@gmail.com
> >
> > Dear Maintainer,
> >
> > I am experiencing an issue where my system completely locks up when
> attempting
> > a large network file transfer from a mounted smb share. When copying a
> file
> > above 1 GB I am consistienly experiencing this behavior. I tried going
> back to
> > the 6.12.3 kernel which is the oldest I have on the system and the
> probelm is
> > there as well. Looking at the dump below my guess is that it was
> introduced
> > with 6.12 and netfs/read_collect.c. I have been unable to get a dump
> with
> > 6.12.15, but the behavior is consistient. The transfer starts, but after
> a few
> > seconds the whole system locks up.
> >
> >
> > 2/22/25 8:38 AM ------------[ cut here ]------------
> > 2/22/25 8:38 AM WARNING CPU: 4 PID: 291 at fs/netfs/read_collect.c:110
> > netfs_consume_read_data.isra.0+0x67f/0xb50 [netfs]
> > 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4
> > nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer
> snd_seq
> > snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT
> nf_reject_ipv4
> > xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6
> > nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash
> algif_skcipher
> > af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc
> > edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek
> snd_hda_codec_generic
> > snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii
> crct10dif_pclmul
> > nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel
> snd_intel_dspcfg btusb
> > snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl
> > sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep
> gf128mul
> > r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer
> libata
> > rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore
> > i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp
> > 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc
> ppdev lp
> > parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4
> mbcache
> > jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic
> > async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq
> raid1
> > raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu
> video
> > amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched
> drm_suballoc_helper
> > drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme
> usbcore
> > cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt
> gpio_generic
> > 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G OE
> > 6.12.3-amd64 #1 Debian 6.12.3-1
> > 2/22/25 8:38 AM Tainted [O]=OOT_MODULE, [E]=UNSIGNED_MODULE
> > 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS
> > PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021
> > 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs]
> > 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x67f/0xb50
> [netfs]
> > 2/22/25 8:38 AM Code 43 28 48 39 c8 0f 84 04 02 00 00 4c 89 40 58 0f
> 1f 44
> > 00 00 0f 1f 44 00 00 48 8b 43 78 48 89 43 68 48 89 43 70 e9 6e fe ff ff
> <0f> 0b
> > 49 8b 47 70 48 8b 74 24 30 8b 7c 24 38 41 0f b7 97 96 00 00
> > 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010246
> > 2/22/25 8:38 AM RAX 0000000000000000 RBX: 0000000000000000 RCX:
> > 000000003b200000
> > 2/22/25 8:38 AM RDX 000000003b600000 RSI: 000000003b600000 RDI:
> > ffffdd69cfb90000
> > 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000002 R09:
> > 0000000000400000
> > 2/22/25 8:38 AM R10 0000000000000008 R11: 0000000000000008 R12:
> > ffff9cbba2abdaa8
> > 2/22/25 8:38 AM R13 0000000000200000 R14: 000000003b400000 R15:
> > ffff9cbd00ce2280
> > 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> > knlGS:0000000000000000
> > 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> > 2/22/25 8:38 AM CR2 00007f724be0412c CR3: 000000010a98e000 CR4:
> > 0000000000f50ef0
> > 2/22/25 8:38 AM PKRU 55555554
> > 2/22/25 8:38 AM Call Trace
> > 2/22/25 8:38 AM <TASK>
> > 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50
> [netfs]
> > 2/22/25 8:38 AM ? __warn.cold+0x93/0xf6
> > 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50
> [netfs]
> > 2/22/25 8:38 AM ? report_bug+0xff/0x140
> > 2/22/25 8:38 AM ? handle_bug+0x58/0x90
> > 2/22/25 8:38 AM ? exc_invalid_op+0x17/0x70
> > 2/22/25 8:38 AM ? asm_exc_invalid_op+0x1a/0x20
> > 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x67f/0xb50
> [netfs]
> > 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x48b/0xb50
> [netfs]
> > 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0
> > 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs]
> > 2/22/25 8:38 AM process_one_work+0x177/0x330
> > 2/22/25 8:38 AM worker_thread+0x252/0x390
> > 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10
> > 2/22/25 8:38 AM kthread+0xd2/0x100
> > 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> > 2/22/25 8:38 AM ret_from_fork+0x34/0x50
> > 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> > 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30
> > 2/22/25 8:38 AM </TASK>
> > 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]---
> > 2/22/25 8:38 AM netfs R=0000003e[2] s=3b200000-3b7fffff
> > ctl=400000/600000/600000 sl=4
> > 2/22/25 8:38 AM netfs folioq: orders=09090909
> > 2/22/25 8:38 AM BUG kernel NULL pointer dereference, address:
> > 0000000000000000
> > 2/22/25 8:38 AM #PF supervisor write access in kernel mode
> > 2/22/25 8:38 AM #PF error_code(0x0002) - not-present page
> > 2/22/25 8:38 AM PGD 0 P4D 0
> > 2/22/25 8:38 AM Oops Oops: 0002 [#1] PREEMPT SMP NOPTI
> > 2/22/25 8:38 AM CPU 4 UID: 0 PID: 291 Comm: kworker/4:2 Tainted: G W
> OE
> > 6.12.3-amd64 #1 Debian 6.12.3-1
> > 2/22/25 8:38 AM Tainted [W]=WARN, [O]=OOT_MODULE, [E]=UNSIGNED_MODULE
> > 2/22/25 8:38 AM Hardware name Gigabyte Technology Co., Ltd. B550M AORUS
> > PRO-P/B550M AORUS PRO-P, BIOS F13 07/08/2021
> > 2/22/25 8:38 AM Workqueue cifsiod smb2_readv_worker [cifs]
> > 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50
> [netfs]
> > 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41
> e8 8b
> > 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08
> <f0> 41
> > 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00
> > 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206
> > 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX:
> > 0000000000000027
> > 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI:
> > ffff9cc93ee21780
> > 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09:
> > ffffab8240b07c50
> > 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12:
> > ffff9cbba2abdaa8
> > 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15:
> > ffff9cbd00ce2280
> > 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> > knlGS:0000000000000000
> > 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> > 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4:
> > 0000000000f50ef0
> > 2/22/25 8:38 AM PKRU 55555554
> > 2/22/25 8:38 AM Call Trace
> > 2/22/25 8:38 AM <TASK>
> > 2/22/25 8:38 AM ? __die_body.cold+0x19/0x27
> > 2/22/25 8:38 AM ? page_fault_oops+0x15a/0x2d0
> > 2/22/25 8:38 AM ? exc_page_fault+0x7e/0x180
> > 2/22/25 8:38 AM ? asm_exc_page_fault+0x26/0x30
> > 2/22/25 8:38 AM ? netfs_consume_read_data.isra.0+0x2db/0xb50
> [netfs]
> > 2/22/25 8:38 AM ? finish_task_switch.isra.0+0x97/0x2c0
> > 2/22/25 8:38 AM netfs_read_subreq_terminated+0x2ab/0x3f0 [netfs]
> > 2/22/25 8:38 AM process_one_work+0x177/0x330
> > 2/22/25 8:38 AM worker_thread+0x252/0x390
> > 2/22/25 8:38 AM ? __pfx_worker_thread+0x10/0x10
> > 2/22/25 8:38 AM kthread+0xd2/0x100
> > 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> > 2/22/25 8:38 AM ret_from_fork+0x34/0x50
> > 2/22/25 8:38 AM ? __pfx_kthread+0x10/0x10
> > 2/22/25 8:38 AM ret_from_fork_asm+0x1a/0x30
> > 2/22/25 8:38 AM </TASK>
> > 2/22/25 8:38 AM Modules linked in ccm nls_utf8 cifs cifs_arc4
> > nls_ucs2_utils cifs_md4 dns_resolver netfs snd_seq_dummy snd_hrtimer
> snd_seq
> > snd_seq_device xt_CHECKSUM xt_MASQUERADE xt_conntrack ipt_REJECT
> nf_reject_ipv4
> > xt_tcpudp nft_compat nft_chain_nat nf_nat nf_conntrack nf_defrag_ipv6
> > nf_defrag_ipv4 nf_tables bridge stp llc rfcomm cmac algif_hash
> algif_skcipher
> > af_alg overlay qrtr bnep amd_atl intel_rapl_msr intel_rapl_common sunrpc
> > edac_mce_amd binfmt_misc kvm_amd snd_hda_codec_realtek
> snd_hda_codec_generic
> > snd_hda_scodec_component kvm snd_hda_codec_hdmi nls_ascii
> crct10dif_pclmul
> > nls_cp437 snd_hda_intel crc32_pclmul ghash_clmulni_intel
> snd_intel_dspcfg btusb
> > snd_intel_sdw_acpi sha512_ssse3 vfat fat snd_hda_codec sha256_ssse3 btrtl
> > sha1_ssse3 btintel snd_hda_core aesni_intel btbcm ahci btmtk snd_hwdep
> gf128mul
> > r8169 crypto_simd libahci snd_pcm bluetooth cryptd realtek snd_timer
> libata
> > rapl sp5100_tco snd watchdog wmi_bmof mdio_devres gigabyte_wmi soundcore
> > i2c_piix4 pcspkr i2c_smbus rfkill libphy scsi_mod ccp k10temp
> > 2/22/25 8:38 AM scsi_common button lm92 msr dm_mod parport_pc
> ppdev lp
> > parport efi_pstore configfs nfnetlink ip_tables x_tables autofs4 ext4
> mbcache
> > jbd2 razerkbd(OE) efivarfs raid10 raid456 libcrc32c crc32c_generic
> > async_raid6_recov async_memcpy async_pq async_xor xor async_tx raid6_pq
> raid1
> > raid0 md_mod evdev joydev razermouse(OE) hid_generic usbhid hid amdgpu
> video
> > amdxcp i2c_algo_bit drm_ttm_helper ttm drm_exec gpu_sched
> drm_suballoc_helper
> > drm_buddy drm_display_helper xhci_pci xhci_hcd drm_kms_helper drm nvme
> usbcore
> > cec rc_core nvme_core crc32c_intel crc16 usb_common wmi gpio_amdpt
> gpio_generic
> > 2/22/25 8:38 AM CR2 0000000000000000
> > 2/22/25 8:38 AM ---[ end trace 0000000000000000 ]---
> > 2/22/25 8:38 AM RIP 0010:netfs_consume_read_data.isra.0+0x2db/0xb50
> [netfs]
> > 2/22/25 8:38 AM Code c4 40 5b 5d 41 5c 41 5d 41 5e 41 5f e9 e9 86 41
> e8 8b
> > 6c 24 38 48 8b 44 24 28 48 89 f3 49 2b 5f 60 49 89 5f 78 4c 8b 6c e8 08
> <f0> 41
> > 80 4d 00 08 48 8b 44 24 30 48 8b 80 58 02 00 00 a9 00 00 00
> > 2/22/25 8:38 AM RSP 0018:ffffab8240b07dd8 EFLAGS: 00010206
> > 2/22/25 8:38 AM RAX ffff9cbdf4cb7200 RBX: 0000000000600000 RCX:
> > 0000000000000027
> > 2/22/25 8:38 AM RDX 0000000000000000 RSI: 000000003b800000 RDI:
> > ffff9cc93ee21780
> > 2/22/25 8:38 AM RBP 0000000000000004 R08: 0000000000000000 R09:
> > ffffab8240b07c50
> > 2/22/25 8:38 AM R10 ffffffffab4b42c8 R11: 0000000000000003 R12:
> > ffff9cbba2abdaa8
> > 2/22/25 8:38 AM R13 0000000000000000 R14: 000000003b600000 R15:
> > ffff9cbd00ce2280
> > 2/22/25 8:38 AM FS 0000000000000000(0000) GS:ffff9cc93ee00000(0000)
> > knlGS:0000000000000000
> > 2/22/25 8:38 AM CS 0010 DS: 0000 ES: 0000 CR0: 0000000080050033
> > 2/22/25 8:38 AM CR2 0000000000000000 CR3: 000000010a98e000 CR4:
> > 0000000000f50ef0
> > 2/22/25 8:38 AM PKRU 55555554
> > 2/22/25 8:38 AM note kworker/4:2[291] exited with irqs disabled
>
> Thanks for the report. This very much sounded at first like
>
> https://lore.kernel.org/all/CANT5p=qBwjBm-D8soFVVtswGEfmMtQXVW83=TNfUtvyHeFQZBA@mail.gmail.com/
> which has a fix c8b90d40d5bb ("netfs: Fix non-contiguous donation
> between completed reads") which OTOH has landed in 6.13-rc7 and
> 6.12.11.
>
> So can you confirm: With the most recent kernel in unstable you do not
> get anymore above trace, but you observe stalls in transfering a large
> file. In which case this might be orthogonal to the above.
>
> Regards,
> Salvatore
>
[toc] | [prev] | [next] | [standalone]
| From | Salvatore Bonaccorso <carnil@debian.org> |
|---|---|
| Date | 2025-02-27 17:50 +0100 |
| Message-ID | <KkPfb-2nI4-3@gated-at.bofh.it> |
| In reply to | #86126 |
Hi Paul, On Mon, Feb 24, 2025 at 09:12:43PM -0500, Paul DeKraker wrote: > I just tried 6.12.16 and as with 6.12.15 the system locks up / becomes > completely unresponsive and needs to be hard reset when attempting a large > transfer from a network share to the local computer, so this probably is a > separate issue. If there is something more I can do to capture this > behavior please let me know. Thanks for confirming. Yes now we have to get to fresh logs. As I understand you cannot access the system (not even remotely via SSH) to gather the kernel logs after the issue appears? In this case we might be successfull if you attach a netconsole and have another system available which can capture the logs. Documentation can be found here: https://www.kernel.org/doc/html/latest/networking/netconsole.html#netconsole Would you be able to try that and provide fresh logs for the recent 6.12.y kernels? Regards, Salvatore
[toc] | [prev] | [next] | [standalone]
| From | "Debian Bug Tracking System" <owner@bugs.debian.org> |
|---|---|
| Date | 2025-02-23 16:40 +0100 |
| Subject | Processed: Re: Bug#1098698: linux: Segfault and system hang on larger network file transfers |
| Message-ID | <Kjmfg-1oKE-19@gated-at.bofh.it> |
| In reply to | #86088 |
Processing control commands: > tags -1 + moreinfo Bug #1098698 [src:linux] linux: Segfault and system hang on larger network file transfers Added tag(s) moreinfo. -- 1098698: https://bugs.debian.org/cgi-bin/bugreport.cgi?bug=1098698 Debian Bug Tracking System Contact owner@bugs.debian.org with problems
[toc] | [prev] | [next] | [standalone]
| From | Salvatore Bonaccorso <carnil@debian.org> |
|---|---|
| Date | 2025-03-01 15:20 +0100 |
| Message-ID | <KlvR8-2QJI-5@gated-at.bofh.it> |
| In reply to | #86088 |
Hi Paul, On Sat, Mar 01, 2025 at 08:21:58AM -0500, Paul DeKraker wrote: > Here is a log collected via netconsole. > > Thanks, > Paul > > On Thu, Feb 27, 2025 at 11:39 AM Salvatore Bonaccorso <carnil@debian.org> > wrote: > > > Hi Paul, > > > > On Mon, Feb 24, 2025 at 09:12:43PM -0500, Paul DeKraker wrote: > > > I just tried 6.12.16 and as with 6.12.15 the system locks up / becomes > > > completely unresponsive and needs to be hard reset when attempting a > > large > > > transfer from a network share to the local computer, so this probably is > > a > > > separate issue. If there is something more I can do to capture this > > > behavior please let me know. > > > > Thanks for confirming. Yes now we have to get to fresh logs. As I > > understand you cannot access the system (not even remotely via SSH) to > > gather the kernel logs after the issue appears? > > > > In this case we might be successfull if you attach a netconsole and > > have another system available which can capture the logs. > > Documentation can be found here: > > > > https://www.kernel.org/doc/html/latest/networking/netconsole.html#netconsole > > > > Would you be able to try that and provide fresh logs for the recent > > 6.12.y kernels? Thanks a lot for the log, this was very helpful. At first glance it looks like this issue: https://lore.kernel.org/netfs/CAKPOu+_4mUwYgQtRTbXCmi+-k3PGvLysnPadkmHOyB7Gz0iSMA@mail.gmail.com/ There is a submitted patch but it does not look like it got applied already, checking. Might you be able to test the patch from https://lore.kernel.org/netfs/20250210191118.3444416-1-max.kellermann@ionos.com/ to see if it fixes the issue? You can follow the procedure in https://kernel-team.pages.debian.net/kernel-handbook/ch-common-tasks.html#id-1.6.6.4 to do the "simple patching and building". Regards, Salvatore
[toc] | [prev] | [next] | [standalone]
| From | Norbert Lange <nolange79@gmail.com> |
|---|---|
| Date | 2025-03-07 17:10 +0100 |
| Message-ID | <KnIqS-4l04-17@gated-at.bofh.it> |
| In reply to | #86088 |
On Sat, 1 Mar 2025 15:10:17 +0100 Salvatore Bonaccorso <carnil@debian.org> wrote: > Hi Paul, > > On Sat, Mar 01, 2025 at 08:21:58AM -0500, Paul DeKraker wrote: > > Here is a log collected via netconsole. > > > > Thanks, > > Paul > > > > On Thu, Feb 27, 2025 at 11:39 AM Salvatore Bonaccorso <carnil@debian.org> > > wrote: > > > > > Hi Paul, > > > > > > On Mon, Feb 24, 2025 at 09:12:43PM -0500, Paul DeKraker wrote: > > > > I just tried 6.12.16 and as with 6.12.15 the system locks up / becomes > > > > completely unresponsive and needs to be hard reset when attempting a > > > large > > > > transfer from a network share to the local computer, so this probably is > > > a > > > > separate issue. If there is something more I can do to capture this > > > > behavior please let me know. > > > > > > Thanks for confirming. Yes now we have to get to fresh logs. As I > > > understand you cannot access the system (not even remotely via SSH) to > > > gather the kernel logs after the issue appears? > > > > > > In this case we might be successfull if you attach a netconsole and > > > have another system available which can capture the logs. > > > Documentation can be found here: > > > > > > https://www.kernel.org/doc/html/latest/networking/netconsole.html#netconsole > > > > > > Would you be able to try that and provide fresh logs for the recent > > > 6.12.y kernels? > > Thanks a lot for the log, this was very helpful. At first glance it > looks like this issue: > https://lore.kernel.org/netfs/CAKPOu+_4mUwYgQtRTbXCmi+-k3PGvLysnPadkmHOyB7Gz0iSMA@mail.gmail.com/ > > There is a submitted patch but it does not look like it got applied > already, checking. > > Might you be able to test the patch from > https://lore.kernel.org/netfs/20250210191118.3444416-1-max.kellermann@ionos.com/ > to see if it fixes the issue? > > You can follow the procedure in > https://kernel-team.pages.debian.net/kernel-handbook/ch-common-tasks.html#id-1.6.6.4 > to do the "simple patching and building". > > Regards, > Salvatore (followed up from https://bugs.debian.org/cgi-bin/bugreport.cgi?bug=1099591) I reproduced with the available versions in debian: linux-image-6.12.12-amd64 6.12.12-1 -> segfault linux-image-6.12.17-amd64 6.12.17-1 -> segfault linux-image-6.13-amd64 6.13.5-1~exp1 -> 'kernel BUG at fs/netfs/read_collect.c:316!' Then I took the debian 6.12.17-1 kernel (latest LTS), added those 3 patches: https://lore.kernel.org/netfs/20250211093432.3524035-1-max.kellermann@ionos.com/ https://lore.kernel.org/netfs/20250210223144.3481766-1-max.kellermann@ionos.com/ https://lore.kernel.org/netfs/20250210191118.3444416-1-max.kellermann@ionos.com/ The resulting kernel apparently fixed the issue, I just testet in Qemu so far (no signed kernel for secure boot). Please provide a fixed version ASAP. regards, Norbert
[toc] | [prev] | [next] | [standalone]
| From | "Debian Bug Tracking System" <owner@bugs.debian.org> |
|---|---|
| Date | 2025-03-23 07:20 +0100 |
| Subject | Bug#1098698: marked as done (linux: Segfault and system hang on larger network file transfers) |
| Message-ID | <KtmQF-83Zl-13@gated-at.bofh.it> |
| In reply to | #86088 |
[Multipart message — attachments visible in raw view] — view raw
Your message dated Sun, 23 Mar 2025 06:11:29 +0000 with message-id <E1twEYT-008rb4-2L@fasolo.debian.org> and subject line Bug#1099591: fixed in linux 6.13.8-1~exp1 has caused the Debian Bug report #1099591, regarding linux: Segfault and system hang on larger network file transfers to be marked as done. This means that you claim that the problem has been dealt with. If this is not the case it is now your responsibility to reopen the Bug report if necessary, and/or fix the problem forthwith. (NB: If you are a system administrator and have no idea what this message is talking about, this may indicate a serious mail system misconfiguration somewhere. Please contact owner@bugs.debian.org immediately.) -- 1099591: https://bugs.debian.org/cgi-bin/bugreport.cgi?bug=1099591 Debian Bug Tracking System Contact owner@bugs.debian.org with problems
[toc] | [prev] | [standalone]
Back to top | Article view | linux.debian.kernel
csiph-web