Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > linux.debian.kernel > #74411 > unrolled thread
| Started by | Eric Valette <eric.valette@free.fr> |
|---|---|
| First post | 2022-02-15 19:20 +0100 |
| Last post | 2022-02-16 21:10 +0100 |
| Articles | 7 — 2 participants |
Back to article view | Back to linux.debian.kernel
This discussion starts older than the indexed window; earlier articles aren't shown. The article labeled Started by
below is the oldest one visible, not the original post.
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Eric Valette <eric.valette@free.fr> - 2022-02-15 19:20 +0100
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Eric Valette <eric.valette@free.fr> - 2022-02-15 19:40 +0100
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Alex Deucher <alexdeucher@gmail.com> - 2022-02-15 21:50 +0100
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Alex Deucher <alexdeucher@gmail.com> - 2022-02-16 21:10 +0100
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Eric Valette <eric.valette@free.fr> - 2022-02-16 22:40 +0100
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Eric Valette <eric.valette@free.fr> - 2022-04-10 12:30 +0200
Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces Eric Valette <eric.valette@free.fr> - 2022-02-16 21:10 +0100
| From | Eric Valette <eric.valette@free.fr> |
|---|---|
| Date | 2022-02-15 19:20 +0100 |
| Subject | Bug#1005005: same behavior on my MSI Bravo 17 A4DDR : second resume fails and exception in kernel traces |
| Message-ID | <DRaAF-21Dp-1@gated-at.bofh.it> |
Since I upgraded from 5.10 (own compiled kernel or debian kernel) to 5.15 (own compiled from same config or debian kernel) and even 5.16 kernel from debian, I get this behavior : 1) First suspend and resume works, 2) But later suspendend always fails. I have kernel exceptions in amdgpu : Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803952] Hardware name: Micro-Star International Co., L td. Bravo 17 A4DDR/MS-17FK, BIOS E17FKAMS.117 10/29/2020 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803956] Workqueue: pm pm_runtime_work Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803966] RIP: 0010:dm_suspend+0x241/0x260 [amdgpu] Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804340] Code: 4c 89 e6 4c 89 ef e8 ee 8a 16 00 83 f8 0 1 74 21 89 c2 48 c7 c6 a0 59 69 c1 48 c7 c7 60 83 76 c1 e8 14 ba 06 ff e9 6d ff ff ff <0f> 0b e9 f8 fd ff ff 4c 89 e6 4c 89 ef e8 6d ba 15 00 e9 56 ff ff Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804344] RSP: 0018:ffffae6040647c90 EFLAGS: 00010286 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804348] RAX: 0000000000000000 RBX: ffff973088a40000 RC X: 0000000000000000 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804351] RDX: 000000000000000a RSI: 0000000000000000 RD I: ffff973088a40000 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804353] RBP: 0000000000000000 R08: 0000000003c0ca00 R0 9: 0000000080380002 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804356] R10: ffff9730860f25a0 R11: 000000000000005f R1 2: ffff973088a40000 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804358] R13: ffff97308132a0d0 R14: 0000000000000008 R1 5: 0000000000000000 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804361] FS: 0000000000000000(0000) GS:ffff97339f68000 0(0000) knlGS:0000000000000000 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804364] CS: 0010 DS: 0000 ES: 0000 CR0: 0000000080050 033 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804366] CR2: 00007fb8c9ce6000 CR3: 00000001115fa000 CR 4: 0000000000350ee0 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804369] Call Trace: Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804375] <TASK> Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804380] ? nv_common_set_clockgating_state+0xa3/0xb0 [ amdgpu] Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804693] amdgpu_device_ip_suspend_phase1+0x63/0xc0 [am dgpu] Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804977] amdgpu_device_suspend+0x66/0x110 [amdgpu] Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805260] amdgpu_pmops_runtime_suspend+0xad/0x180 [amdg pu] Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805542] pci_pm_runtime_suspend+0x5a/0x160 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805549] ? pci_dev_put+0x20/0x20 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805553] __rpm_callback+0x44/0x150 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805558] ? pci_dev_put+0x20/0x20 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805561] rpm_callback+0x59/0x70 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805565] ? pci_dev_put+0x20/0x20 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805568] rpm_suspend+0x14a/0x720 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805572] ? _raw_spin_unlock+0x16/0x30 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805580] ? finish_task_switch.isra.0+0xc1/0x2f0 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805586] ? __switch_to+0x114/0x440 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805593] pm_runtime_work+0x94/0xa0 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805597] process_one_work+0x1e8/0x3c0 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805604] worker_thread+0x50/0x3b0 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805608] ? rescuer_thread+0x370/0x370 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805611] kthread+0x16b/0x190 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805616] ? set_kthread_struct+0x40/0x40 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805621] ret_from_fork+0x22/0x30 Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805630] </TASK> Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805632] ---[ end trace 8d77579b410d926d ]--- Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142133] amdgpu 0000:03:00.0: [drm:amdgpu_ring_test_hel per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142444] [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KGQ d isable failed Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462157] amdgpu 0000:03:00.0: [drm:amdgpu_ring_test_hel per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462465] [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KCQ d isable failed Feb 14 15:40:59 pink-floyd3 kernel: [ 831.782375] [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* faile d to halt cp gfx Feb 14 15:41:05 pink-floyd3 kernel: [ 837.297839] amdgpu 0000:03:00.0: amdgpu: SMU: I'm not done with your previous command: SMN_C2PMSG_66:0x0000003A SMN_C2PMSG_82:0x00000000
[toc] | [next] | [standalone]
| From | Eric Valette <eric.valette@free.fr> |
|---|---|
| Date | 2022-02-15 19:40 +0100 |
| Message-ID | <DRaU2-21Jy-3@gated-at.bofh.it> |
| In reply to | #74411 |
On 15/02/2022 19:14, Eric Valette wrote: > Since I upgraded from 5.10 (own compiled kernel or debian kernel) to > 5.15 (own compiled from same config or debian kernel) and even 5.16 > kernel from debian, I get this behavior : > > 1) First suspend and resume works, > 2) But later suspendend always fails. I have kernel exceptions > in amdgpu : In addition when booting 5.10 I get this message : /var/log/syslog.1:Feb 13 11:11:48 pink-floyd3 kernel: [ 3.035073] amdgpu 0000:03:00.0: amdgpu: ACPI VFCT table present but broken (too short #2),skipping and I do not see it on console with later kernels. > > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803952] Hardware name: > Micro-Star International Co., L > td. Bravo 17 A4DDR/MS-17FK, BIOS E17FKAMS.117 10/29/2020 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803956] Workqueue: pm > pm_runtime_work > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803966] RIP: > 0010:dm_suspend+0x241/0x260 [amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804340] Code: 4c 89 e6 4c 89 > ef e8 ee 8a 16 00 83 f8 0 > 1 74 21 89 c2 48 c7 c6 a0 59 69 c1 48 c7 c7 60 83 76 c1 e8 14 ba 06 ff > e9 6d ff ff ff <0f> 0b e9 > f8 fd ff ff 4c 89 e6 4c 89 ef e8 6d ba 15 00 e9 56 ff ff > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804344] RSP: > 0018:ffffae6040647c90 EFLAGS: 00010286 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804348] RAX: 0000000000000000 > RBX: ffff973088a40000 RC > X: 0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804351] RDX: 000000000000000a > RSI: 0000000000000000 RD > I: ffff973088a40000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804353] RBP: 0000000000000000 > R08: 0000000003c0ca00 R0 > 9: 0000000080380002 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804356] R10: ffff9730860f25a0 > R11: 000000000000005f R1 > 2: ffff973088a40000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804358] R13: ffff97308132a0d0 > R14: 0000000000000008 R1 > 5: 0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804361] FS: > 0000000000000000(0000) GS:ffff97339f68000 > 0(0000) knlGS:0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804364] CS: 0010 DS: 0000 > ES: 0000 CR0: 0000000080050 > 033 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804366] CR2: 00007fb8c9ce6000 > CR3: 00000001115fa000 CR > 4: 0000000000350ee0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804369] Call Trace: > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804375] <TASK> > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804380] ? > nv_common_set_clockgating_state+0xa3/0xb0 [ > amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804693] > amdgpu_device_ip_suspend_phase1+0x63/0xc0 [am > dgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804977] > amdgpu_device_suspend+0x66/0x110 [amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805260] > amdgpu_pmops_runtime_suspend+0xad/0x180 [amdg > pu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805542] > pci_pm_runtime_suspend+0x5a/0x160 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805549] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805553] > __rpm_callback+0x44/0x150 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805558] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805561] rpm_callback+0x59/0x70 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805565] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805568] rpm_suspend+0x14a/0x720 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805572] ? > _raw_spin_unlock+0x16/0x30 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805580] ? > finish_task_switch.isra.0+0xc1/0x2f0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805586] ? > __switch_to+0x114/0x440 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805593] > pm_runtime_work+0x94/0xa0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805597] > process_one_work+0x1e8/0x3c0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805604] > worker_thread+0x50/0x3b0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805608] ? > rescuer_thread+0x370/0x370 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805611] kthread+0x16b/0x190 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805616] ? > set_kthread_struct+0x40/0x40 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805621] ret_from_fork+0x22/0x30 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805630] </TASK> > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805632] ---[ end trace > 8d77579b410d926d ]--- > Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142133] amdgpu 0000:03:00.0: > [drm:amdgpu_ring_test_hel > per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) > Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142444] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KGQ d > isable failed > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462157] amdgpu 0000:03:00.0: > [drm:amdgpu_ring_test_hel > per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462465] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KCQ d > isable failed > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.782375] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* faile > d to halt cp gfx > Feb 14 15:41:05 pink-floyd3 kernel: [ 837.297839] amdgpu 0000:03:00.0: > amdgpu: SMU: I'm not done > with your previous command: SMN_C2PMSG_66:0x0000003A > SMN_C2PMSG_82:0x00000000 > > >
[toc] | [prev] | [next] | [standalone]
| From | Alex Deucher <alexdeucher@gmail.com> |
|---|---|
| Date | 2022-02-15 21:50 +0100 |
| Message-ID | <DRcVP-22Vo-7@gated-at.bofh.it> |
| In reply to | #74411 |
What chip is this? Can you provide the full dmesg output? Alex On Tue, Feb 15, 2022 at 1:15 PM Eric Valette <eric.valette@free.fr> wrote: > > Since I upgraded from 5.10 (own compiled kernel or debian kernel) to > 5.15 (own compiled from same config or debian kernel) and even 5.16 > kernel from debian, I get this behavior : > > 1) First suspend and resume works, > 2) But later suspendend always fails. I have kernel exceptions in amdgpu : > > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803952] Hardware name: > Micro-Star International Co., L > td. Bravo 17 A4DDR/MS-17FK, BIOS E17FKAMS.117 10/29/2020 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803956] Workqueue: pm > pm_runtime_work > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.803966] RIP: > 0010:dm_suspend+0x241/0x260 [amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804340] Code: 4c 89 e6 4c 89 > ef e8 ee 8a 16 00 83 f8 0 > 1 74 21 89 c2 48 c7 c6 a0 59 69 c1 48 c7 c7 60 83 76 c1 e8 14 ba 06 ff > e9 6d ff ff ff <0f> 0b e9 > f8 fd ff ff 4c 89 e6 4c 89 ef e8 6d ba 15 00 e9 56 ff ff > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804344] RSP: > 0018:ffffae6040647c90 EFLAGS: 00010286 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804348] RAX: 0000000000000000 > RBX: ffff973088a40000 RC > X: 0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804351] RDX: 000000000000000a > RSI: 0000000000000000 RD > I: ffff973088a40000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804353] RBP: 0000000000000000 > R08: 0000000003c0ca00 R0 > 9: 0000000080380002 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804356] R10: ffff9730860f25a0 > R11: 000000000000005f R1 > 2: ffff973088a40000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804358] R13: ffff97308132a0d0 > R14: 0000000000000008 R1 > 5: 0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804361] FS: > 0000000000000000(0000) GS:ffff97339f68000 > 0(0000) knlGS:0000000000000000 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804364] CS: 0010 DS: 0000 > ES: 0000 CR0: 0000000080050 > 033 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804366] CR2: 00007fb8c9ce6000 > CR3: 00000001115fa000 CR > 4: 0000000000350ee0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804369] Call Trace: > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804375] <TASK> > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804380] ? > nv_common_set_clockgating_state+0xa3/0xb0 [ > amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804693] > amdgpu_device_ip_suspend_phase1+0x63/0xc0 [am > dgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.804977] > amdgpu_device_suspend+0x66/0x110 [amdgpu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805260] > amdgpu_pmops_runtime_suspend+0xad/0x180 [amdg > pu] > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805542] > pci_pm_runtime_suspend+0x5a/0x160 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805549] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805553] > __rpm_callback+0x44/0x150 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805558] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805561] rpm_callback+0x59/0x70 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805565] ? pci_dev_put+0x20/0x20 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805568] rpm_suspend+0x14a/0x720 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805572] ? > _raw_spin_unlock+0x16/0x30 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805580] ? > finish_task_switch.isra.0+0xc1/0x2f0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805586] ? > __switch_to+0x114/0x440 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805593] > pm_runtime_work+0x94/0xa0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805597] > process_one_work+0x1e8/0x3c0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805604] worker_thread+0x50/0x3b0 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805608] ? > rescuer_thread+0x370/0x370 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805611] kthread+0x16b/0x190 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805616] ? > set_kthread_struct+0x40/0x40 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805621] ret_from_fork+0x22/0x30 > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805630] </TASK> > Feb 14 15:40:58 pink-floyd3 kernel: [ 830.805632] ---[ end trace > 8d77579b410d926d ]--- > Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142133] amdgpu 0000:03:00.0: > [drm:amdgpu_ring_test_hel > per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) > Feb 14 15:40:58 pink-floyd3 kernel: [ 831.142444] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KGQ d > isable failed > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462157] amdgpu 0000:03:00.0: > [drm:amdgpu_ring_test_hel > per [amdgpu]] *ERROR* ring kiq_2.1.0 test failed (-110) > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.462465] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* KCQ d > isable failed > Feb 14 15:40:59 pink-floyd3 kernel: [ 831.782375] > [drm:gfx_v10_0_hw_fini [amdgpu]] *ERROR* faile > d to halt cp gfx > Feb 14 15:41:05 pink-floyd3 kernel: [ 837.297839] amdgpu 0000:03:00.0: > amdgpu: SMU: I'm not done > with your previous command: SMN_C2PMSG_66:0x0000003A > SMN_C2PMSG_82:0x00000000 > > >
[toc] | [prev] | [next] | [standalone]
| From | Alex Deucher <alexdeucher@gmail.com> |
|---|---|
| Date | 2022-02-16 21:10 +0100 |
| Message-ID | <DRyMF-2gNz-5@gated-at.bofh.it> |
| In reply to | #74413 |
On Wed, Feb 16, 2022 at 2:56 PM Eric Valette <eric.valette@free.fr> wrote: > > On Tue, 15 Feb 2022 15:42:27 -0500 Alex Deucher <alexdeucher@gmail.com> > wrote: > > What chip is this? Can you provide the full dmesg output? > > Its a a renoir chip with a second gpu : [Radeon RX 5500/5500M / Pro 5500M]. Thanks. Can you attach the dmesg output? Can you bisect? Alex > > > Processor Information > Socket Designation: FP6 > Type: Central Processor > Family: Zen > Manufacturer: Advanced Micro Devices, Inc. > ID: 01 0F 86 00 FF FB 8B 17 > Signature: Family 23, Model 96, Stepping 1 > Flags: > FPU (Floating-point unit on-chip) > VME (Virtual mode extension) > DE (Debugging extension) > PSE (Page size extension) > TSC (Time stamp counter) > MSR (Model specific registers) > PAE (Physical address extension) > MCE (Machine check exception) > CX8 (CMPXCHG8 instruction supported) > APIC (On-chip APIC hardware supported) > SEP (Fast system call) > MTRR (Memory type range registers) > PGE (Page global enable) > MCA (Machine check architecture) > CMOV (Conditional move instruction supported) > PAT (Page attribute table) > PSE-36 (36-bit page size extension) > CLFSH (CLFLUSH instruction supported) > MMX (MMX technology supported) > FXSR (FXSAVE and FXSTOR instructions supported) > SSE (Streaming SIMD extensions) > SSE2 (Streaming SIMD extensions 2) > HTT (Multi-threading) > Version: AMD Ryzen 7 4800H with Radeon Graphics > Voltage: 1.2 V > External Clock: 100 MHz > Max Speed: 4300 MHz > Current Speed: 2900 MHz > Status: Populated, Enabled > Upgrade: None > L1 Cache Handle: 0x000D > L2 Cache Handle: 0x000E > L3 Cache Handle: 0x000F > Serial Number: Unknown > Asset Tag: Unknown > Part Number: Unknown > Core Count: 8 > Core Enabled: 8 > Thread Count: 16 > Characteristics: > 64-bit capable > Multi-Core > Hardware Thread > Execute Protection > Enhanced Virtualization > Power/Performance Control
[toc] | [prev] | [next] | [standalone]
| From | Eric Valette <eric.valette@free.fr> |
|---|---|
| Date | 2022-02-16 22:40 +0100 |
| Message-ID | <DRAbL-2hyq-3@gated-at.bofh.it> |
| In reply to | #74418 |
On 16/02/2022 21:01, Eric Valette wrote: > On 16/02/2022 20:58, Alex Deucher wrote: >> On Wed, Feb 16, 2022 at 2:56 PM Eric Valette <eric.valette@free.fr> >> wrote: >>> > Here is the dmesg. For bisecting, I'm not home and the Intrenet > connection is just too slow. In order to help a bit, I started to build kernel with the kernel patches set I had on my laptop so here it is: 5.11 suspend/resume is ok 5.12 suspend/resume is ok 5.13 suspend OK, resume KO, PC is dead I have to reboot. 5.14 suspend/resume is ok multiple times but I do have the exceptions at various places sudo dmesg | grep RIP [ 19.918769] RIP: 0010:ieee80211_reconfig+0x9a/0x1300 [ 19.918976] RIP: 0010:drv_remove_interface+0xd8/0xe0 [ 19.919086] RIP: 0010:drv_stop+0xb8/0xc0 [ 20.320553] RIP: 0010:iwl_mvm_mac_ctxt_init+0x1e2/0x220 [iwlmvm] [ 20.320801] RIP: 0033:0x7f0d4632536d 5.15 you have the result with latest one 5.15.24. RIP is at a given place. Sorry to be unable to do more. --eric
[toc] | [prev] | [next] | [standalone]
| From | Eric Valette <eric.valette@free.fr> |
|---|---|
| Date | 2022-04-10 12:30 +0200 |
| Message-ID | <EaCZr-7MxD-1@gated-at.bofh.it> |
| In reply to | #74420 |
On 16/02/2022 22:36, Eric Valette wrote: > On 16/02/2022 21:01, Eric Valette wrote: >> On 16/02/2022 20:58, Alex Deucher wrote: >>> On Wed, Feb 16, 2022 at 2:56 PM Eric Valette <eric.valette@free.fr> >>> wrote: >>>> > >> Here is the dmesg. For bisecting, I'm not home and the Intrenet >> connection is just too slow. > > In order to help a bit, I started to build kernel with the kernel > patches set I had on my laptop so here it is: > > 5.11 suspend/resume is ok > 5.12 suspend/resume is ok > 5.13 suspend OK, resume KO, PC is dead I have to reboot. > 5.14 suspend/resume is ok multiple times but I do have the exceptions at > various places > sudo dmesg | grep RIP > [ 19.918769] RIP: 0010:ieee80211_reconfig+0x9a/0x1300 > [ 19.918976] RIP: 0010:drv_remove_interface+0xd8/0xe0 > [ 19.919086] RIP: 0010:drv_stop+0xb8/0xc0 > [ 20.320553] RIP: 0010:iwl_mvm_mac_ctxt_init+0x1e2/0x220 [iwlmvm] > [ 20.320801] RIP: 0033:0x7f0d4632536d > > 5.15 you have the result with latest one 5.15.24. RIP is at a given place. > > > Sorry to be unable to do more. As a follow up : 5.17.1 from debian fixes the problem. Unfortunately it is not a long term kernel... -- eric
[toc] | [prev] | [next] | [standalone]
| From | Eric Valette <eric.valette@free.fr> |
|---|---|
| Date | 2022-02-16 21:10 +0100 |
| Message-ID | <DRyMF-2gNz-7@gated-at.bofh.it> |
| In reply to | #74413 |
On Tue, 15 Feb 2022 15:42:27 -0500 Alex Deucher <alexdeucher@gmail.com>
wrote:
> What chip is this? Can you provide the full dmesg output?
Its a a renoir chip with a second gpu : [Radeon RX 5500/5500M / Pro 5500M].
Processor Information
Socket Designation: FP6
Type: Central Processor
Family: Zen
Manufacturer: Advanced Micro Devices, Inc.
ID: 01 0F 86 00 FF FB 8B 17
Signature: Family 23, Model 96, Stepping 1
Flags:
FPU (Floating-point unit on-chip)
VME (Virtual mode extension)
DE (Debugging extension)
PSE (Page size extension)
TSC (Time stamp counter)
MSR (Model specific registers)
PAE (Physical address extension)
MCE (Machine check exception)
CX8 (CMPXCHG8 instruction supported)
APIC (On-chip APIC hardware supported)
SEP (Fast system call)
MTRR (Memory type range registers)
PGE (Page global enable)
MCA (Machine check architecture)
CMOV (Conditional move instruction supported)
PAT (Page attribute table)
PSE-36 (36-bit page size extension)
CLFSH (CLFLUSH instruction supported)
MMX (MMX technology supported)
FXSR (FXSAVE and FXSTOR instructions supported)
SSE (Streaming SIMD extensions)
SSE2 (Streaming SIMD extensions 2)
HTT (Multi-threading)
Version: AMD Ryzen 7 4800H with Radeon Graphics
Voltage: 1.2 V
External Clock: 100 MHz
Max Speed: 4300 MHz
Current Speed: 2900 MHz
Status: Populated, Enabled
Upgrade: None
L1 Cache Handle: 0x000D
L2 Cache Handle: 0x000E
L3 Cache Handle: 0x000F
Serial Number: Unknown
Asset Tag: Unknown
Part Number: Unknown
Core Count: 8
Core Enabled: 8
Thread Count: 16
Characteristics:
64-bit capable
Multi-Core
Hardware Thread
Execute Protection
Enhanced Virtualization
Power/Performance Control
[toc] | [prev] | [standalone]
Back to top | Article view | linux.debian.kernel
csiph-web