Patch "drm/amdgpu: Fix call trace warning and hang when removing amdgpu device" has been added to the 6.2-stable tree

[Date Prev][Date Next][Thread Prev][Thread Next][Date Index][Thread Index]

 



This is a note to let you know that I've just added the patch titled

    drm/amdgpu: Fix call trace warning and hang when removing amdgpu device

to the 6.2-stable tree which can be found at:
    http://www.kernel.org/git/?p=linux/kernel/git/stable/stable-queue.git;a=summary

The filename of the patch is:
     drm-amdgpu-fix-call-trace-warning-and-hang-when-remo.patch
and it can be found in the queue-6.2 subdirectory.

If you, or anyone else, feels it should not be added to the stable tree,
please let <stable@xxxxxxxxxxxxxxx> know about it.



commit 93b828b6e8ee6a859da12faa08137ddcc803c87d
Author: lyndonli <Lyndon.Li@xxxxxxx>
Date:   Thu Mar 2 14:18:12 2023 +0800

    drm/amdgpu: Fix call trace warning and hang when removing amdgpu device
    
    [ Upstream commit 93bb18d2a873d2fa9625c8ea927723660a868b95 ]
    
    On GPUs with RAS enabled, below call trace and hang are observed when
    shutting down device.
    
    v2: use DRM device unplugged flag instead of shutdown flag as the check to
    prevent memory wipe in shutdown stage.
    
    [ +0.000000] RIP: 0010:amdgpu_vram_mgr_fini+0x18d/0x1c0 [amdgpu]
    [ +0.000001] PKRU: 55555554
    [ +0.000001] Call Trace:
    [ +0.000001] <TASK>
    [ +0.000002] amdgpu_ttm_fini+0x140/0x1c0 [amdgpu]
    [ +0.000183] amdgpu_bo_fini+0x27/0xa0 [amdgpu]
    [ +0.000184] gmc_v11_0_sw_fini+0x2b/0x40 [amdgpu]
    [ +0.000163] amdgpu_device_fini_sw+0xb6/0x510 [amdgpu]
    [ +0.000152] amdgpu_driver_release_kms+0x16/0x30 [amdgpu]
    [ +0.000090] drm_dev_release+0x28/0x50 [drm]
    [ +0.000016] devm_drm_dev_init_release+0x38/0x60 [drm]
    [ +0.000011] devm_action_release+0x15/0x20
    [ +0.000003] release_nodes+0x40/0xc0
    [ +0.000001] devres_release_all+0x9e/0xe0
    [ +0.000001] device_unbind_cleanup+0x12/0x80
    [ +0.000003] device_release_driver_internal+0xff/0x160
    [ +0.000001] driver_detach+0x4a/0x90
    [ +0.000001] bus_remove_driver+0x6c/0xf0
    [ +0.000001] driver_unregister+0x31/0x50
    [ +0.000001] pci_unregister_driver+0x40/0x90
    [ +0.000003] amdgpu_exit+0x15/0x120 [amdgpu]
    
    Signed-off-by: lyndonli <Lyndon.Li@xxxxxxx>
    Reviewed-by: Guchun Chen <guchun.chen@xxxxxxx>
    Reviewed-by: Christian König <christian.koenig@xxxxxxx>
    Signed-off-by: Alex Deucher <alexander.deucher@xxxxxxx>
    Signed-off-by: Sasha Levin <sashal@xxxxxxxxxx>

diff --git a/drivers/gpu/drm/amd/amdgpu/amdgpu_object.c b/drivers/gpu/drm/amd/amdgpu/amdgpu_object.c
index 25a68d8888e0d..5d4649b8bfd33 100644
--- a/drivers/gpu/drm/amd/amdgpu/amdgpu_object.c
+++ b/drivers/gpu/drm/amd/amdgpu/amdgpu_object.c
@@ -1315,7 +1315,7 @@ void amdgpu_bo_release_notify(struct ttm_buffer_object *bo)
 
 	if (!bo->resource || bo->resource->mem_type != TTM_PL_VRAM ||
 	    !(abo->flags & AMDGPU_GEM_CREATE_VRAM_WIPE_ON_RELEASE) ||
-	    adev->in_suspend || adev->shutdown)
+	    adev->in_suspend || drm_dev_is_unplugged(adev_to_drm(adev)))
 		return;
 
 	if (WARN_ON_ONCE(!dma_resv_trylock(bo->base.resv)))



[Date Prev][Date Next][Thread Prev][Thread Next][Date Index][Thread Index]
[Index of Archives]     [Linux USB Devel]     [Linux Audio Users]     [Yosemite News]     [Linux Kernel]     [Linux SCSI]

  Powered by Linux