UPDATE: the issue was caused by an outdated BIOS and unrelated to Fedora Linux. If you’re experiencing issues almost identical to mine, see what I marked as solution. If you’re not in the same boat, you can still skim the thread for some useful insight if you’re having a similar AMDGPU problem.
The title is self-explanatory… my GPU randomly resets while gaming, crashing the game and restarting KWin effects.
Now follows some useful context. To begin, here is some relevant system information of mine, copy-pasted from fastfetch:
OS: Fedora Linux 43 (Forty Three) x86_64
Host: 83DD (IdeaPad Slim 5 16AHP9)
Kernel: Linux 7.1.5-101.fc43.x86_64
Display (AGO0001): 1920x1080 in 13", 60 Hz [External]
DE: KDE Plasma 6.7.3
WM: KWin (Wayland)
WM Theme: SMOD
Theme: Windows7Aero (Aero) [Qt], Windows-7-Better [GTK2/3]
Icons: Windows 7 Aero [Qt], Windows 7 Aero [GTK2/3/4]
Font: Noto Sans (10pt) [Qt], Noto Sans (10pt) [GTK2/3/4]
Cursor: aero-drop (24px)
CPU: AMD Ryzen 7 8845HS (16) @ 5.14 GHz
GPU: AMD Radeon 780M Graphics [Integrated]
Memory: 3.43 GiB / 27.20 GiB (13%)
Swap: 0 B / 8.00 GiB (0%)
Disk (/): 845.93 GiB / 952.28 GiB (89%) - btrfs
Local IP (wlp2s0): 192.168.0.203/24
Battery (L23M3PK1): 100% [AC Connected]
Locale: en_GB.UTF-8
Second of all, I must talk about the issue itself as well. I don’t remember exactly when it started or what could have caused it, as it began happening quite some time ago and I simply put up with it. I’m asking here to get some insight into what the most likely cause could be and how I could go about troubleshooting and fixing this annoying issue.
Here are some of the things that I’ve done to my system that I recall to be around when the issue started, sorted from what I think is the most likely cause to the least likely.
- Installing AMD ROCm
- Installing COSMIC over my existing install (which was Fedora Sway). Sometime after, I ditched COSMIC for KDE Plasma, but my GPU was already having issues before I installed KDE. Obviously, I’m also having the same problem on KDE, otherwise I wouldn’t be here. Right now, I have all three desktops installed on my computer.
- Using an external monitor (that being an Acer V223HQ, a 1080p LED monitor like any other you’d find in early 2010s, it really doesn’t matter in contrast to the HDMI>VGA adapter which is what’s actually plugged into my laptop. Though I doubt this as I remember my games crashing on the internal display as well, even if not sure).
What I do know for sure is that I experienced the same random resets when connected to my TV over HDMI. - Running local LLMs via LocalAI (related to point 1)
Those are the things that I’ve done personally that might’ve led me here. And, of course, I can’t rule out the dreaded possibility of bad hardware, which I’d hope is not the case…
I know this report is really vague but I’d like to be pointed into the right direction for troubleshooting this mess… I am also not reimaging my computer, I have so many configs and hundreds of gigabytes of data and programs collected on a single BTRFS partition, maybe I could figure out how to reinstall Fedora over my existing root partition without deleting any data?
Third of all, let me cover what finally tipped me over to open a forum thread about this… I had just installed the Windows version of Minecraft Bedrock via this launcher called BedrockOnLinux and it was my second time playing it. I was on a server when, as usual, my monitor went black for a few seconds, displaying the “no signal” message. Then it goes back to my desktop but the game is frozen, my mouse is completely gone when inside the window and I get an “application not responding” dialog on top of the game which I can only OK away using the keyboard, after which I am back to the launcher… sometimes, but not in this case (probably due to DND settings or something), this is accompanied by a “KWin effects were restarted due to a graphics reset” notification after the fact.
I was absolutely done for after this incident, so I went to my terminal and printed out the dmesg. Sure enough, a bunch of blah blah about the GPU and it having successfully reset, this log will be quoted at the bottom. Read on for now.
As I said before, this issue is not KDE-exclusive and also not game specific. It also has nothing to do with my Plasma theme as I was having identical problems on COSMIC as well without any custom theme.
Other games that have caused my GPU driver to crash, non-exhaustive list: Minecraft Bedrock (but through mcpelauncher), BeamNG.drive, Titanfall 2…
There are others but I can’t say for sure, I literally can’t remember on which games it happens as I’ve been trying so hard to ignore the blatant problem.
There is absolutely no pattern to this, it is completely random! Sometimes, Minecraft would lock up before I can even connect to a server. Or a few minutes after I join. I could have been playing BeamNG perfectly fine for 1-2 hours, then I get the dreaded black screen and a not responding dialog followed by the game’s error reporter. All of these are accompanied by the KWin effects in a synchronised harmony.
Another thing to mention is that this has never happened during non-gaming use of my computer, even during GPU-intensive operations like prompting a local LLM.
And finally, here is the dmesg collected right after I got kicked out of Minecraft for Windows… I know for sure it’s the same thing for all the other games even though I haven’t seen dmesg, as the same graphics reset is mentioned after KWin panics.
Though it is a great excuse for procrastination, playing every game on my computer until my GPU resets then collecting dmesg, as opposed to doing the important work I don’t feel like doing.
I tend to ramble a lot but I’m done now, here is the dmesg.
[ 53.932175] amdgpu 0000:04:00.0: [drm] REG_WAIT timeout 1us * 100 tries - dcn31_program_compbuf_size line:141
[ 331.410461] amdgpu 0000:04:00.0: Dumping IP State
[ 331.412783] amdgpu 0000:04:00.0: Dumping IP State Completed
[ 331.412794] amdgpu 0000:04:00.0: [drm] AMDGPU device coredump file has been created
[ 331.412796] amdgpu 0000:04:00.0: [drm] Check your /sys/class/drm/card1/device/devcoredump/data
[ 331.412798] amdgpu 0000:04:00.0: ring gfx_0.0.0 timeout, signaled seq=54225, emitted seq=54227
[ 331.412800] amdgpu 0000:04:00.0: Process MINECRAFT MAIN pid 4990 thread vkd3d_queue pid 5081
[ 331.412802] amdgpu 0000:04:00.0: Starting gfx_0.0.0 ring reset
[ 333.416612] amdgpu 0000:04:00.0: MES failed to respond to msg=RESET
[ 333.416619] amdgpu 0000:04:00.0: failed to reset legacy queue
[ 333.416621] amdgpu 0000:04:00.0: reset via MES failed and try pipe reset -110
[ 333.416623] amdgpu 0000:04:00.0: The CPFW hasn't support pipe reset yet.
[ 333.416625] amdgpu 0000:04:00.0: Ring gfx_0.0.0 reset failed
[ 333.416628] amdgpu 0000:04:00.0: GPU reset begin!. Source: 1
[ 333.424239] amdgpu 0000:04:00.0: [drm] *ERROR* Failed to initialize parser -125!
[ 335.463265] amdgpu 0000:04:00.0: MES failed to respond to msg=REMOVE_QUEUE
[ 335.463279] amdgpu 0000:04:00.0: failed to unmap legacy queue
[ 335.737206] [drm:gfx_v11_0_hw_fini [amdgpu]] *ERROR* failed to halt cp gfx
[ 335.739173] amdgpu 0000:04:00.0: MODE2 reset
[ 335.777176] amdgpu 0000:04:00.0: GPU reset succeeded, trying to resume
[ 335.777808] amdgpu 0000:04:00.0: [drm] PCIE GART of 512M enabled (table at 0x00000080FFD00000).
[ 335.777895] amdgpu 0000:04:00.0: SMU is resuming...
[ 335.779928] amdgpu 0000:04:00.0: SMU is resumed successfully!
[ 335.787014] amdgpu 0000:04:00.0: [drm] DMUB hardware initialized: version=0x08005D00
[ 336.044962] amdgpu 0000:04:00.0: ring gfx_0.0.0 uses VM inv eng 0 on hub 0
[ 336.044968] amdgpu 0000:04:00.0: ring comp_1.0.0 uses VM inv eng 1 on hub 0
[ 336.044971] amdgpu 0000:04:00.0: ring comp_1.1.0 uses VM inv eng 4 on hub 0
[ 336.044972] amdgpu 0000:04:00.0: ring comp_1.2.0 uses VM inv eng 6 on hub 0
[ 336.044974] amdgpu 0000:04:00.0: ring comp_1.3.0 uses VM inv eng 7 on hub 0
[ 336.044975] amdgpu 0000:04:00.0: ring comp_1.0.1 uses VM inv eng 8 on hub 0
[ 336.044977] amdgpu 0000:04:00.0: ring comp_1.1.1 uses VM inv eng 9 on hub 0
[ 336.044978] amdgpu 0000:04:00.0: ring comp_1.2.1 uses VM inv eng 10 on hub 0
[ 336.044979] amdgpu 0000:04:00.0: ring comp_1.3.1 uses VM inv eng 11 on hub 0
[ 336.044981] amdgpu 0000:04:00.0: ring sdma0 uses VM inv eng 12 on hub 0
[ 336.044982] amdgpu 0000:04:00.0: ring vcn_unified_0 uses VM inv eng 0 on hub 8
[ 336.044984] amdgpu 0000:04:00.0: ring jpeg_dec uses VM inv eng 1 on hub 8
[ 336.044986] amdgpu 0000:04:00.0: ring mes_kiq_3.1.0 uses VM inv eng 13 on hub 0
[ 336.046784] amdgpu 0000:04:00.0: GPU reset(1) succeeded!
[ 336.046798] amdgpu 0000:04:00.0: [drm] device wedged, but no recovery needed
[ 343.697676] clocksource: Watchdog remote CPU 12 read timed out
[ 382.501108] evm: overlay not supported
[ 386.184202] usb 1-4: reset full-speed USB device number 3 using xhci_hcd
And to clarify, I have tried checking my /sys/class/drm/card1/device/devcoredump/data by opening the folder in Dolphin, and while there were some virtual files inside, the folder completely vanished seconds after opening it and I got kicked out… how am I supposed to check it then? There was also no chance to retry as the thing was already gone without a trace.
Please point me in the right direction… also excuse me for any potential mistakes or confusion in all the text above, English is not my first language.
Thanks in advance for your advice!



