Hello! So I’ve been running EOS for about a year and things have been going pretty decent, but with my last system update things aren’t working so well. I think part of this issue is due to the fact I left the PC in hibernate for 3 weeks while I was dealing with some personal stuff and when I got back I decided to update the system instead of giving it a reboot first.
From my current digging, it appears that my system is no longer recognizing my gpu. When I put nvidia-smi into the terminal it gives me “no devices were found” and during boot I get a nvidia-drm error of “Failed to allocate NvKmsKapiDevice.” Due to the gpu not actually working, my second monitor also isn’t working. I’ve had similar issues with this in the past, but managed to fix it after some random tinkering, but I’m not finding much info I can wrap my head around on how to fix this one. I also tried using eos-shifttime but I was running to several issues with that. Figured it would be best to reach out to people who would know of a better way to fix this issue before I go trying to figure out shifttime.
My output from: lspci -k | grep -A 2 -E “(VGA|3D)”
08:00.0 VGA compatible controller: NVIDIA Corporation TU106 [GeForce RTX 2070] (rev a1)
Subsystem: Micro-Star International Co., Ltd. [MSI] Device 3731
Kernel driver in use: nvidia
Here’s the errors from: journalctl -p err -e
Sep 29 14:48:10 The-Wizard-Hut kernel: virt/tdx: TDX not supported by the host platform
Sep 29 14:48:10 The-Wizard-Hut kernel:
Sep 29 14:48:11 The-Wizard-Hut kernel: [drm:nv_drm_dev_load [nvidia_drm]] *ERROR* [nvidia-drm] [GPU ID 0x00000800] Failed t>
Sep 29 14:48:23 The-Wizard-Hut kernel: nvidia-gpu 0000:08:00.3: i2c timeout error e0000000
Sep 29 14:48:23 The-Wizard-Hut kernel: ucsi_ccg 3-0008: i2c_transfer failed -110
Sep 29 14:48:23 The-Wizard-Hut kernel: ucsi_ccg 3-0008: ucsi_ccg_init failed - -110
Sep 29 14:48:23 The-Wizard-Hut kernel: ucsi_ccg 3-0008: probe with driver ucsi_ccg failed with error -110
Sep 29 14:48:34 The-Wizard-Hut wpa_supplicant[1194]: bgscan simple: Failed to enable signal strength monitoring
Here’s my “inxi -FAZ” with hopefully all of the reliant info (I can send more if needed obvi).
System:
Kernel: 7.2.7-arch1-1 arch: x86_64 bits: 64
Desktop: KDE Plasma v: 6.7.5 Distro: EndeavourOS
Machine:
Type: Desktop Mobo: ASUSTeK model: PRIME X470-PRO v: Rev X.0x
serial: <superuser required> Firmware: UEFI vendor: American Megatrends
v: 4024 date: 09/07/2018
CPU:
Info: 8-core model: AMD Ryzen 7 2700X bits: 64 type: MT MCP cache: L2: 4 MiB
Speed (MHz): avg: 2200 min/max: 2200/4350 cores: 1: 2200 2: 2200 3: 2200
4: 2200 5: 2200 6: 2200 7: 2200 8: 2200 9: 2200 10: 2200 11: 2200 12: 2200
13: 2200 14: 2200 15: 2200 16: 2200
Graphics:
Device-1: NVIDIA TU106 [GeForce RTX 2070] driver: nvidia v: 615.71.09
Display: wayland server: X.org v: 1.21.1.24 with: Xwayland v: 24.1.13
compositor: kwin_wayland driver: X: loaded: modesetting
gpu: simple-framebuffer resolution: 1024x768~60Hz
API: EGL v: 1.5 drivers: swrast platforms: wayland,x11,surfaceless,device
API: OpenGL v: 4.6 vendor: mesa v: 26.2.3-arch1.1 renderer: llvmpipe
(LLVM 22.1.8 256 bits)
API: Vulkan Message: No Vulkan data available.
Info: Tools: api: clinfo, eglinfo, glxinfo, vulkaninfo
de: kscreen-console,kscreen-doctor gpu: nvidia-smi wl: wayland-info
x11: xdpyinfo, xprop, xrandr
Drives:
Local Storage: total: 1.82 TiB used: 499.18 GiB (26.8%)
ID-1: /dev/sda vendor: Western Digital model: WDS100T2B0A-00SM50
size: 931.51 GiB
Partition:
ID-1: / size: 913.83 GiB used: 498.9 GiB (54.6%) fs: ext4 dev: /dev/dm-0
Swap:
ID-1: swap-1 type: file size: 24 GiB used: 0 KiB (0.0%) file: /swapfile
Sensors:
System Temperatures: cpu: 68.2 C mobo: 37.0 C
Fan Speeds (rpm): cpu: 0
Info:
Memory: total: 16 GiB available: 15.53 GiB used: 4.61 GiB (29.7%)
Processes: 337 Uptime: 55m Shell: fish inxi: 3.3.41
I dug into things a bit more, it appears that I’m having similar issues and outputs to this thread from the arch forum, but my understanding of computers is lagging behind a bit to wrap my head around it.
I also attempted downgrading the kernel and nvidia drivers with no change besides nvidia-smi outputting this now. Attempted to do eos-shifttime and ran into issues downgrading both cuda and pipewire-pulse
NVIDIA-SMI has failed because it couldn't communicate with the NVIDIA driver. Make sure that the latest NVIDIA driver is installed and running.