← r/LocalLLaMA
▲
4
+1
2👁
r/LocalLLaMA · u/ParaboloidalCrest · 17h ago

Llama.cpp-vulkan: What's the best strategy to use iGPU alongside dGPU(s)

...without slowing everything to a halt?

Edit: I realize this might not be clear, but I'm refering to the iGPU within a consumer CPU, eg Ryzen 9950x, rather than Halo.

For example, is there a kind of buffer or operation that could be safely and specifically offloaded to iGPU's RAM? And how?

4 0 4 10/8 15:32 10/8 19:30 UTC
scorecomments2 sightings
first seen 2026-10-08 15:32 UTClast seen 2026-10-08 19:30 UTCscore then 3score now 4gained +1sightings 2
open on reddit ↗ 💬 19 (+11)