← r/LocalLLaMA
▲
0
 
17👁
r/LocalLLaMA · u/mcgeezy-e · 7d ago

Gemma 4 31b on 32Gb of vram (3 cards)

Looking for some points on if i can realistically get this model to run via vllm (preferably as it will be multi access at times) on 3 cards. 1 3070 8gb and 2x 3060 12gb.

If not, other suggestions?

1 0 0 10/3 06:28 10/8 16:15 UTC
scorecomments17 sightings
first seen 2026-10-03 06:28 UTClast seen 2026-10-08 16:15 UTCscore then 0score now 0gained 0sightings 17
open on reddit ↗ 💬 14 (+5)