← r/LocalLLaMA
▲
5
-1
13👁
r/LocalLLaMA · u/Competitive-Scar-627 · 7d ago

Model weight inferencing

I have 4050 6gb gpu, 24 gb ram which model should i choose to run i need speed. i try qwen 3.8 27b and feel too slow tried from onslot studio.
I have heard of weight inferencing does it helpful what should i do to try weight inferencing.

6 0 5 10/3 06:28 10/8 12:10 UTC
scorecomments13 sightings
first seen 2026-10-03 06:28 UTClast seen 2026-10-08 12:10 UTCscore then 6score now 5gained -1sightings 13
open on reddit ↗ 💬 23 (+2)