← r/LocalLLaMA
▲
0
 
8👁
r/LocalLLaMA · u/dampflokfreund · 28h ago

Qwen Flash Q2_0 vs IQ2_XS GSQ-RCO using Strata

Hello,

so lately I have been testing those two quants, since those are the ones that run decently enough on my old laptop. IQ3 destroys prefill.

I have noticed IQ2\_XS definately has better preserved world knowledge, but is much more prone to looping than q2\_0 at the same recommended sampler settings, especially without thinking.

What are your experiences running these quants? The benchmarks are also pretty interesting, there's some clear advantages for q2\_0 but also for iq2\_xs in specific areas.

1 0 0 10/8 07:28 10/9 07:36 UTC
scorecomments8 sightings
first seen 2026-10-08 07:28 UTClast seen 2026-10-09 07:36 UTCscore then 0score now 0gained 0sightings 8
open on reddit ↗ 💬 6 (+5)