← r/LocalLLaMA
▲
5
+5
15👁
r/LocalLLaMA · u/politefella0 · 9d ago

What’s better? A very small quant or a large model distilled into a small one?

I see people asking (begging, take it as humor) for 0.000001 bit quants but isn’t a lower quant essentially going to hurt model’s quality and tool calls?

5 0 5 10/3 06:29 10/7 05:47 UTC
scorecomments15 sightings
first seen 2026-10-03 06:29 UTClast seen 2026-10-07 05:47 UTCscore then 0score now 5gained +5sightings 15
open on reddit ↗ 💬 20 (+2)