← r/LocalLLaMA
▲
112
+108
34👁
r/LocalLLaMA · u/jacek2023 · 6d ago

qwen4exp : halve the indexer score memory by ServeurpersoCom · Pull Request #29825 · ggml-org/llama.cpp

Qwen Flash Next now uses less VRAM

112 0 112 10/3 06:28 10/7 07:41 UTC
scorecomments34 sightings
first seen 2026-10-03 06:28 UTClast seen 2026-10-07 07:41 UTCscore then 4score now 112gained +108sightings 34
open on reddit ↗ 💬 14 (+12)