← r/LocalLLaMA
▲
100
-3
15👁
r/LocalLLaMA · u/jacek2023 · 19d ago

CUDA: enable sparse fa for qwen4 by am17an · Pull Request #28770 · ggml-org/llama.cpp

Another day, another Qwen Flash Next speedup

107 0 100 10/3 04:45 10/5 07:29 UTC
scorecomments15 sightings
first seen 2026-10-03 04:45 UTClast seen 2026-10-05 07:29 UTCscore then 103score now 100gained -3sightings 15
open on reddit ↗ 💬 13