CUDA: enable sparse fa for qwen4 by am17an · Pull Request #28770 · ggml-org/llama.cpp
Another day, another Qwen Flash Next speedup
scorecomments15 sightings
first seen 2026-10-03 04:45 UTClast seen 2026-10-05 07:29 UTCscore then 103score now 100gained -3sightings 15
open on reddit ↗
💬 13