← r/LocalLLaMA
▲
168
+140
44👁
r/LocalLLaMA · u/KnownAd4832 · 3d ago

Qwen3.8-Flash-Next on Strata

post image

Hey! 👋

I have released an official support for Strix Halo machines on Strata for Qwen3.8-Flash-Next.

Currently numbers are the best on long context decode and ppts using typical Unsloth’s Q4 and GSQ-RCO model weights.

Can go up to 1M context length without big speed loss. Currently support is marked as experimental and was done on Linux only.

https://github.com/Niko1221/Strata/

Will be happy for any feedback and pull requests you could give! 👀

170 0 168 10/6 15:47 10/9 07:04 UTC
scorecomments44 sightings
first seen 2026-10-06 15:47 UTClast seen 2026-10-09 07:04 UTCscore then 28score now 168gained +140sightings 44
open on reddit ↗ 💬 85 (+70)