← r/LocalLLaMA
▲
89
+6
44👁
r/LocalLLaMA · u/pubudeux · 10d ago

First few days of qwen3.8-flash-next on 4x R9700 - it's been really interesting so far

post image

Here's a metric dashboard giving an idea of the last few days.

Been testing with a variety of different agentic coding use-cases, mostly using a pi harness.

qwen3.8-flash-next has seriously exceeded my expectations (used https://huggingface.co/tcclaviger/Qwen3.8-Flash-Next-MXFP4-FP8)

Both speed and quality have surprised me, given that I can get 3-5 concurrent streams going with \~100t/s gen each, and single stream easily gets to 150+t/s. Prefill is 10k+t/s

89 0 89 10/3 04:47 10/8 07:10 UTC
scorecomments44 sightings
first seen 2026-10-03 04:47 UTClast seen 2026-10-08 07:10 UTCscore then 83score now 89gained +6sightings 44
open on reddit ↗ 💬 67 (+2)