← r/LocalLLaMA
▲
161
 
18👁
r/LocalLLaMA · u/Zeeplankton · 32d ago

Are you running Qwen 3.8 27b or Qwen Flash Next?

Curious about what people are preferring, if you have the hardware. I have m3 Max 96gb and both run, and largely feel identical, but prefill on qwen 27b is faster. Is there anything / anyone working on anything to improve pp with mlx?

Branching question: is anyone working on a harness that works with no reasoning? This interests me ever since Jetbrains shared that they're using 3.6 with reasoning off entirely: https://blog.jetbrains.com/junie/2026/08/qwen-for-junie/

Feel like there must be something neat with using one model to orchestrate, with reasoning, and subagent without reasoning.

165 0 161 10/3 04:43 10/7 06:56 UTC
scorecomments18 sightings
first seen 2026-10-03 04:43 UTClast seen 2026-10-07 06:56 UTCscore then 161score now 161gained 0sightings 18
open on reddit ↗ 💬 248