Are you running Qwen 3.8 27b or Qwen Flash Next?
Curious about what people are preferring, if you have the hardware. I have m3 Max 96gb and both run, and largely feel identical, but prefill on qwen 27b is faster. Is there anything / anyone working on anything to improve pp with mlx?
Branching question: is anyone working on a harness that works with no reasoning? This interests me ever since Jetbrains shared that they're using 3.6 with reasoning off entirely: https://blog.jetbrains.com/junie/2026/08/qwen-for-junie/
Feel like there must be something neat with using one model to orchestrate, with reasoning, and subagent without reasoning.
scorecomments18 sightings
first seen 2026-10-03 04:43 UTClast seen 2026-10-07 06:56 UTCscore then 161score now 161gained 0sightings 18
open on reddit ↗
💬 248