← r/LocalLLaMA
▲
57
-1
23👁
r/LocalLLaMA · u/Character-Result-281 · 21d ago

Built this yesterday with Qwen3.8-Flash-Next (NVFP4, 262K context) on a single NVIDIA DGX Spark

post image

Planning, coding, testing = 8h total.
Stack: VSCode Copilot in autopilot mode + SGLang
Stats: ∼10k lines generated, ∼800k tokens consumed

Sure, it's not GPT-6 Astra level, but for a 100% local ∼180B MoE running on a single DGX Spark at ∼35 tok/s. Not bad...

61 0 57 10/3 04:50 10/9 03:25 UTC
scorecomments23 sightings
first seen 2026-10-03 04:50 UTClast seen 2026-10-09 03:25 UTCscore then 58score now 57gained -1sightings 23
open on reddit ↗ 💬 28