← r/LocalLLaMA
▲
76
-1
32👁
r/LocalLLaMA · u/kirisoraa · 27d ago

Anybody use frontier models like Astra/Fable for planning/judging, and qwen3.8 as the main workhorse? Curious to hear about your setups!

Hey everyone!

I'm curious to hear from people that use a combination of cloud-based frontier models and local ones for development. I'm planning to set something similar up and wanted to hear about actual examples of this in action.

Currently my plan is to use my chatgpt plus subscription purely for planning and judging with Astra, and then run a local qwen3.8-27b model for the actual coding gruntwork - i.e Astra plans -> qwen implements -> Astra critiques the implementation -> qwen fixes and so on.
This way I keep cloud usage down and cheap, while retaining the high-parameter intelligence for architecture decisions and optimization.

For those of you who have a similar setup, how is it? How do you switch between the two, what harness/settings/etc? Anything you would suggest?

81 0 76 10/3 04:48 10/8 21:12 UTC
scorecomments32 sightings
first seen 2026-10-03 04:48 UTClast seen 2026-10-08 21:12 UTCscore then 77score now 76gained -1sightings 32
open on reddit ↗ 💬 79 (+1)