← r/LocalLLaMA
▲
0
-3
26👁
r/LocalLLaMA · u/soyalemujica · 6d ago

Thanks to Strata I have quit 27b for Qwen Flash (24gb VRAM plus 64gb ram)

Using Strata on a 7900xtx plus 64 gb ddr5 ram, 60t per second even at 250k context, can finally use my pc while working AI in the background, even game as well, smarter and more precise than dense model, it follows orders more accurately, follows plan more versatile, it goes around doing a lot of tests for tasks I request in frontend and also in backend.

The best thing is that it's faster, I can fit more context at q8 precision, it's smarter and I can get to use my pc without worrying about an OOM error due to dense model.

I no longer have to use Linux as well, it's working as fast in Windows 11 as it did in Linux.

I use it with a 6gb VRAM reserve so I can have Windows 11 with 4gb available.

Edit:

The "people" saying I am a bot, or that people commenting are bots, are completely clueless, seriously, even down voting something that benefits ALL of us.

3 0 0 10/3 11:26 10/8 16:05 UTC
scorecomments26 sightings
first seen 2026-10-03 11:26 UTClast seen 2026-10-08 16:05 UTCscore then 3score now 0gained -3sightings 26
open on reddit ↗ 💬 57 (+56)