← r/LocalLLaMA
▲
10
+5
13👁
r/LocalLLaMA · u/bakatristan · 3d ago

I quantized GLM-5.3-UNCENSORED to MXFP4 for AMD GPUs - weights available on Hugging Face

Made an MXFP4 quant of dealignai’s GLM-5.3-UNCENSORED-FP8 for anyone looking to run it on AMD GPUs. Figured some of you might find it useful because I was looking for it and couldn't find any version for AMD GPU's so I uploaded the weights and conversion scripts.

Download on Hugging Face

  • 423.75 GB / 394.65 GiB, about 44% smaller than the FP8 source
  • Converted using AMD Quark on an MI355X server
  • Expert weights use MXFP4; attention, routers and other sensitive layers stay at higher precision
  • README includes the source revision, quantization details, measured stats and validation results
11 0 10 10/6 15:47 10/9 03:55 UTC
scorecomments13 sightings
first seen 2026-10-06 15:47 UTClast seen 2026-10-09 03:55 UTCscore then 5score now 10gained +5sightings 13
open on reddit ↗ 💬 3