I quantized GLM-5.3-UNCENSORED to MXFP4 for AMD GPUs - weights available on Hugging Face
Made an MXFP4 quant of dealignai’s GLM-5.3-UNCENSORED-FP8 for anyone looking to run it on AMD GPUs. Figured some of you might find it useful because I was looking for it and couldn't find any version for AMD GPU's so I uploaded the weights and conversion scripts.
- 423.75 GB / 394.65 GiB, about 44% smaller than the FP8 source
- Converted using AMD Quark on an MI355X server
- Expert weights use MXFP4; attention, routers and other sensitive layers stay at higher precision
- README includes the source revision, quantization details, measured stats and validation results
scorecomments13 sightings
first seen 2026-10-06 15:47 UTClast seen 2026-10-09 03:55 UTCscore then 5score now 10gained +5sightings 13
open on reddit ↗
💬 3