llama: add Maple 20B-A1B ternary MoE architecture (CPU) by AlexGabbia · Pull Request #27000 · ggml-org/llama.cpp
20B-A1B model is coming, good for low VRAM people?
scorecomments21 sightings
first seen 2026-10-03 04:43 UTClast seen 2026-10-08 09:01 UTCscore then 155score now 152gained -3sightings 21
open on reddit ↗
💬 46