← r/LocalLLaMA
▲
205
+3
22👁
r/LocalLLaMA · u/netikas · 29d ago

GigaChat-3.5-Reasoning

Hey y'all!

We've released a new model in our lineup: GigaChat-3.5 Reasoning. It's a 432B-A28B MoE with Gated DeltaNet for long-context efficiency.

We trained domain experts (code, math, general, etc.) with CISPO and then distilled them into a single model via on-policy distillation.

In our evals the resulting model lands close to DeepSeek V4 Flash Preview while using 37% fewer tokens in its reasoning traces.

Weights are on Hugging Face under MIT: https://huggingface.co/collections/ai-sage/gigachat-35-reasoning. You can also try it at giga.chat — pick the reasoning tab (rightmost one).

206 0 205 10/3 04:20 10/9 01:03 UTC
scorecomments22 sightings
first seen 2026-10-03 04:20 UTClast seen 2026-10-09 01:03 UTCscore then 202score now 205gained +3sightings 22
open on reddit ↗ 💬 39