← r/LocalLLaMA
▲
0
 
10👁
r/LocalLLaMA · u/AIFrontierReads · 11d ago

Laya: replace LLM-as-a-judge with a 322M-parameter decision engine (26,639 stars in 9 days, hands-on test)

It turns decisions — routing, triage, yes/no calls — into typed outputs from a small model instead of generated text, with a routing-only CLI, triage presets, and an abstention gate when confidence falls below a threshold. I ran through the tutorial on CPU end to end, including a French ticket classification, and with min\_confidence=0.90 it abstained on one case it would otherwise have misclassified — the honest highlight. Warm latency was about 0.7s per question on CPU; the calibration caveat (over-confident checkpoints) is worth knowing before trusting the scores blindly.

1 0 0 10/3 06:31 10/4 17:44 UTC
scorecomments10 sightings
first seen 2026-10-03 06:31 UTClast seen 2026-10-04 17:44 UTCscore then 0score now 0gained 0sightings 10
open on reddit ↗ 💬 5