Update: Yandex/AliceAI 80B-A3B fine tune progress
loss curve \(taken from the last micro of every step, to explain the variation\)
some help from gemini 3.8 flash high
About 40% of the way done with the initial fine tune. The loss is so spiky because I accidentally used the last loss of each micro, rather than the average of each step
The training live stream is at: https://figure-bios-expect-cio.trycloudflare.com/ \- and it allows you to inspect any and all of the training data I'm using, if you're interested - I can also provide those roughly 3.5k examples as a dataset. It was generated from sftmill
Original thread: https://www.reddit.com/r/LocalLLaMA/comments/1wslgkw/watch\_me\_posttrain\_aliceaifoundation80ba3b\_from/
scorecomments22 sightings
first seen 2026-10-03 06:29 UTClast seen 2026-10-07 15:55 UTCscore then 51score now 54gained +3sightings 22
open on reddit ↗
💬 8