Redditor Post-Trains Yandex AliceAI-80B-A3B From Scratch on 3 Local V100s, Update #5

jjusko20 · reddit · 2026-10-06

Reddit user jjusko20 shares update #5 of their open post-training project: an instruct finetune of yandex/AliceAI-80B-A3B-Base targeting agentic work and conversation, with a shallow distill of Qwen 3.8 27B to teach chain-of-thought reasoning.

Key details:

A rare hands-on log of low-budget local finetuning of a large MoE model.

Original post →

More from Models

Models channel →