Dev Fine-Tuned a 2B LLM on WhatsApp Group Chat, Simulating Six Friends on an M1 Pro

BarisSayit · reddit · 2026-09-12

A developer spent months fine-tuning a 2B local model on his own six-person WhatsApp group chat, training and running it on an M1 Pro to simulate the whole group. All experiments were in Turkish.

How good is it:

On paper: Using a judge-LLM, human-anchored evaluation, the best version achieved an 80% human win rate in human-vs-model tests (ideal would be <50%).

The author open-sourced the reproducible local pipeline, chat UI, evaluation method, results, and an experiment PDF — but not the private chat data or fine-tuned weights — and reminds you to ask for consent before training on others' chats.

Original post →

More from Fun

Fun channel →