ML engineer: writing a bs=1 trainer myself beats explaining it to an LLM

cephaloform · x · 2026-10-01

User cephaloform argues that implementing a bs=1 version of a trainer is easier than explaining how to do it to an LLM — tacit engineering intuition LLMs can't yet absorb. In the follow-up he offers the flip side: agents shine at parallelizing whatever wacky bs=1 trainer you come up with, which used to be the most annoying part of multi-component training setups. A crisp take on where LLMs help and fail in ML engineering.

Related event: Developer: Build a bs=1 Trainer and Let AI Parallelize It(4 posts)→

Original post →

More from coding & agent

coding & agent channel →