Researcher's agent workflow: label 100 examples, have the agent scale to 10k pseudo-labels

ducha_aiki · x · 2026-10-08

Responding to skepticism that fine-tuning is simpler, duchaaiki describes his months-long workflow: hand-label 100 examples, have an agent find heuristics to extend to 10k pseudo-labels, train models on them, verify predictions manually, and let the agent refine the pseudo-labeling heuristics iteratively.

Related event: Researcher Uses Agents to Scale 100 Labels into 10K Training Samples(2 posts)→

Original post →

More from coding & agent

coding & agent channel →