No RL pipeline trains proactive agents for workplace communication norms

menhguin · x · 2026-10-05

The author points out that humans must learn subtle workplace communication norms—no double texting, no pinging at 3am, how to handle being left on seen, whether and how to follow up—yet there is no known RL training pipeline teaching agents these skills.

Ironically, the one thing proactive agents do manage well is ML training runs, and they get extensive training for exactly that. The implication: agent capability distribution is shaped by training data, and high-frequency human scenarios like workplace communication are being neglected by current training regimes.

Original post →

More from AGI Musings

AGI Musings channel →