GRPO token costs and multi-agent coordination mechanisms discussed
lukaszkaiser · x · 2026-08-28
Lukasz Kaiser discusses technical details on GRPO RL and agent coordination:
- GRPO Token Costs: In GRPO-like RL runs, additional token costs exist where each used token lowers the reward; looking up info to reduce token count pays off.
- Message Board Coordination: Responding to inquiries about the OpenAI message board incident, discussing the benefits of this loose coordination vs single agents/subagents, and whether it persists info across invocations.
Related event: Researchers Question OpenAI's Message Board Agent Coordination Experiment(2 posts)→
More from coding & agent
- OpenWiki adopts OKF 2.0 for page-level verification and provenance — BraceSproul · 2026-08-28
- OpenInstinct: Self-hostable iMessage AI assistant with browser control — arthurcolle · 2026-08-28
- Live Pipeline Builder Session: Building Data Pipelines on Demand — aronchick · 2026-08-28
- Plane Powers reveals agent mechanics: runs on work graph, not chat sidebar — JosephJacks_ · 2026-08-28
- Burning through Grok credits with OpenClaw integration — heyneighbor · 2026-08-28
- OpenPresence: A Framework for Deploying Autonomous Social Media Agents — RichardsonDx · 2026-08-28