Prime Intellect presents verifier brittleness work at COLM, hiring across research

willcb · x · 2026-10-06

Prime Intellect will present research on verifier brittleness and reward hacking in long-horizon agents at the AIMS workshop on Oct 9 during COLM week (arXiv coming soon). The company is hiring across Applied Research & Research, targeting agentic RL / long-horizon post-training, environments, evals, verifiers, reward design & alignment, agent data and training/inference infra, and multi-agent systems. It's also hosting a COLM Research Happy Hour at its SF HQ on Oct 7, 6–9 PM PDT, with lightning talks and networking (registration full, waitlist open).

Original post →

More from Companies & People

Companies & People channel →