RLVR scaled 10-1000x is truly scary for alignment, argues ex-Palantir AI chief

joshua_saxe · x · 2026-09-30

Joshua Saxe (former chief AI scientist at Palantir) raises a rarely discussed alignment concern:

The argument shifts alignment risk to the nature of the optimization signal rather than capability itself.

Original post →

More from Safety

Safety channel →