Coding models are running out of data — PL researchers propose 'intent computing' as the fix

LingmingZhang · x · 2026-09-03

Grigore Rosu warns that LLM coding ability depends heavily on code data, and the industry has essentially consumed all high-quality code data for training and evaluation — progress is visibly slowing. Whoever finds the next big data pool holds the key.

His group, rooted in programming languages and formal methods, proposes a new paradigm called intent computing:

The core idea: formally verified code-plus-proof artifacts could be the next major pool of high-quality training data.

Original post →

More from coding & agent

coding & agent channel →