A Verse code corpus could reshape pretraining, but a model-specific rubric may poison it
mike64_t · x · 2026-07-28
The author worries that Epic’s new scripting language Verse may be too elegant for many humans to actually use, which could mean its code never becomes a substantial part of pretraining data.
They then raise a broader concern: if the base corpus of source code is never shaped by real human preferences, a model-specific rubric could end up contaminating the training pool and create winner-take-all lock-in effects.
More from AGI Musings
- A blunt pro-distillation argument says shared knowledge should not stay locked up — ns123abc · 2026-07-28
- The phrase “Artificial Superintelligence” may be losing steam, prompting new AI-era labels — kellerjordan0 · 2026-07-28
- Sebastien Bubeck calls AI the most profound scientific revolution in a century — danintheory · 2026-07-28
- AI models will become every company’s new website, but data flywheels will decide the winners — aigclink · 2026-07-28
- Anduril’s containerized data center can be deployed by two people in under 10 minutes — damianplayer · 2026-07-28
- Who keeps human values relevant when AI starts defining “optimal” decisions? — PierceLilholt · 2026-07-28