AI Market Needs Independent Scorekeepers, Vals Sees Opportunity

JenniferHli · x · 2026-08-14

The author argues that as AI models are increasingly applied to complex workflows like finance, legal research, and software development, traditional standardized benchmarks (like SATs and LSATs) can no longer prove real-world capability and safety. We need actual "driving tests on the road."

However, building evaluation systems for real expert workflows, auto-grading to an expert standard, and keeping pace with frontier model releases is technically very difficult. Just as credit and public markets need independent ratings and audit agencies, the AI market needs an independent scorekeeper. This is the market gap Vals aims to fill: providing a trusted compass for organizations to understand model capabilities.

Original post →

More from Companies & People

Companies & People channel →