Evaluation Becomes a Standalone Software Market
Model evaluation, grading and benchmarking separate from the labs into an independent tooling and services market, with cheap open-weight judges collapsing the unit cost of grading; expect recurring vendor launches, funding rounds and consolidation in eval infrastructure as buyers demand verification they do not have to take from the model vendor.
forming · confidence 38 · Emerging (watchlist) · tracking since August 4, 2026 · updated August 4, 2026
Why the conviction moved
- Aug 4Strengthened
A published forecast sizes model evaluation and benchmarking tools at roughly $9.6 billion by 2035 against about $1.15 billion last year, and new research shows cheap open-weight models can grade proofs in place of expensive frontier graders. Both the demand-side forecast and the supply-side cost collapse point at an independent eval vendor layer.
Source trail
Supporting · August 4, 2026
AI Evaluation Turns Into a Product Market as Researchers Show Cheap Open Models Can Grade Proofs
A published forecast sizes model evaluation and benchmarking tools at roughly $9.6 billion by 2035 against about $1.15 billion last year, and new research shows cheap open-weight models can grade proofs in place of expensive frontier graders. Both the demand-side forecast and the supply-side cost collapse point at an independent eval vendor layer.
AI Post (Telegram)
Unlock full source trail, score history, and daily updates.
Unlock Trends