# AI Evaluation Turns Into a Product Market as Researchers Show Cheap Open Models Can Grade Proofs

One forecast puts model evaluation and benchmarking tools at roughly $9.6 billion by 2035, up from about $1.15 billion last year. New research tests whether small open-weight judges can replace expensive frontier graders.

- Published: 2026-08-04T06:03:13.954Z
- Canonical: https://polylog.news/ai/2026-08-04/ai-evaluation-turns-into-a-product-market-as-researchers-sho
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [Polylog editors](https://polylog.news), [arXiv cs.CL](https://arxiv.org/abs/2608.00004), [arXiv cs.CL](https://arxiv.org/abs/2608.00005)

The tooling layer around model evaluation is being priced as an industry rather than treated as a research chore. Precedence Research values model evaluation and benchmarking tools at about $1.15 billion in 2025 and projects roughly $9.57 b…

This story is for subscribers. Read it in full at https://polylog.news/ai/2026-08-04/ai-evaluation-turns-into-a-product-market-as-researchers-sho (subscription information: https://polylog.news/pricing).