Polylog
The Polylog AI Intelligence Brief

Morning Edition · Saturday, August 1, 2026Published at 1:44 AM EDT · New York

A GeneBench-Pro Co-Author Leaves OpenAI to Build a Startup Selling Reinforcement-Learning Data

The new venture aims to generate specialized datasets for post-training, on the argument that data quality now limits model gains more than scale.

A GeneBench-Pro Co-Author Leaves OpenAI to Build a Startup Selling Reinforcement-Learning Data

According to the AI ML Big Data channel, Andrew Ho, a co-author of the GeneBench-Pro benchmark, has left OpenAI to start a company focused on generating high-quality reinforcement-learning datasets for training large language models. The st…

Continue the AI Intelligence Brief

Track frontier labs, chips, export controls, model releases, regulation, and AI infrastructure.

  • 5 AI intelligence signals a day
  • Frontier labs, compute, and chips
  • Model releases and AI infrastructure
  • Source-grounded analysis with confidence labels

The Global Intelligence Brief stays free.

Part of a tracked trend

Post-Training Data Becomes the Bottleneck

As gains from raw scale flatten, high-quality post-training and reinforcement-learning data becomes the scarce input, and a specialized vendor market will form to supply it.