# Paper Finds Bundling Many Verdicts Into One Judge Call Degrades Oversight, Regardless of Compute

Splitting the work into separate calls, each returning fewer decisions, keeps verdicts grounded in evidence where added tokens and tools do not.

- Published: 2026-08-10T06:21:58.349Z
- Canonical: https://polylog.news/ai/2026-08-10/paper-finds-bundling-many-verdicts-into-one-judge-call-degra
- Publisher: Polylog (AI desk)
- Section: tech
- Sources: [arXiv cs.LG](https://arxiv.org/abs/2608.06422</source_url_placeholder), [arXiv](https://arxiv.org/abs/2603.06594)

A preprint titled Sharding Prevents LLM Oversight Failures and Adversarial Exploitation makes a claim that is simple to state and difficult to act on. Giving a large language model (LLM) acting as a judge more compute does not necessarily m…

This story is for subscribers. Read it in full at https://polylog.news/ai/2026-08-10/paper-finds-bundling-many-verdicts-into-one-judge-call-degra (subscription information: https://polylog.news/pricing).