The Polylog AI Intelligence Brief

Morning Edition · Tuesday, August 4, 2026

Alibaba Ships Qwen3.8-Max at 2.4 Trillion Parameters and Promises the Weights Next Week

Tech

Alibaba Ships Qwen3.8-Max at 2.4 Trillion Parameters and Promises the Weights Next Week

The model is priced at $2 and $6 per million input and output tokens, roughly a fifth of Anthropic's top tier, and Alibaba says a Max-class Qwen will be downloadable for the first time.

1 source
Corroborated
OpenAI Publishes Lean-Checked Proofs for Ten Open Mathematics Problems From an Unreleased Model

Tech

OpenAI Publishes Lean-Checked Proofs for Ten Open Mathematics Problems From an Unreleased Model

The company puts the compute cost at about $2,000 and released machine-checkable certificates. No result has been through refereed review, and OpenAI is judging the novelty of its own work.

1 source
Corroborated
OpenAI Publishes Internal Messages to Rebut Apple's Trade-Secret Suit

Tech

OpenAI Publishes Internal Messages to Rebut Apple's Trade-Secret Suit

Apple's July complaint names Tang Tan, a former design leader, and Chang Liu, an engineer, and alleges recruiting aimed at unreleased hardware. The case reaches OpenAI's consumer device plans.

1 source
Plausible
Tencent's Hyra Agent Claims Record Results on 29 of 55 Open Mathematics Problems

Tech

Tencent's Hyra Agent Claims Record Results on 29 of 55 Open Mathematics Problems

The discovery agent runs on the openly released Hy3 model, 295 billion parameters with 21 billion active, and publishes its research artifacts to a public repository.

1 source
Corroborated
OpenAI Details the Full-Duplex Stack Behind GPT-Live, Built in Six Months

Tech

OpenAI Details the Full-Duplex Stack Behind GPT-Live, Built in Six Months

The system removes the separate voice-activity detector entirely and trains a speech-native model that listens and speaks at the same time.

1 source
An Open-Source Runtime Streams Mixture-of-Experts Weights From SSD to Run an 80-Billion-Parameter Qwen on a Mac

Tech

An Open-Source Runtime Streams Mixture-of-Experts Weights From SSD to Run an 80-Billion-Parameter Qwen on a Mac

Swiftlet keeps only the dense core in memory and fetches each routed expert with a single disk read, putting a 35-billion-parameter model on an iPhone at roughly one token per second.

1 source
Meta Sells Its Best Model by the Token After Years of Giving Weights Away

Tech

Meta Sells Its Best Model by the Token After Years of Giving Weights Away

Muse Spark 1.1 reaches outside developers through the Meta Model API at $1.25 and $4.25 per million tokens, undercutting the frontier tiers it is aimed at.

2 sources
New Papers Push Back on Paying Frontier Prices to Grade Model Output

Tech

New Papers Push Back on Paying Frontier Prices to Grade Model Output

One study asks whether cheap open-weight models can judge natural-language mathematical proofs reliably, as forecasters put the model evaluation tools market near $1.15 billion in 2025.

3 sources
Meta's Segmentation and Vision Models Move Into Assistive Robotics and National Laboratory Science

Tech

Meta's Segmentation and Vision Models Move Into Assistive Robotics and National Laboratory Science

The University of Pittsburgh applies Segment Anything and DINO to assistive manipulation, while Lawrence Berkeley National Laboratory uses the same models in early Genesis Mission projects.

2 sources
Researchers Propose an Executable Benchmark for the Decisions Agents Make Before They Answer

Tech

Researchers Propose an Executable Benchmark for the Decisions Agents Make Before They Answer

Two arXiv papers target the choices an agent makes before answering (answer directly, decompose, retrieve, run code, delegate, verify or recover), which drive most of its cost.

2 sources
Agent Tooling Turns Toward Traces and Verified Skill Claims

Tech

Agent Tooling Turns Toward Traces and Verified Skill Claims

Yandex adds execution-trace analysis to its AI Studio, while an arXiv paper proposes registering agent capabilities on a ledger to close the gap between what a third-party skill claims and what it does.

2 sources
OpenAI Publishes a Telco Deployment With Revenue Numbers Attached

Tech

OpenAI Publishes a Telco Deployment With Revenue Numbers Attached

The operator Circles reports average revenue per user up 22 percent and churn down 9 percent using the OpenAI API and Codex, figures OpenAI published and no outside party has audited.

1 source
Plausible