A New Paper Proposes Sparse, Block-Denoising Diffusion to Cut Language-Model Inference Cost · Polylog