Skip to content
arXiv cs.CL · Papers

Lost in Interpolation: Why Predictive Feedback Fails in Diffusion Language Models

arXiv:2608.06529v1 Announce Type: new Abstract: Soft-masking accelerates the convergence of Masked Diffusion Language Models (MDLMs). Existing formulations build this blend with linear interpolation (LERP) in the raw embedding space, which implicitly treats that space as Euclidean. We analyze the embedding space of MDL