Ir al contenido principalSaltar al contenido

CaRE Compute-aware Remasking Evaluation Protocol for Masked Diffusion Language Models

Abstract

arXiv:2607.24763v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are advancing rapidly, yet the evaluation standards needed to reliably interpret their progress have not kept pace. Despite MDLMs becoming competitive with autoregressive language models, seven recent remasking papers evaluate under incompatible settings, varying nominal step counts, metrics, and sampling temperatures without jointly controlling these factors, rendering their strategy rankings largely incomp

Más info: Política de Cookies