Authors

Ruikang Zhang

Publications 1

Preprint Open access

Informed Masking: Structure-Aware Perturbation for Reinforcement Learning in Diffusion Large Language Models

Xiaoyi Yu, Enver Sangineto, Pei Fu et al. · 2026

Diffusion Large Language Models (dLLMs) have emerged as an efficient alternative to autoregressive models, yet aligning them via Reinforcement Learning (RL) requires likelihood surrogates estimated from masked reconstruction subproblems under a small Monte Carlo budget per rollout. Existing methods construct these subp …

Co-authors