الباحثون

Pei Fu

المنشورات 2

نسخة أولية وصول مفتوح

Xiaomi-OCR-0 Technical Report

Xin Chen, Anan Du, Feng Feng وآخرون · 2026

Compact OCR-specific vision-language models achieve strong document parsing performance, but often rely on costly supervision and focus primarily on visual-text reconstruction. We introduce Xiaomi-OCR-0, a unified 0.8B model for document parsing and OCR-centric understanding. We build an approximately 170M-sample OCR-c …

نسخة أولية وصول مفتوح

Informed Masking: Structure-Aware Perturbation for Reinforcement Learning in Diffusion Large Language Models

Xiaoyi Yu, Enver Sangineto, Pei Fu وآخرون · 2026

Diffusion Large Language Models (dLLMs) have emerged as an efficient alternative to autoregressive models, yet aligning them via Reinforcement Learning (RL) requires likelihood surrogates estimated from masked reconstruction subproblems under a small Monte Carlo budget per rollout. Existing methods construct these subp …

المؤلفون المشاركون