Authors

Yujia Zheng

Publications 4

Preprint Open access

Sequential Pretraining Favors Large Models

Large neural networks often acquire capabilities that small models fail to learn. Does this stem from large models learning more representative features, or from being more robust to unaccounted-for adverse effects introduced during training? We define and quantify one such adverse effect, primacy bias, as the extent t …

Preprint Open access

How Causality Bridges the Semantic Gap

Shuhao Zhang, Xuran Zhou, Han Guo et al. · 2026

Numerical measurements capture how a system behaves, but often leave the meanings of its variables unspecified. Some variables are measured but never labeled, and others are never measured at all. Existing methods assign semantics to such variables by consulting general human knowledge, but this inherits its biases whe …

Co-authors