Balmaseda, V., Lin, C. L., & Yang, T. (2026). From Pixels, Without Pre-training: Joint Generative and Self-Supervised Representation Learning in One Model. https://omanscience.com/ar/articles/from-pixels-without-pre-training-joint-generative-and-self-supervised-representation-learning-in-one-model