Authors

Jin Ma

Publications 3

Preprint Open access

Learning Multimodal Embeddings with Evidence-Aligned Readout

Zirong Chen, Fuda Ye, Enjun Du et al. · 2026

Multimodal large language models can expose task-relevant evidence through generation, but producing useful evidence does not by itself determine how it enters a retrieval embedding. We study whether the semantic organization of that evidence can also specify where representations are read. To address this question, we …

Co-authors