Abstract

Recent research on Multimodal Sentiment Analysis (MSA) has focused on learning from language, visual, and acoustic modalities with incomplete data to infer human sentiment. Most studies typically compensate for missing information by reconstructing modality features or designing complicated fusion mechanisms. However, these methods still suffer from spurious generation and noisy guidance due to the lack of high-level semantic grounding in partially observed multimodal evidence. To address these issues, we propose SemMSA, a latent semantic-aided framework that constructs rich sentiment-relevant semantics with LLMs, fully integrating with all modalities via anchor-free spectral alignment. It mainly consists of Cross-modal Semantic Refinement (CSR) and Cross-modal Spectral Alignment (CSA). Specifically, CSR first adaptively extracts visual and acoustic representations by corresponding adapters to form a unified multimodal prefix with language in the frozen LLM embedding space. It then iteratively produces continuous discriminative semantic states through a token-efficient latent refinement process without decoding explicit text. Next, CSA simultaneously aligns the refined semantics with all modalities by enhancing the dominant spectral component of their kernel Gram matrix. This captures global nonlinear dependencies among all representations without relying on a predefined anchor modality. In addition, an instance-level spectral separation constraint preserves cross-sample discriminability and mitigates representation collapse. Extensive experiments on SIMS, MOSI, and MOSEI benchmarks demonstrate that SemMSA achieves state-of-the-art performance.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Li, W., Wu, Z., Xiao, C., & Wang, Q. (2026). SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data. https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data

MLA 9

Li, Wenhao, et al. "SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data." https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data.

Chicago (author–date)

Li, Wenhao, Zhibin Wu, Chong Xiao, and Qiangchang Wang. 2026. "SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data." https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data.

Harvard

Li, W., Wu, Z., Xiao, C. and Wang, Q. (2026) 'SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data', Available at: https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data.

Vancouver

Li W, Wu Z, Xiao C, Wang Q. SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data. https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data

IEEE

W. Li, Z. Wu, C. Xiao, and Q. Wang, "SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data," https://omanscience.com/en/articles/semmsa-latent-semantic-aided-robust-multimodal-sentiment-analysis-with-incomplete-data.