Authors

Shrikanth Narayanan

Publications 3

Preprint Open access

Mitigating Accent-Language Confusion in Self-Supervised Speech Representations for Language Identification

Spoken language identification (LID) aims to recognize the target language regardless of accent. In practice, however, LID models fine-tuned from self-supervised speech representations frequently confuse accents with languages, misclassifying non-native (L2) speech as the speaker's first language (L1). We show that non …

Preprint Open access

Paired Multimodal Scaling Laws

Existing multimodal scaling laws fit multimodality terms empirically after testing and never vary how much data is multimodally paired at fixed data budgets. We investigate how, under the same total data per modality, changing the number of paired data affects loss curves in multimodal classification tasks. We train mo …

Co-authors