Abstract

Medical question answering spans diverse specialties and modalities, and individual medical large language models (LLMs) exhibit distinct strengths across tasks and domains. This heterogeneity suggests that combining specialists may enable broader coverage of medical questions than relying on any single model. However, existing LLM routing methods primarily seek to balance answer quality and inference cost, leaving open how to exploit differences in specialist competence to improve medical reasoning. In this paper, we introduce MedRouter, an agentic system that uses an embedding-based multi-label router to select and query specialist LLMs, then passes their responses to a generator to produce the final answer. We further propose SCALE (Specialist Competence-Aware Learning), a two-stage training framework that first trains the Router with specialist correctness supervision and then optimizes its selections through reinforcement learning. The second stage uses a Performance Gain Reward (PGR) that measures how specialist information affects the generator's answer correctness relative to answering without that information. Experiments on eight text-based and multimodal medical QA benchmarks show that MedRouter outperforms the strongest routing baseline by 8% in average accuracy. Our analysis of specialist outputs further reveals distinct strengths and complementary question-level coverage, motivating learned routing to combine these capabilities for more comprehensive medical reasoning.

Keywords

Subject

Publication details

Journal
Not available
Open access
Green open access

Cite this article

APA 7

Cao, L., Lu, B., Shen, Y., & Guo, Y. (2026). MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning. https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning

MLA 9

Cao, Lang, et al. "MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning." https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning.

Chicago (author–date)

Cao, Lang, Binghang Lu, Yuhao Shen, and Yue Guo. 2026. "MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning." https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning.

Harvard

Cao, L., Lu, B., Shen, Y. and Guo, Y. (2026) 'MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning', Available at: https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning.

Vancouver

Cao L, Lu B, Shen Y, Guo Y. MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning. https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning

IEEE

L. Cao, B. Lu, Y. Shen, and Y. Guo, "MedRouter: Demystifying Knowledge Differences Across Medical LLMs for Routing-Based Reasoning," https://omanscience.com/en/articles/medrouter-demystifying-knowledge-differences-across-medical-llms-for-routing-based-reasoning.