Wang, Z., Xie, M., Liu, Y., Wu, D., Li, Y., & Dai, P. (2026). Seeing and Solving Are Not Enough for Vision-Language Models. https://omanscience.com/ar/articles/seeing-and-solving-are-not-enough-for-vision-language-models