Esfahani, M. M., Salem, S., Alser, M., & Calhoun, V. (2026). Do Vision Models Learn Physical Constraints or Rendering Shortcuts? A Counterfactual Benchmark for Grounded Physical Consistency. https://omanscience.com/en/articles/do-vision-models-learn-physical-constraints-or-rendering-shortcuts-a-counterfactual-benchmark-for-grounded-physical-consistency