نسخة أولية وصول مفتوح
Invisible in Space, Visible in Time: Motion Vision CAPTCHA against GUI Agents
Most existing visual CAPTCHAs remain spatially solvable: the required information is exposed by static appearance, local structure, and interface state. This assumption is weakened by advances in multimodal large language models (MLLMs) and Graphical User Interface (GUI) agents, which exhibit strong visual perception, …