الباحثون

Zijian Chen

المنشورات 3

نسخة أولية وصول مفتوح

From Pixel to Coding: Evaluating the Figure Reproduction Capabilities of MLLMs

Zijian Chen, Zhengyu Chen, Bohan Liang وآخرون · 2026

Multimodal Large Language Models (MLLMs) have demonstrated impressive capabilities in both visual understanding and code generation. However, existing benchmarks typically evaluate these two modalities in isolation, lacking a dedicated assessment of their unification, i.e., how a model can perceive complex visual struc …

نسخة أولية وصول مفتوح

From Learner Behavior to Reusable Skills for Effective and Efficient Learner Simulation

Zijian Chen, Zheng Zhang, Miao Jia وآخرون · 2026

Learner simulation aims to reproduce how a particular learner behaves on new tasks. Although Large Language Models (LLMs) can generate increasingly fine-grained learning behaviors, existing approaches often need to repeatedly process a growing interaction history to reconstruct the learner. This introduces additional c …

نسخة أولية وصول مفتوح

Invisible in Space, Visible in Time: Motion Vision CAPTCHA against GUI Agents

Zeyu Zhang, Dingyi Rong, Zijian Chen وآخرون · 2026

Most existing visual CAPTCHAs remain spatially solvable: the required information is exposed by static appearance, local structure, and interface state. This assumption is weakened by advances in multimodal large language models (MLLMs) and Graphical User Interface (GUI) agents, which exhibit strong visual perception, …

المؤلفون المشاركون