الباحثون

Yao Zhang

المنشورات 6

نسخة أولية وصول مفتوح

Inverting Multi-Vector Visual Document Indices

Prevailing multi-vector visual document retrievers store each page as about a thousand patch vectors, often in vector databases run by a third party. Since no one can read a page from its vectors, this index is easily treated as less sensitive than the page. However, because the index keeps one vector per patch in rast …

نسخة أولية وصول مفتوح

Mapping E-textiles Design Pain Points and Generative AI Opportunities: Insights from Workshops in Shanghai and Winchester

Zhuchenyang Liu, Nianchong Qu, Yao Zhang وآخرون · 2026

E-textile design involves complex decisions across materials, sensor and actuator structures, fabrication, garment integration, and data processing. It typically requires iterative prototyping and testing, which are time- and labour-intensive, while few practitioners possess cross-disciplinary expertise across all rele …

نسخة أولية وصول مفتوح

Knit-Structure Effects on Electromechanical Metrics and Their Correlation with Joint-Angle Estimation Error in Knitted Strain Sensors

Knitted resistive strain sensors show strong promise for joint motion sensing in sports and rehabilitation, but the linkage between sensor design and in situ performance remains unclear. We investigate how knit structure and machine settings (e.g., stitch size) shape electromechanical properties and which metrics predi …

نسخة أولية وصول مفتوح

ColNanoVDR: Document-Free Query Distillation for Multi-Vector Visual Document Retrieval via Optimal Transport

Zhuchenyang Liu, Ziyi Wang, Yao Zhang وآخرون · 2026

Multi-vector retrievers built on vision-language models lead visual document retrieval (VDR), but they run a multi-billion-parameter query encoder on every search. Distilling this encoder into a small student that queries the teacher's existing index would remove the bottleneck. The standard recipe, however, matches th …

نسخة أولية وصول مفتوح

The Earth in One Gaze: Training-Free Active Focus for UHR Remote Sensing Understanding

Yao Zhang, Pengyu Dai, Wei Guo وآخرون · 2026

Multimodal large language models (MLLMs) must balance local detail against scene context when interpreting ultra-high-resolution (UHR) remote sensing (RS) imagery within a limited visual-input budget. Existing selection-based methods either prune tokens and select patches through relevance scoring, or crop actively thr …

المؤلفون المشاركون