الباحثون

Bo Zhou

المنشورات 3

نسخة أولية وصول مفتوح

FREA: A Multi-Source Expert Benchmark for Reaction Feasibility Verification

Botao Yu, Bo Zhou, Daniel Adu-Ampratwum وآخرون · 2026

As generative models and AI agents propose chemical reactions at a scale beyond expert review, feasibility verifiers decide which proposals enter synthesis planning. But do their decisions agree with chemists across different kinds of candidates? We introduce FREA, a benchmark of 751 reactions labeled by expert chemist …

نسخة أولية وصول مفتوح

OSWorld-Science: A Benchmark of Computer Use Agents for Learning and Using Scientific Software

Dingyuan Dai, Heli Qi, Lei Liu وآخرون · 2026

Scientific software presents a demanding test for computer-using agents based on visual language models (VLMs): completing a research workflow requires interpreting specialized interfaces, manipulating scientific objects, and producing verifiable results. We thus introduce OSWorld-Science, a benchmark and evaluation en …

المؤلفون المشاركون