Authors

Liang Lin

Publications 6

Preprint Open access

FAST: Flow Any Scene Transformer

Scaling has become a primary driver of progress in language and vision foundation models, yet its role in precise correspondence matching remains underexplored. In this work, we present Flow Any Scene Transformer (FAST), a scalable correspondence model driven by two key insights. First, we reveal that the query-key pro …

Preprint Open access

Recovering the View: Benchmarking Physical Active Vision for Occlusion Recovery in Robotic Manipulation

Kaijun Luo, Yudi Huang, Qijun Zhong et al. · 2026

Physical active vision allows robots to change their viewpoint when task-relevant observations become unreliable, yet existing manipulation benchmarks provide limited support for studying how policies recover from occlusion during execution. We introduce BAVO-Bench (Bimanual Active Vision under Occlusion), a bimanual a …

Co-authors