الباحثون

شين سو

المنشورات 3

نسخة أولية وصول مفتوح

Discriminating Fixture Coverage in Agent-Infrastructure Verification Suites

Invariant suites and runtime monitors increasingly gate agent deployment decisions, and the evidence offered for any particular suite is almost always a single observation: it passes an implementation believed correct and fails one believed broken. We measure what that observation is worth. Applying mutation analysis t …

نسخة أولية وصول مفتوح

Hard-Gate Candidacy in a Deployed Validator Suite

شين سو · 2026

Before a validator can be promoted to a hard gate on a deployment pipeline, it has to be shown that its firing separates outputs that reach users in working order from those that do not. We run that screen on 13 validators in a deployed generative agent, against 550 runtime and 350 static builds labelled by downstream …

نسخة أولية وصول مفتوح

Learning to Steer, Steering to See: Unveiling the Geometry of RLVR in Large Language Models via Trainable Vectors

Reinforcement learning (RL) has become a key paradigm for enhancing the reasoning of large language models, yet the high dimensionality of parameter updates makes its training dynamics hard to analyze. We study reinforcement learning with verifiable rewards (RLVR) and use vector steering to identify a low-dimensional e …

المؤلفون المشاركون