Preprint Open access
Verification Trap: Understanding Test-Time Selection Failures under False Premises in Code Generation
Test-time compute has become a central way to improve code generation: systems sample multiple candidate programs and use verifier-visible evidence to select the final output. This paradigm implicitly assumes that the verifier provides a corrective signal independent from the generator. We challenge this assumption und …