Authors

Kangjun Noh

Publications 2

Preprint Open access

Reliable Self-Evolution with Imperfect Proxy Rewards

Large language model (LLM)-based self-evolving search is a promising approach to scientific discovery. However, high-fidelity evaluation of every candidate is prohibitively expensive in some domains. Self-evolving systems in such settings therefore rely on low-cost but imperfect proxy rewards, which may assign high sco …

Co-authors