Preprint Open access
Learn Now, Use Next, Trust Later: Prequential Test-Time Learning for LLM Agents
Adapting large language model agents during deployment requires not only retaining past experience, but also turning new observations into timely guidance. Many test-time learning methods, however, acquire knowledge from completed episodes. Feedback from an ongoing interaction may therefore not be distilled into knowle …