نسخة أولية وصول مفتوح
RubricArmor: Adversarial Evolution Improves LLM-Based Rubric Generation
Rubric-based reinforcement learning (RL) provides interpretable rewards for aligning large language models (LLMs) by evaluating responses against query-specific evaluation criteria. To construct rubrics at scale, a straightforward approach to LLM-based rubric generation is to prompt an LLM to generate a rubric directly …