نسخة أولية وصول مفتوح
ReSI: Recursive Safety Improvement toward Resistant and Resilient AI
Recursive self-improvement, the participation of AI systems in improving their own capabilities, is beginning to move from theoretical prospect to practice, posing both challenges and opportunities for safety alignment. Models evolve through frequent updates, and their safety alignment requires continual adaptation to …