نسخة أولية وصول مفتوح
Safe at One Loop, Risky at Another: Aligning Safety Across Recurrent Depths in Looped Language Models
Looped Language Models (LoopLMs) provide a parameter efficient approach to scaling model capabilities through repeated use of shared parameters across recurrent steps. Since each recurrent depth can be read out independently, a single LoopLM exposes a broader output space across inference depths, raising an important q …