Preprint Open access
UBA-ORL: Unlearning-Activated Backdoor Attacks on Offline Reinforcement Learning
Offline reinforcement learning (offline RL) enables policy learning from pre-collected static datasets without online exploration, and is increasingly deployed not only in safety-critical domains such as autonomous driving and robotic control but also in data-mining applications such as recommendation and behavior anal …