view article Article Red Teaming with RL: Exploiting Tinker API for Harmful RL on 235B Model 19 days ago โข 16