Skip to yearly menu bar Skip to main content


Poster
in
Workshop: Trustworthy AI for Good Workshop

Does Moral Reasoning Training Help or Hurt? Red-Teaming RL-Trained Ethical Agents with Persona Attacks

Arth Ashish Singh

Abstract

Chat is not available.