Skip to yearly menu bar Skip to main content


Poster
in
Workshop: Trustworthy AI for Good Workshop

Beyond the Prompt: Leveraging Pre-Decoding States for Jailbreak Detection in dLLMs

Adam Hazimeh ⋅ Amel Abdelraheem ⋅ Ke Wang ⋅ Mariam Salman ⋅ Ljiljana Dolamic ⋅ Gérôme Bovet ⋅ Pascal Frossard

Abstract

Chat is not available.