Skip to yearly menu bar Skip to main content


Poster
in
Workshop: 2nd Workshop on Compositional Learning: Safety, Interpretability, and Agents
Sat, Jul 11, 2026 • 4:00 PM – 5:00 PM KST

Introspective Coupling: LMs Explain Themselves Better Than Training Targets

Carl Guo ⋅ Laura Ruis ⋅ Jacob Andreas ⋅ Belinda Li

Abstract

Chat is not available.