Skip to yearly menu bar Skip to main content


Making LLMs Say What They Think: Measuring and Improving CoT-Interpretability Alignment

Yihuai Hong ⋅ Shauli Ravfogel ⋅ Chen Zhao ⋅ Eunsol Choi

Abstract

Chat is not available.