Skip to yearly menu bar Skip to main content


Poster Thu, Jul 9, 2026 • 2:30 PM – 4:15 PM KST HALL A #205

Expected Return Causes Outcome-Level Mode Collapse in Reinforcement Learning and How to Fix It with Inverse Probability Scaling

Abhijeet Sinha ⋅ Sundari Elango ⋅ Dianbo Liu

Abstract

Lay Summary

Video

Chat is not available.