Skip to yearly menu bar Skip to main content


Poster Wed, Jul 8, 2026 • 10:30 AM – 12:15 PM KST HALL A #4504

Provable Benefits of RLVR over SFT for Reasoning Models: Learning to Backtrack Efficiently

Stanley Wei ⋅ Juno Kim

Abstract

Lay Summary

Video

Chat is not available.