Skip to yearly menu bar Skip to main content


Poster Thu, Dec 10, 2026 • 8:30 AM – 11:30 AM AEDT Hall C1

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages

Vishnu Teja Kunde ⋅ Fatemeh Doudi ⋅ Mahdi Farahbakhsh ⋅ Dileep Kalathil ⋅ Krishna Narayanan ⋅ JF Chamberland

Abstract

Chat is not available.