Skip to yearly menu bar Skip to main content


Poster Thu, Dec 10, 2026 • 5:00 PM – 8:00 PM AEDT Hall 1-4

Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor’s Internal States

Yunho Choi ⋅ Jongwon Lim ⋅ woojin Ahn ⋅ Minjae Oh ⋅ Jeonghoon Shim ⋅ Yohan Jo

Abstract

Chat is not available.