Skip to yearly menu bar Skip to main content


Poster Thu, Dec 10, 2026 • 2:00 AM – 5:00 AM AEDT Hall C1

Stop Calling It Reinforcement Learning in Language Models Without Clear Improvement Claims: Decision-Process Cards as a Reporting Standard

Anvi Kohli ⋅ Vedant Khandelwal ⋅ Rishit Agarwal ⋅ Rahul Maity ⋅ Amit Sheth

Abstract

Chat is not available.