Fetching the paper…

Rewarding Doubt: A Reinforcement Learning Approach to Calibrated Confidence Expression of Large Language Models · Around