Pinned
๐Thrilled to share that RLMF is headed to #COLM2026! In this work we introduce ๐ฅ๐ ๐๐ถ๐๐ต ๐บ๐ฒ๐๐ฎ๐ฐ๐ผ๐ด๐ป๐ถ๐๐ถ๐๐ฒ ๐ณ๐ฒ๐ฒ๐ฑ๐ฏ๐ฎ๐ฐ๐ธ (๐ฅ๐๐ ๐) & metacognitive data selection. It outperforms standard RL while improving metacognitive monitoring in LLMs ๐ง โจ
Details below!๐
๐ฅ New paper alert!! Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs ๐ฅ
Can a model that learns to ๐ซ๐ถ๐ฅ๐จ๐ฆ ๐ช๐ต๐ด๐ฆ๐ญ๐ง get even better at tasksโand learn to be honest about what it doesn't know? Turns out: yes๐คฏ
Details๐งต๐

