Pinned
🎉Thrilled to share that RLMF is headed to #COLM2026! In this work we introduce 𝗥𝗟 𝘄𝗶𝘁𝗵 𝗺𝗲𝘁𝗮𝗰𝗼𝗴𝗻𝗶𝘁𝗶𝘃𝗲 𝗳𝗲𝗲𝗱𝗯𝗮𝗰𝗸 (𝗥𝗟𝗠𝗙) & metacognitive data selection. It outperforms standard RL while improving metacognitive monitoring in LLMs 🧠✨
Details below!👇
🔥 New paper alert!! Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs 🔥
Can a model that learns to 𝘫𝘶𝘥𝘨𝘦 𝘪𝘵𝘴𝘦𝘭𝘧 get even better at tasks—and learn to be honest about what it doesn't know? Turns out: yes🤯
Details🧵👇

