Pinned
🎉 Excited to share that RECAP has been accepted to #EMNLP2026 Main!
The key idea: Flawed thinking can actually help reasoning models learn better alignment. 🧠🛡️
We find that injecting just a small amount of flawed reasoning can reduce model safety by up to 36%. 📉
Instead of

