It’s great to see so much excitement around self-distillation!
However, some recent work highlights an important failure mode: when the teacher is given the solution, it can suppress verification, backtracking, and exploration, which make reasoning models effective.
We
Excited to be heading to ICML! I will be presenting our paper on rethinking how to train models with expert solutions using self-distillation.
🗓️ July 8, 2026, 5:00 PM – 6:45, Hall A #2502



