Happy to introduce our new paper "Diversity-Rewarded CFG Distillation".
We combine distillation, a novel diversity reward, and model merging to improve the quality-diversity tradeoff of MusicLM.
arxiv: arxiv.org/abs/2410.06084
More info:
An AI will win a Nobel price someday✨. Yet currently, alignment reduces creativity. Our new @GoogleDeepMind paper "diversity-rewarded CFG distillation" improves quality AND diversity for music, via distillation of test-time compute, RL with a diversity reward, and model merging.



