🚨 New preprint on creative chess puzzle generation with diffusion models is out! ♟️
In his exceptional MSc thesis, @AatuSelkee introduced a new RL recipe on top of a masked diffusion model improving the previous SOTA.
We're open-sourcing the model and releasing a demo 🚀
1/n
I had the privilege of sitting down with @kchonyc for our new AI podcast "And humanity created intelligence". we discussed the backstory behind the attention paper as well as... 🧵
if you're at #icml2025, come check out our spotlight poster on "Mastering Board Games by External and Internal Planning with Language Models" ♟️
📜: arxiv.org/abs/2412.12119
⏲️: Wed 16 Jul 11 am - 1:30 pm PDT
📍: East Exhibition Hall A-B #E-2508
demo: goo.gle/ChessChamp
thank you for the recognition @GaryMarcus!
there's room for improvement, but I find it quite remarkable that an LLM learns to play creative sacrifices like this (best move according to Stockfish)
@ericmalmi has kind of done that and it does pretty well except in weird positions - where it still sometimes make illegal moves.
Confirming your conjecture and mine, if I understand his results correctly.
arxiv.org/pdf/2412.12119…