Log inSign up
Hadi Alzayer
248 posts
@HadiZayer

Hadi Alzayer

@HadiZayer
Research Scientist @GoogleDeepMind. Previously @Stanford @umdcs @Cornell
hadizayer.github.io
Joined February 2012
276
Following
559
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @HadiZayer
    Hadi Alzayer
    @HadiZayer
    Jul 24
    We trained a video world model on just 15 hours of video of a single-arm robot. 🧵 It generalizes zero-shot to unseen embodiments (and even orangutans). And it picked up something we never trained for: give it the object motion you want, and it synthesizes the robot motion that
    Image
    00:00
    13
  • @HadiZayer
    Hadi Alzayer
    @HadiZayer
    Aug 18
    Life update: Happy to share that I've wrapped up my PhD! Special thanks to @jbhuang0604 and @jiajunwu_cs for advising me throughout this journey. The PhD has been a blast with many incredible memories. Up next, starting as a research scientist at @GoogleDeepMind!
    Image
    Image
    31
  • @HadiZayer
    Hadi Alzayer
    @HadiZayer
    Jun 3
    I'll be presenting Coupled Diffusion at #CVPR2026!! If you are around in Denver, I would be happy to connect and chat!
    @HadiZayer
    Hadi Alzayer
    @HadiZayer
    Dec 4, 2025
    Our new work, coupled diffusion sampling, allows fast and diverse multi-view editing! Not by training — but by having one diffusion model guide another. 🤝 As a bonus: it can construct video-editing datasets and can make video models generate longer videos. 🧵👇
    Image
    00:00
    1
  • @HadiZayer
    Hadi Alzayer
    @HadiZayer
    May 31
    excited to finally see a standardized text-to-image benchmark!! this has been due for too long and seriously bottlenecking scientific analysis on scaling and generalization beyond basic class image conditioning. Most importantly, no more measuring FID on the training set!!! (I
    @KyleSargentAI
    Kyle Sargent
    @KyleSargentAI
    May 29
    Today we released “GPIC: A Giant Permissive Image Corpus for Visual Generation.” It’s a 100M image dataset for visual generation, with text captions and 100% known+permissive licenses, hosted on HuggingFace. I’m excited to get this out! Check it out: gpic.stanford.edu
    Image
  • @HadiZayer
    Hadi Alzayer
    @HadiZayer
    Mar 8
    we had so much fun playing together!! trying to find glitches in the game play reminded me of some of the good old days with video games
    @Po_lhr
    Ryan Po
    @Po_lhr
    Mar 6
    🎮 Real-time multiplayer world model 👥 Arbitrary number of players 🧠 Generated entirely by a neural network MultiGen is a real-time multiplayer diffusion game engine that supports an arbitrary number of players through a shared memory-based world model, rather than limiting
    Image
    00:00
Advertisement
Advertisement