user_image

Yidong Huang

I’m a Second Year CS Ph.D. student at UNC Chapel Hill, advised by Prof. Mohit Bansal. I obtained my master degree from University of Michigan advised by Prof. Joyce Chai. Before that, I got my bachelor degree from University of Michigan and Shanghai Jiao Tong University.

I specialize in Embodied Artificial Intelligence and the Generative AI that interact with humans and their environments. My current research goal is to create AI agents beyond mere perception and reactive generation, to develop rich representations of the world and human partners, enabling deliberative planning and collaboration with humans.

Beyond work, I enjoy all kinds of sports, building and playing games, watching animations, and connecting with people from diverse backgrounds. If we share any interests, feel free to reach out!

I’m always open to research collaborations, project ideas, or just a good conversation. Feel free to contact me via email!

EDUCATION

  • 2025-08-01 –

    The University of North Carolina at Chapel Hill

    Ph.D. in Computer Science

  • 2023-09-01 – 2025-04-31

    University of Michigan

    MS in Computer Science

  • 2021-09-01 – 2023-04-30

    University of Michigan

    B.S.E in Computer Science

  • 2019-09-01 – 2023-08-01

    Shanghai Jiao Tong Univeristy

    B.S.E in Electronic and Computer Engineering

  • 2016-09-01 – 2019-06-30

    No. 2 High School of East China Normal University

PUBLICATIONS

05/2026

PhyMotion, Structured 3D Motion Reward for Physics-Grounded Human Video Generation

Yidong Huang*, Zun Wang*, Han Lin, Dong-Ki Kim, Shayegan Omidshafiei, Jaehong Yoon, Jaemin Cho, Yue Zhang, Mohit Bansal

* Equal contribution

arXiv preprint, 2026

Image

11/2025

SketchVerify, Planning with Sketch-Guided Verificationfor Physics-Aware Video Generation

Yidong Huang, Zun Wang, Han Lin, Dong-Ki Kim, Shayegan Omidshafiei, Jaehong Yoon, Yue Zhang, Mohit Bansal

arXiv preprint, 2025

Image

06/2024

DriVLMe, Enhancing LLM-based Autonomous Driving Agents with Embodied and Social Experiences

Yidong Huang, Jacob Sansom, Ziqiao Ma, Felix Gervits, Joyce Chai

IROS 2024

Image

12/2023

Inversion-Free Image Editing with Natural Language

Sihan Xu*, Yidong Huang*, Jiayi Pan, Ziqiao Ma, Joyce Chai

* Equal contribution

CVPR 2024

Image

10/2023

CycleNet, Rethinking Cycle Consistency in Text-Guided Diffusion for Image Manipulation

Sihan Xu*, Ziqiao Ma*, Yidong Huang, Honglak Lee, Joyce Chai

* Equal contribution

NeurIPS 2023

The Double Wizard Setup

10/2022

DOROTHIE, Spoken Dialogue for Handling Unexpected Situations in Interactive Autonomous Driving Agents

Ziqiao Ma*, Benjamin VanDerPloeg*, Cristian-Paul Bara*, Yidong Huang*, Eui-In Kim, Felix Gervits, Matthew Marge, Joyce Chai

* Equal contribution

Findings of EMNLP 2022

Image

12/2021

A-ESRGAN, Training Real-World Blind Super-Resolution with Attention U-Net Discriminators

Zihao Wei*, Yidong Huang*, Yuang Chen, Chenhao Zheng, Jingnan Gao

* Equal contribution

PRICAI 2023

RECENT PUBLICATIONS

[All Publications]
  1. PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation
    Huang, Yidong, Wang, Zun, Lin, Han, Kim, Dong-Ki, Omidshafiei, Shayegan, Yoon, Jaehong, Cho, Jaemin, Zhang, Yue, and Bansal, Mohit
    arXiv preprint arXiv:2605.14269, 2026
  2. The Mechanistic Emergence of Symbol Grounding in Language Models
    Wu, Shuyu, Ma, Ziqiao, Luo, Xiaoxi,  Huang, Yidong, Torres-Fonseca, Josue, Shi, Freda, and Chai, Joyce
    arXiv preprint arXiv:2510.13796, 2025
  3. Planning with Sketch-Guided Verification for Physics-Aware Video Generation
    Huang, Yidong, Wang, Zun, Lin, Han, Kim, Dong-Ki, Omidshafiei, Shayegan, Yoon, Jaehong, Zhang, Yue, and Bansal, Mohit
    arXiv preprint arXiv:2511.17450, 2025