1. X
  2. Chenghao Yang
Log inSign up
Chenghao Yang
238 posts
user avatar
Chenghao Yang
@chrome1996
Senior AS @Microsoft. Ph.D. @UChicago Ex-SR @google Ex-Scientist @AWS. Ex-RA @jhuCLSP @columbianlp @TsinghuaNLP. Ex-Intern @IBM @AWS. Opinions are my own.
Redmond, WA
yangalan123.github.io
Born December 16, 1996
Joined March 2017
760
Following
2,218
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Chenghao Yang
    @chrome1996
    Jun 24, 2025
    Have you noticed… 🔍 Aligned LLM generations feel less diverse? 🎯 Base models are decoding-sensitive? 🤔 Generations get more predictable as they progress? 🌲 Tree search fails mid-generation (esp. for reasoning)? We trace these mysteries to LLM probability concentration, and
    Image
    00:00
  • user avatar
    Chenghao Yang
    @chrome1996
    Jun 15
    Excited to contribute to this work! 👁️📝 As someone relatively new to the multimodal world, I've always wanted to explore how we can better ground LLMs. But before diving deep into modeling, I believe we need a better understanding of what our benchmarks are actually evaluating.
    Image
    Image
    user avatar
    Harvey Yiyun Fu
    @harveyiyun
    Jun 15
    Seeing👀 is not reasoning🤔? Introducing our new blog post. We tested 8 open VLMs on 9 VQA benchmarks, and found that a lot of "visual" accuracy isn't visual at all. Many benchmarks can be solved without ever looking at the image, and some VLMs reason better from text than from
  • user avatar
    Chenghao Yang
    @chrome1996
    Apr 29
    What happens when we build an "AI society" with LLMs? Turns out, it's far more stereotyped than we thought. 🤖📉 Excited to share our new work on Persona Collapse—where distinct AI agents regress into narrow behavioral modes. Surprisingly, the models with the highest per-persona
    user avatar
    Lorenzo Xiao
    @lrzneedresearch
    Apr 28
    We have some concerns about the current state of LLM-based social simulation. We benchmarked 10 LLMs on persona simulation. Every model collapses. The "best" ones are the worst offenders. And RLHF actively makes it worse. arxiv.org/pdf/2604.24698
    Image
  • user avatar
    Chenghao Yang
    @chrome1996
    Mar 21
    BranchingFactor v1.1 just dropped! 🚀 (Yes — it’s an actively updated paper.) (arxiv.org/abs/2506.17871) As models rely more on post-training, understanding the synergy between pre-training and alignment becomes crucial. Branching Factor (BF) offers a simple way to track the
    Image
    00:00
  • user avatar
    Chenghao Yang
    @chrome1996
    Jan 30
    I will be doing my PhD defense today! Come and learn about my Grounded Alignment works! Detailed information (w/Zoom) below: Candidate: Chenghao Yang Date: Friday, January 30, 2026 Time:  2 pm CST Location: John Crerar Library 298 Zoom:  uchicago.zoom.us/j/96014992390?… Meeting ID:
    Image
    Image

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement