1. X
  2. Brian Chao
Log inSign up
Brian Chao
141 posts
user avatar

Brian Chao

@BrianCChao
Ph.D. student @Stanford · @NSF Graduate Fellow
bchao1.github.io
Joined August 2021
432
Following
667
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Brian Chao
    @BrianCChao
    Jun 11
    Aside from the official code release, I am thrilled to share that Spectral Progressive Diffusion a.k.a. SPEED (arxiv.org/abs/2605.18736) is now integrated into @lmsysorg's SGLang (@sgl_project)! 🚀 Instead of always running diffusion at full resolution, SPEED progressively grows
    Image
    Image
    00:30
    user avatar
    Howard Xiao
    @howard_xhc
    Jun 11
    Today we release the code and a demo for our recent Spectral Progressive Diffusion paper🎉 Play around with it anytime! Just as what we have been doing also, we hope that it encourages the integration of our plug-and-play framework into latest and greatest image and video
  • user avatar
    Brian Chao
    @BrianCChao
    Jun 9
    Excited to release the code and model weights for our recent paper "Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation"! We hope this inspires continued research on mixed-resolution diffusion and efficient token allocation for image and video generation.
    Image
    00:00
  • user avatar
    Brian Chao
    @BrianCChao
    Jun 5
    I integrated our recent work, Spectral Progressive Diffusion a.k.a. SPEED (arxiv.org/abs/2605.18736), with the open-source @ideogram_ai model released yesterday. Turns out our method works out of the box and speeds up inference by up to 1.6× while preserving the high image
    Image
  • user avatar
    Brian Chao
    @BrianCChao
    Jun 3
    Very cool model by @reve. The details and edibility are mind-blowing! From the researchers' posts, I can gather the following: 1. Pixel-space diffusion: they explicitly said "no latent auto-encoders". This is the main reason behind their text rendering quality and
    user avatar
    Reve
    @reve
    Jun 3
    Today, we’re launching Reve 2.0, the best 4K image model in the world. We invented a new way to generate and edit any image using precise layouts. For the first time, it’s possible to create images you can touch.
    Image
    00:00
  • user avatar
    Brian Chao
    @BrianCChao
    Jun 2
    Active vision has strong implications for robust and efficient autonomous systems. There is a lot of research in VLMs on guiding models to know where to look, and Policy-based Foveated Imaging is the physical realization of exactly that! Super cool work by @howard_xhc!
    user avatar
    Gordon Wetzstein
    @GordonWetzstein
    Jun 2
    The era of ultra-high-resolution imaging has arrived. Modern image sensors exceeding 200 MP resolution are common in smartphones, with over 400 MP sensors under development. However, the large number of pixels poses significant challenges for acquisition and processing,
    Image
    00:00

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement