Log inSign up
Tianyu (Steve) Wang
249 posts
Tianyu (Steve) Wang profile banner
@VisionSteve

Tianyu (Steve) Wang

@VisionSteve
Research Scientist @AdobeResearch | Ph.D. @ CUHK | Prev. Research Intern @AdobeResearch | Photographer | Opinions are my own
California, USA
stevewongv.github.io
Joined November 2019
513
Following
755
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @VisionSteve
    Tianyu (Steve) Wang
    @VisionSteve
    Jul 31
    Excited to share Chimera! As Kimi K3 brings hybrid linear attention into the spotlight, our work explores the same architectural transition from the visual side: adapting KDA-based hybrid attention to token-extensive visual generation, together with a systematic scaling recipe.
    @hanwenjiang1
    Hanwen Jiang
    @hanwenjiang1
    Jul 31
    Interested in how frontier labs pre-train image/video generation models? We were too. Since those recipes are rarely made public in full, we started from the most mature pretraining playbook available in the open: how modern LLMs are built. Introducing Chimera: a visual
    Image
    Image
    Image
    Image
    1
  • @VisionSteve
    Tianyu (Steve) Wang
    @VisionSteve
    Jun 24
    Great work! But it would be great to cite our 2D-Box ObjectMover paper in CVPR 2025.
    @anand_bhattad
    Anand Bhattad
    @anand_bhattad
    Jun 23
    Here is a short video of our Thinking in Boxes. Enjoy! More results on our website: thinking-in-boxes.github.io
    Image
    00:00
    1
  • @VisionSteve
    Tianyu (Steve) Wang
    @VisionSteve
    Apr 25
    If you’re attending, please stop by our oral session for EditVerse: I won’t be there in person, but I’d be very happy if you could check out the work, chat with the team, and join the discussion. 📍 Room 201 A/B 🕚 Apr 25, 11:18–11:28 AM
    @juxuan_27
    Xuan Ju
    @juxuan_27
    Feb 6
    Excited to share our paper EditedVerse is accepted as oral to ICLR 2026! Many thanks to our amazing coauthors!! Paper Link: arxiv.org/pdf/2509.20360 Project Page: …se.s3-website-us-east-1.amazonaws.com
    Image
  • @VisionSteve
    Tianyu (Steve) Wang
    @VisionSteve
    Mar 17
    Check out our project that answers how to train any-step T2I model from scratch. We release the code for everyone to explore this area. Looking forward to seeing more on this!! #CVPR #CVPR2026
    @andy_yx27
    Xin Yu (Andy)
    @andy_yx27
    Mar 17
    😀😀We’re excited to release the training code for Self-E (accepted to CVPR 2026): github.com/XinYu-Andy/Sel… Self-E is a training-from-scratch, self-contained framework for any-step text-to-image generation, without teacher distillation.
  • @VisionSteve
    Tianyu (Steve) Wang
    @VisionSteve
    Nov 13, 2025
    The last version of FSD V13 is good! Confident and smooth. But the v14.1.4 is bad, not confident and not smooth. I put my hand back to the wheel again.😅 Cancel subscription again and wait for a stable version.
    @karpathy
    Andrej Karpathy
    @karpathy
    Nov 12, 2025
    I took delivery of a beautiful new shiny HW4 Tesla Model X today, so I immediately took it out for an FSD test drive, a bit like I used to do almost daily for 5 years. Basically... I'm amazed - it drives really, really well, smooth, confident, noticeably better than what I'm used
    1
Advertisement
Advertisement