1. X
  2. David Fan
Log inSign up
David Fan
AMI Labs
186 posts
user avatar
David Fan
AMI Labs
@DavidJFan
🇰🇷 ICML 2026! AMI Labs | ex-Meta FAIR | @Princeton CS '19 Building the next revolution of AI models that understand the real world.
New York City
scholar.google.com/citations?user…
Joined June 2013
503
Following
1,788
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Mar 4
    [1/9] What happens when you treat vision as a first-class citizen during multimodal pretraining? To find out, we studied the design space of training Transfusion-style models that input and output all modalities, from scratch. Here is what we learned about visual representations,
    arXiv logo
    arxiv.org
    Beyond Language Modeling: An Exploration of Multimodal Pretraining
    The visual world offers a critical axis for advancing foundation models beyond language. Despite growing interest in this direction, the design space for native multimodal models remains opaque....
    76K
  • user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Jul 26
    Nice work!! Really glad to hear WebSSL works well for robot policies 😁 Representation matters
    user avatar
    Jeff Cui
    @jeffacce
    Jul 22
    Your policy doesn't need 7B params. It simply needs dense features. Introducing Patch Policy: pretrained ViT + small transformer beats OpenVLA-OFT with 0.7% of its params, and trains on a 5090. Here it inserts a cable (~2mm tol), and does it again as we unplug mid-rollout. 🧵
    Image
    00:00
    4.7K
  • user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Jul 7
    @__JohnNguyen__ and I are presenting the Beyond Language Modeling paper as an ICML spotlight in 30 minutes at the 10:30 AM poster session! I’ll also be at the AMI Mixer on Thursday and hanging around in Seoul until early next week :)
    user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Mar 4
    [1/9] What happens when you treat vision as a first-class citizen during multimodal pretraining? To find out, we studied the design space of training Transfusion-style models that input and output all modalities, from scratch. Here is what we learned about visual representations,
    25K
  • user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Apr 16
    Congrats Dr. @TongPetersb!! Your research journey has clearly culminated in a very cohesive and inspiring narrative that unifies several areas of work with much scope for future expansion :D It's a testament to your work ethic, good taste in problems, and attention to detail. I'm
    user avatar
    John Nguyen
    AMI Labs
    @__JohnNguyen__
    Apr 16
    Congrats Dr. Tong! Really glad to be a part of your PhD journey @TongPetersb
    Image
    Image
    Image
    00:00
    12K
  • user avatar
    David Fan
    AMI Labs
    @DavidJFan
    Apr 1
    Check out the code and data for DexWM! Training + Inference Code: github.com/facebookresear… Robocasa Data: huggingface.co/datasets/faceb…
    user avatar
    Raktim Gautam Goswami
    @raktimgg
    Apr 1
    The code for DexWM is now publicly available: github.com/facebookresear…. The repository includes the full training and evaluation pipelines, along with custom dexterous manipulation datasets generated in RoboCasa, making it easy to reproduce our results and build on top of this
    3.5K
  • See @DavidJFan's full profile

    Sign up
    Log in

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement