Log inSign up
Ellis Brown
461 posts
@_ellisbrown

Ellis Brown

@_ellisbrown
PhD Student @nyu_courant w/ Profs @sainingxie and @rob_fergus, AIM student @meta FAIR. Prev @ai2prior, @carnegiemellon, @vanderbiltu
NYC
ellisbrown.github.io
Joined January 2016
646
Following
1,016
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @_ellisbrown
    Ellis Brown
    @_ellisbrown
    Jun 4
    introducing PaintBench 🎨 a deterministic eval protocol + benchmark for precise visual editing unlike most generative model evals, no judge model in the loop see @itskaixu's 🧵⬇️ for technical details. I'll provide more context / commentary about why it's exciting here
    Image
    5
  • @_ellisbrown
    Ellis Brown
    @_ellisbrown
    Mar 10
    when I was choosing PhD programs, @sainingxie was my main draw to NYU. he’d led such influential work in representation learning at FAIR and then swam upstream, choosing academia over big labs. I could tell he had a rare sense for what to work on in the new ChatGPT era doing my
    @sainingxie
    Saining Xie
    AMI Labs
    @sainingxie
    Mar 10
    i’m joining forces with @ylecun and an incredible group of people to start AMI Labs @amilabs. AMI isn’t a conventional lab. we don’t intend to become one. a lot to say about why this moment matters, but for now we’re heads down building. join us: amilabs.xyz
    2
  • @_ellisbrown
    Ellis Brown
    @_ellisbrown
    Mar 5
    🤝
    @askalphaxiv
    alphaXiv
    @askalphaxiv
    Mar 5
    Yann LeCun 🤝 Saining Xie insane crossover of the 2 biggest visual representation researchers in the AI field “Beyond Language Modeling: An Exploration of Multimodal Pretraining” Right now, most multimodal models are basically a language model with a vision adapter bolted on,
    Image
  • @_ellisbrown
    Ellis Brown
    @_ellisbrown
    Nov 10, 2025
    🌶️ hot take 🌶️ > we should normalize training on the test set yes, you read that right. no, I'm not joking. and, yes... I have taken ML 101 👉 here's why this is crucial for future multimodal LLM research [1/n] 🧵
    8
  • @_ellisbrown
    Ellis Brown
    @_ellisbrown
    Nov 7, 2025
    MLLMs are great at understanding videos, but struggle with spatial reasoning—like estimating distances or tracking objects across time. the bottleneck? getting precise 3D spatial annotations on real videos is expensive and error-prone. introducing SIMS-V 🤖 [1/n]
    Image
    00:00
    3
Advertisement
Advertisement