Log inSign up
Zhouhan Lin
46 posts
@zhouhan_lin

Zhouhan Lin

@zhouhan_lin
Associate Professor at Shanghai Jiao Tong University @sjtu1896. Formerly @ Facebook AI Research @MetaAI. Ph.D. @Mila_Quebec with Yoshua Bengio.
hantek.github.io
Joined May 2016
258
Following
210
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @zhouhan_lin
    Zhouhan Lin
    @zhouhan_lin
    Jul 29
    Top tech giants iterate LLMs at breakneck speed behind closed doors, only releasing final model weights without revealing the internal trade-offs, trial-and-error, or architectural decisions. Meanwhile, the traditional academic publishing cycle—from paper submission to peer
    Image
    GitHub - InternLM/archspace: Turning LLM architecture exploration into reusable knowledge for the...
    From github.com
  • @zhouhan_lin
    Zhouhan Lin
    @zhouhan_lin
    Apr 25
    Sick of LLMs over-optimizing for a single reasoning path? 📉 Meet FlowRL at #ICLR2026! We use flow balancing to capture the full reward distribution, ensuring more diverse and robust reasoning. 🌈It moves beyond the reward maximization objective of PPO/GRPO, and introduces a new
    Image
    Image
    Image
    Image
  • @zhouhan_lin
    Zhouhan Lin
    @zhouhan_lin
    Apr 24
    🚨 Is that "new" model actually yours? 🕵️‍♂️ Stop by our ICLR poster to see AWM (Accurate Weight-Matrix Fingerprint) in action. We’ve built a high-fidelity metric to verify LLM lineage—even after intensive post-training or structural changes. ✅ No more worrying about model
    Image
    Image
    Image
    1
  • @zhouhan_lin
    Zhouhan Lin
    @zhouhan_lin
    Apr 24
    🚨 Is that "new" model actually yours? 🕵️‍♂️ Stop by our ICLR poster to see AWM (Accurate Weight-Matrix Fingerprint) in action. We’ve built a high-fidelity metric to verify LLM lineage—even after intensive post-training or structural changes. ✅ No more worrying about model
    Image
    Image
    Image
  • @zhouhan_lin
    Zhouhan Lin
    @zhouhan_lin
    Apr 23
    Presenting our #ICLR2026 work on Apr. 23: "MLP Memory: A Retriever-Pretrained Memory for LLMs." We are trying to solve the Long-term Memory challenge by shifting from external document access to a differentiable parametric module.🧠✨ ✅ True Parametric Memory: An external neural
    Image
    Image
    Image
    Image
Advertisement
Advertisement