Log inSign up
Zichen Liu
595 posts
Zichen Liu profile banner
@zzlccc

Zichen Liu

@zzlccc
Gemini RL @GoogleDeepMind
Singapore
lkevinzc.github.io
Joined October 2021
459
Following
6,028
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @zzlccc
    Zichen Liu
    @zzlccc
    Mar 21, 2025
    🪂Understanding R1-Zero-Like Training: A Critical Perspective * DeepSeek-V3-Base already exhibits "Aha moment" before RL-tuning?? * The ever-increasing output length in RL-tuning might be due to a BIAS in GRPO?? * Getting GRPO Done Right, we achieve a 7B AIME sota! 🧵 📜Full
    Image
    29
  • @zzlccc
    Zichen Liu
    @zzlccc
    Aug 6
    Welcome submissions if you work on meta-agents!
    @MetaAgentWkshop
    Meta-Agents Workshop @ NeurIPS 2026
    @MetaAgentWkshop
    Aug 4
    🚀Call for Papers — @NeurIPSConf 2026 Workshop Workshop on Responsible Use of Meta-Agents 📅 December 11/12 · 📍 Sydney, Australia Join us to shape the future of Meta-Agents and Harness. Topics include but are not limited to: 1. Automated Design of Agent Harnesses 2.
    Image
  • @zzlccc
    Zichen Liu
    @zzlccc
    Apr 10
    rl intuition (up-scaled by the correctness of infra) is all you need when cooking with a strong base model such as gemini✨
    3
  • @zzlccc
    Zichen Liu
    @zzlccc
    Mar 17
    🦎🦎 Happy to see two of our works (DrGRPO & DPPO) are highlighted here! I don’t think changing a few terms is worth a new branding, so we respectfully kept predecessors’ name while highlighting the correction/improvement on top of them. Hopefully they inspire RL algo designs.
    @a_weers
    Alex Weers
    @a_weers
    Mar 15
    Finally finished! If you're interested in an overview of recent methods in reinforcement learning for reasoning LLMs, check out this blog post: aweers.de/blog/2026/rl-f… It summarizes ten methods, tries to highlight differences and trends, and has a collection of open problems
    Blog post on the current state of reinforcement learning for reasoning LLMs
    2
  • @zzlccc
    Zichen Liu
    @zzlccc
    Mar 11
    Huge congrats to Min!
    @mavenlin
    Min Lin
    AMI Labs
    @mavenlin
    Mar 10
    The most exciting breakthroughs in intelligence are yet to come. I’m super excited to start this journey with mes amis to make them happen together.
Advertisement
Advertisement