1. X
  2. Alexandre Ramé
Log inSign up
Alexandre Ramé
908 posts
Image
user avatar
Alexandre Ramé
@ramealexandre
Senior research scientist @GoogleDeepMind. Previously PhD @Sorbonne_Univ_. Post-training Gemma LLMs: (self)distillation, RL and merging.
alexrame.github.io
Joined May 2011
791
Following
1,997
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Alexandre Ramé
    @ramealexandre
    Apr 2
    Gemma 💎💎💎💎 Fantastic models, fantastic team!
    user avatar
    Demis Hassabis
    @demishassabis
    Apr 2
    Excited to launch Gemma 4: the best open models in the world for their respective sizes. Available in 4 sizes that can be fine-tuned for your specific task: 31B dense for great raw performance, 26B MoE for low latency, and effective 2B & 4B for edge device use - happy building!
    Image
    1K
  • user avatar
    Alexandre Ramé
    @ramealexandre
    May 22, 2025
    Thanks for leading this post-training, with great engineering and research breakthroughs! Super achievement by the whole team.
    user avatar
    Robert Dadashi
    @robdadashi
    May 22, 2025
    Replying to @robdadashi
    Our post-training team (once again) went beyond to post-craft this fantastic performance-for-size model: @johanferret @ramealexandre @sarah_perrin_ @CdrGeo @nino_vieillard @sabelaraga @angelinepouget V Carbune L Rouillard L Hussenot @liu_gael @sinopalnikov @OlivierBachem 2/2
    2.1K
  • user avatar
    Alexandre Ramé
    @ramealexandre
    May 20, 2025
    Releasing Gemma 3n, our new open-weight model processing audio, images and text (with improved multilingual capabilities), optimized for on-device usage with MatFormer architecture (enabling adaptive compute) and reaching 1283 on Chatbot Arena. Read more: developers.googleblog.com/en/introducing….
    Image
    7.9K
  • user avatar
    Alexandre Ramé
    @ramealexandre
    May 5, 2025
    There are also a few interesting surprised updates in the Lmsys leaderboard 👀 Gemma-3-27B (1341) ~ Qwen3-235B-A22B (1342) Gemma-3-12B (1321) ~ DeepSeek-V3-685B-37B (1318) Gemma-3-4B (1272) ~ Llama-4-Maverick-17B-128E (1270)
    Image
    Image
    user avatar
    Arena.ai
    @arena
    May 5, 2025
    The community votes are in for Qwen3-235B-A22B 🥁 The latest open-source Qwen3 is now on the Arena Top 10 🏆 Congrats to @alibaba_qwen on this achievement! 👏 Highlights: 💠 For Chat: Qwen3-235B-A22B ranks #10, tied with o1 💠 Strong in Coding at #4 and Math #1 💠 For WebDev:
    11K
  • user avatar
    Alexandre Ramé
    @ramealexandre
    Mar 26, 2025
    Hiring two student researchers for Gemma post-training team at @GoogleDeepMind Paris! First topic is about diversity in RL for LLMs (merging, generalization, exploration & creativity), second is about distillation (with @nino_vieillard). Ideal if you're finishing PhD. DMs open!
    21K
  • See @ramealexandre's full profile

    Sign up
    Log in

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement