1. X
  2. Julian Minder
Log inSign up
Julian Minder
323 posts
user avatar

Julian Minder

@jkminder
Anthropic Research Fellow (contract) – PhD at EPFL with Robert West and Ryan Cotterell – MATS 7 with Neel Nanda
Lausanne/Zürich
jkminder.ch
Joined November 2011
611
Following
876
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Julian Minder
    @jkminder
    Aug 14
    Synthetic Persona Pretraining: Alignment From Token Zero – our full paper is finally out. We train up to 3B models and inject synthetic morally-laden reflections into 10% of pretraining documents. Surprisingly, intervening early really shifts the model's value priorities 🧵
    Image
  • user avatar
    Julian Minder
    @jkminder
    2h
    agreed:)
    user avatar
    Alex Turner
    @Turn_Trout
    4h
    Synthetic persona pretraining / alignment pretraining seems promising (see e.g. turntrout.com/self-fulfillin…). Hope labs adopt these techniques. Yes pretraining changes are tougher but may be worth it. modelraising.ai/spp/
    Image
  • user avatar
    Julian Minder
    @jkminder
    Aug 14
    please do;) we are both limited by people and compute (studying safety pretraining and interactions with RL at scale is v expensive)!
    user avatar
    Nathan Helm-Burger
    @nathan84686947
    Aug 14
    Replying to @nathan84686947
    Dear AI safety funders, Please support this! x.com/nathan84686947…
  • user avatar
    Julian Minder
    @jkminder
    Aug 12
    I think understanding the effects of safety pretraining is crucial for things going well, come and help us!
    user avatar
    Bob West
    @cervisiarius
    Aug 12
    We're extremely grateful to @coeff_giving for supporting our work on "model raising" (LLM alignment from pretraining token 0) w/ $1M 🙏, which will fund 3 AI safety engineers for #Apertus. Job posting 👇 We're already screening apps, and you can still apply till Fri 8/14 EOD
  • user avatar
    Julian Minder
    @jkminder
    Aug 12
    Going to represent the view that safety must be considered already in pretraining at AI+X in Zürich this October!
    user avatar
    Valentina Pyatkin
    @valentina__py
    Aug 11
    📢Staging x Misalignment Science: When, Where and How Safety Enters LLM Training @anna_hedstroem and I are co-organizing a workshop at this year's AI+X Summit in Zurich to discuss where in the LLM training pipeline safety is actually decided, and with which methods? This session
    Image

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement