📢Staging x Misalignment Science: When, Where and How Safety Enters LLM Training
@anna_hedstroem and I are co-organizing a workshop at this year's AI+X Summit in Zurich to discuss where in the LLM training pipeline safety is actually decided, and with which methods? This session
- ‼️ Hiring AI Safety Researchers ‼️ I am moving to Berlin to start a new research group on AI safety and societal impacts @HPI_DE. From September this year, I will be hiring fully funded PhDs, postdocs, and research interns. My group will study the risks and benefits that emerge
- 🚨new release: Apertus 1.5, a continued pretraining of the Apertus 1.0 models, with multimodal understanding (text, image, speech), an optional thinking mode, a four times longer context window, and better instruction following and tool use!Hot open-source summer season has gotten to Switzerland 🇨🇭⛰️ Our team at Swiss AI Initiative is happy to release Apertus 1.5 - a multimodal and reasoning update to our first v1 version! The new version builds a foundation for future regular open model development at scale🧵
- It was a pleasure working with you and many congrats on getting the rlhf book out! 💪Replying to @natolambertThe people who did a mix of directly helping me on the book, inspiring me, and or helping the book without knowing it! Luca Soldaini: @soldni Kyle Lo: @kylelostat Dirk Groeneveld: @mechanicaldirk Hanna Hajishirzi: @HannaHajishirzi Saumya Malik: @saumyamalik44 Costa Huang:





