1. X
  2. Yonghoon Dong
Log inSign up
Yonghoon Dong
32 posts
user avatar

Yonghoon Dong

@yhoon96
Ph.D student @KAIST_AI | Reinforcement Learning, Robotics
Seoul, Republic of Korea
yonghdong.github.io
Joined November 2025
326
Following
103
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Yonghoon Dong
    @yhoon96
    May 27
    Introducing TRQAM! Internalizing a KL trust region inside the sampling SDE stabilizes off-policy RL fine-tuning of pretrained flow policies. With TRQAM, we lift offline RL success on 50 OGBench tasks from 46% to 68%. 馃У [1/8] yonghdong.github.io/blog/trqam/
    Image

Log in or sign up for X

See what鈥檚 happening and join the conversation

Continue with phone
or
Log in with username or email
Terms路Privacy路Cookies路Accessibility路Ads Info路漏 2026 X Corp.
Advertisement
Advertisement