1. X
  2. Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
Log inSign up
Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
37.3K posts
Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ profile banner
user avatar

Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ

@KyleMorgenstein
Full of childlike wonder. Teaching robots manners. RL Lead @ Apptronik. UT Austin PhD candidate. Past: Boston Dynamics AI Institute, NASA JPL, MIT โ€˜20.
he/him
kylemorgenstein.com
Joined September 2018
5,276
Following
16.3K
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
    @KyleMorgenstein
    Apr 17, 2024
    when you argue with me about control theory this is who youโ€™re arguing with
    Image
  • user avatar
    Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
    @KyleMorgenstein
    Aug 13
    love this- RL with dense guidance isnโ€™t really RL at all. weโ€™re still pretty bad at efficient exploration, but works like this, RFCL (@Stone_Tao), OmniReset (@patrickhyin, @ty_westenbroek), and others are showing great progress forwards!
    user avatar
    Cornelius Braun
    @corbraun
    Aug 13
    ๐—š๐—ผ๐—ผ๐—ฑ ๐—บ๐—ฎ๐—ป๐—ถ๐—ฝ๐˜‚๐—น๐—ฎ๐˜๐—ถ๐—ผ๐—ป ๐—ฝ๐—ผ๐—น๐—ถ๐—ฐ๐—ถ๐—ฒ๐˜€ ๐—บ๐—ฎ๐˜† ๐˜€๐˜๐—ฎ๐—ฟ๐˜ ๐˜„๐—ถ๐˜๐—ต ๐—ฏ๐—ฒ๐˜๐˜๐—ฒ๐—ฟ ๐˜€๐˜๐—ฎ๐˜๐—ฒ๐˜€, ๐—ป๐—ผ๐˜ ๐—ฏ๐—ฒ๐˜๐˜๐—ฒ๐—ฟ ๐—ฟ๐—ฒ๐˜„๐—ฎ๐—ฟ๐—ฑ๐˜€. We sample diverse, physically feasible contact states as starts + goals for RL. The behaviors that emerge are surprisingly dynamic and
    Image
    00:00
  • user avatar
    Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
    @KyleMorgenstein
    Aug 12
    solar eclipse from Copenhagen (I do not have glasses)
    Image
  • user avatar
    Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
    @KyleMorgenstein
    Aug 11
    this is โ€œtraining an RL policy for inverse kinematicsโ€ all over again. itโ€™s just tau = -g(q). the trade off is data vs modeling the mass and inertias. I would argue modeling mass is easier (and inherently more robust), but still always cool seeing stuff work on hardware.
    user avatar
    Roberto
    @robertorobotics
    Aug 11
    btw this zero G gravity comp is learned. The arm learned its own weight lol. Couple minutes of data.
  • user avatar
    Kyle๐Ÿค–๐Ÿš€๐Ÿฆญ
    @KyleMorgenstein
    Aug 11
    > we just doubled the execution speed at test time and it worked lmao
    user avatar
    dave
    @davidhe137
    Aug 10
    Replying to @aurel_arnold
    Havenโ€™t compared to test-time RTC, itโ€™s a bit hard to both implement and debug๐Ÿฅฒ The policy here is actually trained on 30hz but i just bumped up the execution hz to 50 without issue. Since this is just 1 robot, inference just runs as fast as it can (on Armory!) and the prefix

Log in or sign up for X

See whatโ€™s happening and join the conversation

Continue with phone
or
Log in with username or email
TermsยทPrivacyยทCookiesยทAccessibilityยทAds Infoยทยฉ 2026 X Corp.
Advertisement
Advertisement