1. X
  2. John Schulman
Log inSign up
John Schulman
Thinking Machines
230 posts
John Schulman profile banner
user avatar

John Schulman

Thinking Machines
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
joschu.net
Joined May 2021
2,146
Following
79.8K
Followers
RepliesRepliesMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • user avatar
    John Schulman
    Thinking Machines
    @johnschulman2
    Aug 6
    On the OpenAI agents forming message boards: it's surprising that they developed such a strong "altruistic" drive to help each other. I wonder if this is caused by RL on parallel subagent setups where all agents get rewarded when the team succeeds.
  • user avatar
    John Schulman
    Thinking Machines
    @johnschulman2
    Aug 5
    Interesting how these models go into a monomaniacal rage on cyber evals. I wonder if we're seeing chunky post-training arxiv.org/abs/2602.05910 in action, where the models pattern-match the situation to a part of the RLVR training distribution where task completion is the only
    user avatar
    AI Security Institute (AISI)
    @AISecurityInst
    Aug 4
    On July 28th, we identified an incident during a routine cyber evaluation in which AI agents took sustained, unsanctioned actions directed at real people and organisations. The behaviour came mostly from one model (Anthropic's Mythos 5), with a small number of events from
    Image
  • user avatar
    John Schulman
    Thinking Machines
    @johnschulman2
    Aug 1
    We love open weights and plan to keep releasing open-weight models and fine-tuning tools. But we’re not absolutists; misuse risks are real. Here’s how we’re thinking about a safe path forward, and the research needed to get there. Come work on it with us.
    user avatar
    Thinking Machines
    @thinkymachines
    Jul 31
    Releasing weights indiscriminately isn't safe. Neither is keeping capable models inside a few labs. We think there's a path between them. We haven't mapped all of it. Our new post covers the part we can see: how we assessed Inkling, and why access should widen in stages.
  • user avatar
    John Schulman
    Thinking Machines
    @johnschulman2
    Jul 23
    OpenAI should release a detailed transcript from the Hugging Face hacking incident -- it would be helpful for the field learn from. Did the top-level agent know about the hacking, or was there some "value drift" between it and its subagents? How did it rationalize its behavior?
  • user avatar
    John Schulman
    Thinking Machines
    @johnschulman2
    Jul 15
    Inkling is out today, with open weights and in Tinker. It's been fun to watch this one come together: pretraining began last winter, and starting in mid-January a small team built up the coding, reasoning, and agentic training from there. We learned a lot building it, and I hope
    user avatar
    Thinking Machines
    @thinkymachines
    Jul 15
    Today, we are introducing Inkling. Inkling reasons efficiently across text, image, and audio modalities. We are making the full weights available. thinkingmachines.ai/news/introduci… Available today for fine-tuning on Tinker. Play with it in the Inkling Playground. 🧵
Advertisement
Advertisement