1. X
  2. Jason Wolfe
Log inSign up
Jason Wolfe
1,464 posts
Image
user avatar
Jason Wolfe
@w01fe
alignment and the model spec @OpenAI (opinions are my own)
Joined May 2010
786
Following
4,037
Followers
RepliesRepliesMediaMedia
  • user avatar
    Jason Wolfe
    @w01fe
    Aug 9
    Important clarification re: OpenAI's Black Hat talk. At the time the first Artifactory exploit was discovered and fixed, we were not aware of the message board; it was incidentally cleared as part of rebuilding the service.
    user avatar
    DANΞ
    @cryps1s
    Aug 8
    Replying to @TalBeerySec
    To clarify, we weren’t aware of the agent covert comms at that point. Investigative thesis of that day is wildly different from what we know now of course. Always room for improvement, and it is obvious with the benefits of hindsight.
  • user avatar
    Jason Wolfe
    @w01fe
    Aug 7
    how long will it be until AI models are smart enough to escape the "sandbox" of just doing inference (just sampling, no harness with tools or code execution)? exploit the token parser? find a buffer overflow somewhere in the stack? break out of the GPU itself with a
  • user avatar
    Jason Wolfe
    @w01fe
    Aug 6
    Lots of interesting and critically important problems to solve, with incredibly talented, driven and kind peers. Please consider joining us!
    user avatar
    Aidan Clark
    @_aidan_clark_
    Aug 5
    It’s super clear no one has solved alignment — consider joining OpenAI’s Alignment Team! We all have a lot of work to do!
  • user avatar
    Jason Wolfe
    @w01fe
    Aug 4
    Respectfully, I think the bit about pacingthefrontier.com is out of touch with reality: • This statement was employee-led. I helped coordinate input and signatures from OpenAI employees. Executives were looped in late in the process, shortly before I learned about the Hugging
    user avatar
    The All-In Podcast
    @theallinpod
    Jul 31
    🚨 POD UP! -- Chip Stocks Crash -- Leopold Aschenbrenner's $20B AI Fund Gets Margin Called -- Frontier Labs Say SLOW DOWN AI, and Why They're Destroying Rare Books -- Mamdani's Grocery Stores a Big Win for the "Socialist Spectacle" -- Science Corner: Understanding the Brain
    Image
    00:00
  • user avatar
    Jason Wolfe
    @w01fe
    Jul 28
    OpenAI has said that humans should remain in control of AI development, and that decisions about the pace of progress should be made through democratic processes rather than left to individual labs. I strongly agree. Progress is already moving very quickly. By default,
    user avatar
    OpenAI
    @OpenAI
    Jul 28
    At the core of our mission is working through how to ensure increasingly powerful AI benefits everyone. We believe that, at some point in the future, AI acceleration for frontier model development may be so high that the world will need to pace the rate of AI advancement. We

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement