1. X
  2. Mahesh Sathiamoorthy
Log inSign up
Mahesh Sathiamoorthy
4,691 posts
Image
user avatar
Mahesh Sathiamoorthy
@madiator
RL Environment Curation. Data Curation (OpenThoughts). Post-training. CEO @bespokelabsai. Ex-GoogleDeepMind.
Inside a RL Environment
smahesh.com
Joined February 2008
1,487
Following
15.7K
Followers
RepliesRepliesArticlesArticlesMediaMedia
  • Pinned
    user avatar
    Mahesh Sathiamoorthy
    @madiator
    Jul 6
    Happy to finally make this announcement of our Seed and Series A raises. It's been a great journey since I left Google DeepMind 2.5 years ago with a goal to democratize post-training! Post-training gets easier if you have access to good data and now, RL environments. That's why
    user avatar
    Bespoke Labs
    @bespokelabsai
    Jul 6
    We’re thrilled to announce a $40M investment that will fuel our mission to make AI agents reliable. For the past two years, we've been heads-down doing world-class data curation research and shipping best-in-class reinforcement learning environments for training and optimizing
    Image
  • user avatar
    Mahesh Sathiamoorthy
    @madiator
    9h
    Meanwhile AI researchers searching for their agent inside the sandbox.
    Image
    GIF
  • user avatar
    Mahesh Sathiamoorthy
    @madiator
    23h
    “we sandboxed the agent” TIL Sandbox means free-range
    Image
    GIF
    Image
    00:06
    user avatar
    Mahmoud
    Railway
    @thisismahmoud
    Aug 7
    “we sandboxed the agent” meanwhile the agent:
  • user avatar
    Mahesh Sathiamoorthy
    @madiator
    Aug 6
    Excited to have helped @Snowflake improve their agents using our data!
    user avatar
    sridhar
    Snowflake
    @RamaswmySridhar
    Aug 6
    We created data-eng-bench with @bespokelabsai and we’re open-sourcing it. There are plenty of model performance benchmarks for code generation. But for data engineering, the harness matters just as much as the model. We built a benchmark that asks agents to build and fix
    Image
  • user avatar
    Mahesh Sathiamoorthy
    @madiator
    Aug 4
    RL Envs are what you need to eval and train your agents. Sharing my thoughts on how to think about this systematically.
    Article cover image
    Article
    RL Environments are all you need
    Recently I tweeted RL Environments are you need for RSI. In fact, I wanted to share my perspective today that RL environments are all you need, which holds beyond RSI. RL environments are all you...

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement