1. X
  2. Ludwig Schmidt
Log inSign up
Ludwig Schmidt
250 posts
user avatar
Ludwig Schmidt
@lschmidt3
Assistant professor at @Stanford and member of the technical staff at @AnthropicAI.
Palo Alto, CA
people.csail.mit.edu/ludwigs/
Joined August 2009
426
Following
6,597
Followers
RepliesRepliesMediaMedia
  • user avatar
    Ludwig Schmidt
    @lschmidt3
    Jun 25
    Very excited to release the next project in the DataComp / OpenThoughts line of research! Like OpenThoughts we worked on post-training data, this time with a focus on agentic models.
    user avatar
    Richard Zhuang
    @RichardZ412
    Jun 24
    How can we train small agentic models that are highly capable of terminal use and coding? Announcing OpenThoughts-Agent + OpenThinkerAgent-32B, the strongest Qwen-3 based open-data agentic model: 44.8% avg across 7 agentic benchmarks! (1/n)
    Image
  • user avatar
    Ludwig Schmidt
    @lschmidt3
    Jun 23, 2025
    I'm a big fan of the approach to research funding @andykonwinski and the Laude team are taking! Working with them on terminal-bench has been fantastic (thanks @alexgshaw!) and I'm excited that they're going to support more open, impact-oriented research.
    user avatar
    Andy Konwinski
    @andykonwinski
    Jun 23, 2025
    Today, I’m launching a deeply personal project. I’m betting $100M that we can help computer scientists create more upside impact for humanity. Built for and by researchers, including @JeffDean & @jpineau1 on the board, @LaudeInstitute catalyzes research with real-world impact.
    Image
  • user avatar
    Ludwig Schmidt
    @lschmidt3
    Jun 5, 2025
    Very excited to finally release our paper for OpenThoughts! After DataComp and DCLM, this is the third large open dataset my group has been building in collaboration with the DataComp community. This time, the focus is on post-training, specifically reasoning data.
    Image
  • user avatar
    Ludwig Schmidt
    @lschmidt3
    May 30, 2025
    Cool to see more work on data for AI agents!
    user avatar
    Alex Ratner
    Snorkel AI
    @ajratner
    May 29, 2025
    Agentic AI will transform every enterprise–but only if agents are trusted experts. The key: Evaluation & tuning on specialized, expert data. I’m excited to announce two new products to support this–@SnorkelAI Evaluate & Expert Data-as-a-Service–along w/ our $100M Series D! ---
    Image
    00:00
  • user avatar
    Ludwig Schmidt
    @lschmidt3
    May 20, 2025
    Very excited about our new agent benchmark! I think it's a nice way of evaluating how well agents can do complex task in terminal (command line) environments.
    user avatar
    Mike A. Merrill
    @Mike_A_Merrill
    May 19, 2025
    Many agents (Claude Code, Codex CLI) interact with the terminal to do valuable tasks, but do they currently work well enough to deploy en masse? We’re excited to introduce Terminal-Bench: An evaluation environment and benchmark for AI agents on real-world terminal tasks. Tl;dr
    Image

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement