Log inSign up
Ludwig Schmidt
253 posts
@lschmidt3

Ludwig Schmidt

@lschmidt3
Assistant professor at @Stanford and member of the technical staff at @AnthropicAI.
Palo Alto, CA
people.csail.mit.edu/ludwigs/
Joined August 2009
426
Following
6,701
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @lschmidt3
    Ludwig Schmidt
    @lschmidt3
    Aug 31
    Even more Terminal-Bench!
    @ryan_marten
    Ryan Marten
    @ryan_marten
    Aug 29
    We've pushed a version update to the Terminal-Bench dataset and leaderboard. Terminal-Bench 4.0 calibrates task resources (time, cpu, memory), implements task fixes, and removes saturated tasks.
    Image
    00:00
  • @lschmidt3
    Ludwig Schmidt
    @lschmidt3
    Aug 28
    Very excited about this new direction for Terminal-Bench!
    @StevenDillmann
    Steven Dillmann
    @StevenDillmann
    Aug 27
    We're releasing Terminal-Bench-Science: a benchmark for evaluating AI agents on research workflows across scientific domains. An ongoing Stanford-led community effort, built by the team behind Terminal-Bench together with scientific domain experts at research institutions
    Image
    2
  • @lschmidt3
    Ludwig Schmidt
    @lschmidt3
    Jun 25
    Very excited to release the next project in the DataComp / OpenThoughts line of research! Like OpenThoughts we worked on post-training data, this time with a focus on agentic models.
    @RichardZ412
    Richard Zhuang
    @RichardZ412
    Jun 24
    How can we train small agentic models that are highly capable of terminal use and coding? Announcing OpenThoughts-Agent + OpenThinkerAgent-32B, the strongest Qwen-3 based open-data agentic model: 44.8% avg across 7 agentic benchmarks! (1/n)
    Image
    4
  • @lschmidt3
    Ludwig Schmidt
    @lschmidt3
    Jun 23, 2025
    I'm a big fan of the approach to research funding @andykonwinski and the Laude team are taking! Working with them on terminal-bench has been fantastic (thanks @alexgshaw!) and I'm excited that they're going to support more open, impact-oriented research.
    @andykonwinski
    Andy Konwinski
    @andykonwinski
    Jun 23, 2025
    Today, I’m launching a deeply personal project. I’m betting $100M that we can help computer scientists create more upside impact for humanity. Built for and by researchers, including @JeffDean & @jpineau1 on the board, @LaudeInstitute catalyzes research with real-world impact.
    Image
    2
  • @lschmidt3
    Ludwig Schmidt
    @lschmidt3
    Jun 5, 2025
    Very excited to finally release our paper for OpenThoughts! After DataComp and DCLM, this is the third large open dataset my group has been building in collaboration with the DataComp community. This time, the focus is on post-training, specifically reasoning data.
    Image
    22
Advertisement
Advertisement