1. X
  2. Daniel Weld
Log inSign up
Daniel Weld
1,046 posts
user avatar

Daniel Weld

@dsweld
Computer science prof, entrepreneur & leader at Ai2. Excited by AI for science, human-AI interaction, and Web-scale NLP.
Seattle, WA
cs.washington.edu/homes/weld/
Joined March 2009
271
Following
2,795
Followers
RepliesRepliesMediaMedia
  • user avatar
    Daniel Weld
    @dsweld
    May 4
    This benchmark for Ai scientific capabilities is beautifully thought out - I especially like it's clear enumeration of design principles...
    user avatar
    Ai2
    @allen_ai
    Apr 30
    New AstaBench results show frontier models making progress on scientific research, but the benchmark remains far from solved. Claude Opus 4.7 leads overall at 58.0%, while GPT-5.5 comes within 5.1 points at less than half the measured cost per problem. 🧵
    Image
  • user avatar
    Daniel Weld
    @dsweld
    Feb 4
    Truly open scientific question answering - that's good!
    Image
    Open-source AI program can answer science questions better than humans
    From science.org
  • user avatar
    Daniel Weld
    @dsweld
    Jan 28
    I'm so excited by this! Our system is generating some insightful & novel theories (e.g., internally for LM post-training). And it's still getting better!
    user avatar
    Ai2
    @allen_ai
    Jan 28
    Introducing Theorizer: Turning thousands of papers into scientific laws 📚➡️📜 Most automated discovery systems focus on experimentation. Theorizer tackles the other half of science: theory building—compressing scattered findings into structured, testable claims. 🧵
    Image
  • user avatar
    Daniel Weld
    @dsweld
    Jan 13
    Smart analysis analysis of scholar output when authors adopted LLMs as part of their writing: 1) huge 36% boost in # papers published 2) LLMs mitigate skill disparities, eg native language - enough to shift market share of production toward China bit.ly/4qliJGo @yian_yin
  • user avatar
    Daniel Weld
    @dsweld
    Nov 19, 2025
    Impressive deep-research performance by a tiny & open model!
    user avatar
    Rulin Shao
    @RulinShao
    Nov 18, 2025
    🔥Thrilled to introduce DR Tulu-8B, an open long-form Deep Research model that matches OpenAI DR 💪Yes, just 8B! 🚀 The secret? We present Reinforcement Learning with Evolving Rubrics (RLER) for long-form non-verifiable DR tasks! Our rubrics: - co-evolve with the policy model -
    Image

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement