1. X
  2. Eugene Yan
Log inSign up
Eugene Yan
4,790 posts
Eugene Yan profile banner
user avatar

Eugene Yan

@eugeneyan
MTS @AnthropicAI. Prev: Principal Applied Scientist @Amazon, led ML @ Alibaba, Healthtech startup.
Seattle ⇄ SF
eugeneyan.com
Joined April 2009
617
Following
28.1K
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Eugene Yan
    @eugeneyan
    Jun 25
    How do we eval if a model can find and exploit vulnerabilities? We discuss some benchmarks and the common pattern: • A sandboxed target within Docker containers • Inputs: code only (0-day), with patch (1-day scenario) • Tools such as bash, static analyzers, etc. • A grader to
    Image
    Patterns for Building Cybersecurity Evals
    From eugeneyan.com
  • user avatar
    Eugene Yan
    @eugeneyan
    Jul 22
    When evaling models, we anchor on the median task. But this is like how devs estimate the median task accurately but underestimate the mean which tends to be ~2x estimated (thus "double your estimates"). This applies to models too—we optimize for time/cost on median tasks. But
    Image
    user avatar
    Steve Yegge
    @Steve_Yegge
    Jul 20
    Fable is careful. None of the other models are careful. GPT-5.6 Sol, Opus, Kimi, Grok. You can compare them all day long on capabilities, and it doesn't matter, because in order to use a model for real production work, it must first and foremost be careful. That's the only
  • user avatar
    Eugene Yan
    @eugeneyan
    Jul 1
    I’m at @aiDotEngineer and hanging out around the music corner on the 2nd floor from 1415 - 1515! Come by to chat about eugeneyan.com/writing/workin…, eugeneyan.com/writing/cybers…, claude.com/blog/using-llm…, evals, agents, memory, how to work effectively with claude code, claude tag, etc!
    Image
    How to Work and Compound with AI
    From eugeneyan.com
  • user avatar
    Eugene Yan
    @eugeneyan
    Jun 23
    I've been loving the multiplayer form factor of Claude Tag. Now others can reply on the thread to provide context and direction to Claude!
    user avatar
    Claude
    Anthropic
    @claudeai
    Jun 23
    Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while you focus on other work.
    Image
    00:00
  • user avatar
    Eugene Yan
    @eugeneyan
    Jun 3
    Stronger models have made finding vulnerabilities easier, and the bottleneck has shifted to verification, triage, patching. Here are some lessons from working with security teams to address the new bottlenecks.
    Image
    Using LLMs to secure source code | Claude by Anthropic
    From claude.com

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement