1. X
  2. Andrej Karpathy
Log inSign up
Andrej Karpathy
10.1K posts
Image
user avatar
Andrej Karpathy
@karpathy
deep learning
Earth
Joined April 2009
1,124
Following
3.7M
Followers
RepliesRepliesArticlesArticlesMediaMedia
  • user avatar
    Andrej Karpathy
    @karpathy
    Aug 2
    We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for
    Image
    00:00
  • user avatar
    Andrej Karpathy
    @karpathy
    Jul 21
    One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10
  • user avatar
    Andrej Karpathy
    @karpathy
    Jun 23
    This is a new paradigm for interacting with Claude that is significantly more "inline" with all the other human activity org-wide. Once you do all of the under the hood engineering work to make this "just work" (e.g. across tools, integrations, compute environments, memory,
    user avatar
    Claude
    Anthropic
    @claudeai
    Jun 23
    Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while you focus on other work.
    Image
    00:00
  • user avatar
    Andrej Karpathy
    @karpathy
    Jun 12
    In awe of SpaceX and its story - past, present and the future. You can think about it in 10+ different ways and continue re-blowing your mind in circles. Huge congrats to the team! ๐Ÿš€
  • user avatar
    Andrej Karpathy
    @karpathy
    Jun 9
    This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great and it's SOTA on everything by a margin but I'll add that *qualitatively* also, this is a major-version-bump-deserving step change forward
    user avatar
    Claude
    Anthropic
    @claudeai
    Jun 9
    Replying to @claudeai
    Fable 5 is state-of-the-art on nearly all tested benchmarks, with exceptional performance in software engineering, knowledge work, scientific research, and vision. The longer and more complex the task, the larger Fable 5โ€™s lead over our other models.
    Benchmark table titled Mythos 5 & Fable 5, comparing Claude Mythos 5 and Fable 5 against Claude Mythos Preview, Claude Opus 4.8, GPT 5.5, and Gemini 3.1 Pro.

Log in or sign up for X

See whatโ€™s happening and join the conversation

Continue with phone
or
Log in with username or email
TermsยทPrivacyยทCookiesยทAccessibilityยทAds Infoยทยฉ 2026 X Corp.
Advertisement
Advertisement