Log inSign up
Dan Robinson
Paradigm
10.2K posts
Dan Robinson profile banner
@danrobinson

Dan Robinson

Paradigm
@danrobinson
coder / lawyer. research at @paradigm. automated research reply guy
San Francisco, CA
Joined March 2011
1,419
Following
85.4K
Followers
RepliesRepliesRepostsRepostsMediaMediaArticlesArticles

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • @danrobinson
    Dan Robinson
    Paradigm
    @danrobinson
    Sep 3
    Natural language autoencoders are the discovery that most makes me question my research intuition You jointly train two models: one to turn activations into English, and one to turn English into activations. The objective for the second model is to invert the first one, and the
    @AnthropicAI
    Anthropic
    @AnthropicAI
    May 7
    New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers—called activations—encode Claude’s thoughts, but not in a language we can read. Here, we train Claude to translate its activations into human-readable text.
    Image
    00:00
    13
  • @danrobinson
    Dan Robinson
    Paradigm
    @danrobinson
    Sep 3
    Anti-AI + anti-Pangram is the funniest quadrant of the opinions spectrum You are not fooling anyone
    3
  • @danrobinson
    Dan Robinson
    Paradigm
    @danrobinson
    Sep 2
    Claude loves coming up with things to count and then counting them and displaying the count in the web page it made you
    9
  • @danrobinson
    Dan Robinson
    Paradigm
    @danrobinson
    Sep 2
    I think "category error" is a massively overused term People who throw it around in policy arguments (like "worrying about risk from rogue AI agents is a category error") are usually making a, um, category error
    8
  • @danrobinson
    Dan Robinson
    Paradigm
    @danrobinson
    Sep 2
    I get it but the cybersecurity nerds need to let the alignment nerds have this one Doesn’t really matter how the agents got control of OpenAI infrastructure. “It was a stack overflow in the cache buffer” or w/e. Another zero day; add it to the pile The crazy question is why
    @DrTechlash
    Nirit Weiss-Blatt, PhD
    @DrTechlash
    Sep 1
    This video captures the crux of the debate: Cybersecurity experts are pissed off by the METR/Redwood Research report because AI alignment has sucked the oxygen out of AI security, even though the OpenAI incident is primarily a security issue. Fiction is the product: The
    13
Advertisement
Advertisement