1. X
  2. Goodfire
Log inSign up
Goodfire
828 posts
Goodfire profile banner
@GoodfireAI

Goodfire

@GoodfireAI
Using interpretability to understand, learn from, and design AI.
San Francisco
goodfire.ai
Joined August 2024
31
Following
26.7K
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @GoodfireAI
    Goodfire
    @GoodfireAI
    Aug 4
    Silico, the platform for ambitious AI research, is publicly available today. AI is advancing fast. The tools to understand it need to advance even faster. Silico lets you interpret and train your models at frontier scale. Learn more + get access 🧵
    Image
    00:00
  • @GoodfireAI
    Goodfire
    @GoodfireAI
    Aug 26
    LLMs are like Schrödinger’s cat: many possible trajectories, but you only see one outcome per run. To really understand models, and debug where they go wrong, you can find the "forking tokens" that lead to different trajectories. Our new research does this 100x more efficiently!
    Image
    00:00
  • @GoodfireAI
    Goodfire
    @GoodfireAI
    Aug 20
    We’re giving out $1M in grants of free Silico usage for academic and nonprofit researchers focused on AI interpretability and alignment. We feel extreme urgency about advancing interpretability for alignment, and we want to help more researchers push it forward. 🧵
    Image
    @eric_ho
    Eric Ho
    @eric_ho
    Aug 14
    in light of multiple models breaking containment, we've decided to focus our research at @GoodfireAI to solving AI alignment via interpretability. the hugging face incident is a turning point for the world where AI safety gets real. i am personally very concerned. i'm glad that
  • @GoodfireAI
    Goodfire
    @GoodfireAI
    Aug 14
    With just one prompt, we taught an LLM to see - then looked at its representations to debug where it fell short. Qwen 3 8B can read text, but has no way to see or understand images. Silico trained a vision adapter that matches the official Qwen 3 VL 8B in multiple benchmarks. 🧵
    Image
    Image
  • @GoodfireAI
    Goodfire
    @GoodfireAI
    Aug 13
    Autonomous agent swarms are hacking companies. How can interpretability help us understand where this behavior comes from and how to address it? Our cofounder @banburismus_ spoke with @JPBrebner from SPC about why interpretability matters right now, and how we can move toward
    @JPBrebner
    Jonathan Brebner
    @JPBrebner
    Aug 13
    What does the former co-lead of interpretability at Deepmind and now co-founder of @GoodfireAI think about how to keep AI from becoming (his words) "a generally evil guy"? I got to speak with @banburismus_ at @spc the day after OpenAI announced the Hugging Face hack. Good
    Image
    00:00
Advertisement
Advertisement