1. X
  2. Adam Belfki
Log inSign up
Adam Belfki
43 posts
Adam Belfki profile banner
user avatar

Adam Belfki

@adambelfki
Software Engineer @ndif_team | CS @Northeastern 24' AI Interpretability & Safety.
Joined July 2018
301
Following
102
Followers
RepliesRepliesMediaMedia
  • user avatar
    Adam Belfki
    @adambelfki
    Jul 28
    We’re at the Oculus building in NYC helping the public understand how LLMs work using workbench.ndif.us. Come see us!
    user avatar
    Jaden Fiotto-Kaufman
    @jadenfk23
    Jul 28
    NDIF @ PIT Pop Up. Join me! luma.com/rwbs06zm
  • user avatar
    Adam Belfki
    @adambelfki
    Jun 4
    i know this guy what you should know is that @ndif_team is THE place for how to do interp in the vision space
    user avatar
    How Do Vision Models Work? @ CVPR2026 (Prev: MIV)
    @how_cvpr2026
    Jun 4
    We have our first keynote speaker @jadenfk23 from @ndif_team giving a hands-on tutorial showing how nnsight and ndif make interpretability research easy to use and less scary!
    Image
  • user avatar
    Adam Belfki
    @adambelfki
    Jun 4
    The team has worked very hard to build a white-box hackathon infrastructure so you can investigate and control model internals, compute-free! Join the competition and help develop principled methods for understanding latent model beliefs.
    user avatar
    NDIF
    @ndif_team
    Jun 4
    Can you tell when an AI model is lying? Announcing Aletheia's Quest, an AI lie detection challenge running this summer, organized by @cadenza_labs and @ndif_team. Multiple model organisms to interrogate and probe, $50K prize pool, no local GPU required.
    Image
  • user avatar
    Adam Belfki
    @adambelfki
    May 19
    When @sheridan_feucht first told me about these results I was kind of skeptical, until they mentioned I can trace the modulo base-10 addition in Llama 8B just using Logit Lens. 🔍 So I opened workbench.ndif.us to check it out myself, and this is what I saw:
    Image
    Image
    user avatar
    Sheridan Feucht
    @sheridan_feucht
    May 14
    Replying to @sheridan_feucht
    Turns out, this is because MLP 18 does base-10 addition. Llama first computes a sum (e.g., "four months after October" → 4+10=14), and only in later layers applies modulo to map back to a month (14→Feb). We find that Llama re-uses this addition mechanism across several tasks.
  • user avatar
    Adam Belfki
    @adambelfki
    May 7
    Really fun putting this together! I hope this helps you scale up your experiments on @ndif_team 🚀
    user avatar
    Gabriele Sarti
    @gsarti_
    May 6
    Fellow mechinterp researchers: is @ndif_team NNsight slow, or are you just using it wrong? 🙃 A new tutorial by @adambelfki shows how to speed up large-scale experiments with session, batching, caching and skipping for 130x speedup! 🔥 Check it out ⬇️ nnsight.net/blog/2026/04/3…

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement