1. X
  2. Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
Log inSign up
Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
25.9K posts
Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ² profile banner
user avatar

Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²

@LeonDerczynski
security of models & security with models. research and policy @NVIDIA, prof @ITUkbh. views ostensibly professional. llmsec stan acct
Seattle / Copenhagen
developer.nvidia.com/blog/author/ld…
Joined January 2012
1,204
Following
7,045
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
    @LeonDerczynski
    Jun 13, 2023
    Proud to announce: πŸ’« garak - an LLM vulnerability scannerπŸ’« πŸ”Ž Check if a model is susceptible to common attacks 🦜 Supports HuggingFace, OpenAI, ggml, Cohere, ... πŸ”§ >70 probes: prompt injection, false claims, toxicity, encoding evasion, ..
    Image
    GitHub - NVIDIA/garak: the LLM vulnerability scanner
    From github.com
  • user avatar
    Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
    @LeonDerczynski
    5h
    Nice post-mortem from @Xbow on 1600 agent-derived vulnerabilities. Takeaways: * benchmarks are saturated * multi-step attacks are commonplace * threat is rising * offense drives defense
    Image
  • user avatar
    Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
    @LeonDerczynski
    8h
    Agentic AI Challenges Progress in Confidential Computing securing data on gpus is non-trivial. encryption has to be done at low level to stop accidental exposures. lots of the slowdowns here have now been addressed. maybe confidential computing will finally happen?
    Image
  • user avatar
    Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
    @LeonDerczynski
    Aug 13
    Decent history of distilliation-type events, where information is reconstructed Machine learning distilliation is just another instance of a teacher-student dynamic, where one can reconstruct information efficiently. Even basic synthesis of research is distillation. It's a
    Image
  • user avatar
    Leon Derczynski βš’οΈβ˜οΈπŸ”οΈπŸŒ²
    @LeonDerczynski
    Aug 13
    Every context is different, by definition. Being able to fit tools to context is therefore critical - and this generalisation holds incredibly well. It even holds for models and information technology: being able to tune a model to a task, gets it to perform better in context.
    Image
    Nemotron Labs: How Open Models Give Enterprises and Nations AI They Can Trust, Control and Customize
    From blogs.nvidia.com

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
TermsΒ·PrivacyΒ·CookiesΒ·AccessibilityΒ·Ads InfoΒ·Β© 2026 X Corp.
Advertisement
Advertisement