Log inSign up
Philipp Nazari
32 posts
Philipp Nazari profile banner
@philna00

Philipp Nazari

@philna00
PhD Candidate at Max Planck ETH Center for Learning Systems
phnazari.github.io
Joined August 2015
337
Following
119
Followers
RepliesRepliesRepostsRepostsMediaMedia

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
  • Pinned
    @philna00
    Philipp Nazari
    @philna00
    Feb 17
    🧵[1/5] Qwen3.5 demonstrates the potential of hybrid LLMs. But how well do the Linear Attention layers manage their associative memory? Previous research indicates a low effective rank, which we show: 1. Amplifies query noise 2. Poorly conditions gradients 3. Wastes memory.
    Image
    7
  • @philna00
    Philipp Nazari
    @philna00
    Aug 11
    Insanely cool work, straight out of Tübingen. Congrats to everybody involved!
    @kotekjedi_ml
    Alexander Panfilov
    @kotekjedi_ml
    Aug 11
    We can finally talk about it: We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company. We verified that our reasoning token count matches billed API thinking tokens 1:1 for most of the prompts we queried.
    Image
  • @philna00
    Philipp Nazari
    @philna00
    Jul 15
    Something I've realized using agents for research: they are very good at doing the first 80%. They are bad for the last 20%. But, since I didn't do the first 80% myself, we are cooked, resulting in a slow-down. Wondering to what extent this makes productivity gains an illusion?!
    1
  • @philna00
    Philipp Nazari
    @philna00
    Jun 22
    Check out this super cool paper on fixed-point transformers. Great work by everybody involved!
    @Sajad_Movahedi_
    Sajad Movahedi
    @Sajad_Movahedi_
    Jun 22
    Looped models scale compute by iterating in depth, demonstrating great potential in reasoning. But when should the looping stop, and how to keep it trainable at tens of thousands of unrolled layers? FPRM solves both. (1/7)
    Image
  • @philna00
    Philipp Nazari
    @philna00
    Apr 24
    Great fun! ☀️
    @tk_rusch
    Konstantin Rusch
    Liquid AI
    @tk_rusch
    Apr 24
    @makinai_ and @philna00 presenting our paper on calibration-free in-training compression of SSMs at #ICLR2026 in Rio 🇧🇷 Paper link: lnkd.in/dd3jr89x
    Image
Advertisement
Advertisement