1. X
  2. Rasool Fakoor
Log inSign up
Rasool Fakoor
628 posts
Image
user avatar
Rasool Fakoor
@rasoolfa
Building agents that reason, adapt, and act with RL and friends!
rasoolfa.github.io
Joined December 2012
1,763
Following
499
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Rasool Fakoor
    @rasoolfa
    Jun 15
    Too many ideas die at the edge of the LLM/VLM/VLA training framework. Not anymore. Excited to release FeynRL: built to help you understand, modify, and build new RL methods without fighting the stack. github.com/FeynRL-project… Try it, ⭐ it, and send feedback.
    Image
    GitHub - FeynRL-project/FeynRL: Post-training framework for large models, from new objectives to...
    From github.com
  • user avatar
    Rasool Fakoor
    @rasoolfa
    Jun 10
    This is exactly why open-source frameworks like FeynRL matter more than ever. Open weights are NOT enough. Open AI needs open training recipes, alg, etc. You should be able to understand training, modify it, & build new algorithms, optimizers, recipes. github.com/FeynRL-project…
    user avatar
    elie
    Prime Intellect
    @eliebakouch
    Jun 9
    mythos will be bad ON PURPOSE on ai "frontier llm research" tasks, this is very very sad for the research community also the fact that this is un purpose not visible to the user is crazy
    Image
  • user avatar
    Rasool Fakoor
    @rasoolfa
    Jun 2
    Off-policy data does not have to be a bug in RL. In our work, we shift the question from: Is this data on-policy? -> How much should we trust this batch? That change leads to a adaptive objective for RL LLM-training. blog: feynrl-project.github.io
  • user avatar
    Rasool Fakoor
    @rasoolfa
    May 26
    Join us tmrw if you are around at RLEval, first edition ever. Together with my wonderful co-organizers (@jomulr , @anishathalye, Alina Gavrilov, Aziza Mirsaidova) we have put together an exciting program with a great lineup of speakers and papers. See the full program here:
    user avatar
    Alex Smola
    @smolix
    May 26
    Tomorrow in San Jose: RLEval. Trillions going into LLM agents and we still cannot reliably evaluate them. 19 papers, talks from Alex Dimakis, Corby Rosset, Rasool Fakoor, and others. I'll be presenting Submodular Benchmark Selection from @boson_ai. rl-eval.github.io
  • user avatar
    Rasool Fakoor
    @rasoolfa
    Apr 24
    One thing I keep hearing is that RL for L(L)Ms is "mostly a systems problem now" and the RL part is basically good enough. I really don’t buy that. Current RL algs are still fragile as hell. Better systems help, but they don’t magically make the RL problem go away.

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement