1. X
  2. Philippe Laban
Log inSign up
Philippe Laban
420 posts
user avatar
Philippe Laban
@PhilippeLaban
Research Scientist @MSFTResearch. NLP/HCI Research.
New York City
Joined April 2022
836
Following
1,553
Followers
RepliesRepliesMediaMedia
  • user avatar
    Philippe Laban
    @PhilippeLaban
    Jul 28
    Check out Jihoon's very interesting work on simulating long "situated" conversations where the user can change their mind, switch tasks, underspecify, etc. We need more work evaluating model performance outside of the single-turn, single-task, lab-like setting.
    user avatar
    Jihoon Tack
    @jihoontack
    Jul 27
    Excited to share my first work after joining @MSFTResearch! LLMs have entered the agentic era, and we now collaborate with agents on complex tasks over many turns of interaction. But does your agent actually follow what you intended? We show where agents get lost: LLMs Get Lost
    Image
    1.7K
  • user avatar
    Philippe Laban
    @PhilippeLaban
    Apr 21
    New paper! LLMs Corrupt Your Documents When You Delegate LLMs are enabling a new way of working: delegated work, where users supervise an LLM as it edits documents on their behalf. Delegation requires trust: does the LLM complete tasks without introducing errors? We simulate
    Image
    00:00
    121K
  • user avatar
    Philippe Laban
    @PhilippeLaban
    Feb 24
    LLMs *Still* Get Lost In Multi-Turn Conversation. We re-ran experiments with newer models. Performance still drops, but with modest gains: mostly from improvements on the Python coding task. Also: Lost in Conversation will be presented at ICLR 2026 šŸŽ‰šŸ‡§šŸ‡·
    Image
    23K
  • user avatar
    Philippe Laban
    @PhilippeLaban
    Feb 10
    Very cool work!
    user avatar
    Ai2
    @allen_ai
    Feb 10
    LLMs often generate step-by-step instructions, from real-world tasks (how do I file taxes?) to plans for AI agents. Improving this is hard: outputs can sound fluent for steps that don't work, and current datasets cover few domains. How2Everything evals/trains for this at scale.
    Image
    919
  • user avatar
    Philippe Laban
    @PhilippeLaban
    Oct 6, 2025
    Come see us in COLM!! More importantly, if you're thinking of doing a PhD, go work with wonderful Tuhin!
    user avatar
    Tuhin Chakrabarty
    @TuhinChakr
    Oct 6, 2025
    I am @COLM_conf in Montreal. @PhilippeLaban and I will present work on #AISlop and Calibrated Reward Models for Writing. I will also be admitting 1 PhD student next fall at @sbucompsc to work on Human Centered AI / AI detection / Copyright and Creative Labor. Reach out !!
    Image
    7.7K
  • See @PhilippeLaban's full profile

    Sign up
    Log in

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
TermsĀ·PrivacyĀ·CookiesĀ·AccessibilityĀ·Ads InfoĀ·Ā© 2026 X Corp.
Advertisement
Advertisement