1. X
  2. byron wallace
Log inSign up
byron wallace
517 posts
Image
user avatar
byron wallace
@byron_c_wallace
Prof @Northeastern in Computer Science. NLP / ML / + Health &etc. he/him.
byronwallace.com
Joined July 2014
767
Following
1,474
Followers
RepliesRepliesMediaMedia
  • user avatar
    byron wallace
    @byron_c_wallace
    Oct 21, 2024
    Come chat about @somin's intriguing CoT distillation results — put reasoning after labels, permute or keep only a few key CoT tokens — in Miami!
    user avatar
    Somin W
    @SominW
    Oct 21, 2024
    Planning to attend @emnlpmeeting? Come check out our work/poster in Miami on 11/12 🏝️
    1.1K
  • user avatar
    byron wallace
    @byron_c_wallace
    Jul 1, 2024
    Sheridan has some cool results on tokenization in LLMs and their "implicit vocabularies"👇
    user avatar
    Sheridan Feucht
    @sheridan_feucht
    Jul 1, 2024
    (new preprint) LLMs live in a strange tokenized world. We find that LLMs learn to deal with the weirdness of tokenization by converting tokens into word-like representations and then "forgetting about" those tokens. Maybe this is why tokenization isn't an issue, until it is... 🧵
    1.1K
  • user avatar
    byron wallace
    @byron_c_wallace
    Jun 24, 2024
    Somin has some cool results on CoT + distillation👇
    user avatar
    Somin W
    @SominW
    Jun 24, 2024
    📢We know that including CoT rationales as supervision improves model distillation. But why? New work (w/ @silvio_amir and @byron_c_wallace) research unveils surprising insights! 🔗 Full paper: arxiv.org/abs/2406.14511 [1/5]
    1.1K
  • user avatar
    byron wallace
    @byron_c_wallace
    Jun 4, 2024
    + @HibaAhsan5!
    user avatar
    chilconference
    @CHILconference
    Jun 4, 2024
    Learn about "Retrieving Evidence from EHRs with LLMs: Possibilities and Challenges" by @JeredMcinerney, @silvio_amir, @byron_c_wallace at #CHIL2024!
    630
  • user avatar
    byron wallace
    @byron_c_wallace
    May 3, 2024
    Come chat with me or @ChantalShaib about @SunJiuding's work on the sensitivity of models to instruction phrasings at #ICLR2024 next week ↓
    user avatar
    Jiuding Sun
    @SunJiuding
    Jun 21, 2023
    How robust are the instructions in your instruction-tuned model? In our most recent work (w/ @ChantalShaib and @byron_c_wallace), we show that there is a considerable dip in performance on in-domain tasks when you slightly vary the instruction. arxiv.org/abs/2306.11270
    Image
    1.5K

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement