1. X
  2. Dawid Kopiczko
Log inSign up
Dawid Kopiczko
45 posts
user avatar
Dawid Kopiczko
@dawkopi
PhD in progress
dkopi.github.io
Joined April 2020
505
Following
136
Followers
RepliesRepliesMediaMedia
  • Pinned
    user avatar
    Dawid Kopiczko
    @dawkopi
    Feb 16
    Replying to @dawkopi
    Why repetition works so well is still an open question. There's a lot to uncover about training dynamics of SFT, and we hope this is a useful data point. Joint work with co-authors @Sagar_Vaze @TiRune @y_m_asano Paper: arxiv.org/abs/2602.11149 Code: github.com/dkopi/data-rep…
    arXiv logo
    arxiv.org
    Data Repetition Beats Data Scaling in Long-CoT Supervised Fine-Tuning
    Supervised fine-tuning (SFT) on chain-of-thought data is an essential post-training step for reasoning language models. Standard machine learning intuition suggests that training with more unique...
    1.4K
  • user avatar
    Dawid Kopiczko
    @dawkopi
    Feb 16
    Common knowledge in ML: more unique training data → better generalization. Turns out this doesn't hold for long-CoT SFT. Under a fixed update budget, repeating a small dataset multiple times beats training on more unique samples. And it's not even close.
    Image
    20K

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email
Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.
Advertisement
Advertisement