Pinned
If I had to compress my PhD into one idea, it is this
"The data a model sees early in training leaves an imprint on its representations that is very hard to undo later"
This thread runs through
- Rephrasing the Web
- Safety Pretraining
- TOFU
This is the Finetuner’s Fallacy🧵






