Great new work, spearheaded by @xiulin_yang, about how to fairly compare models across languages. To be presented at #EMNLP2026!!
🍎🍊 How would you know if a language model is better at one language than another? Our #EMNLP2026 paper argues that only one metric can actually lead to fair crosslingual evaluation.
This work is a collaboration with @weGotlieb & @linguist_cat! (1/5)



