Sign in

Wessel Poelman

@wpoelman.bsky.social
19 followers 3 following 1 posts

Working with languages @ KU Leuven

PostsRepliesMedia
Wessel Poelman @wpoelman.bsky.social · 28/01/2026
New EACL paper (with @mdlhx.bsky.social)! We tested if comparing perplexity of parallel data across languages is fair. Turns out: it depends. We show the choice of test set (even with consistent meaning) can flip conclusions about which language is easier to model. Paper: arxiv.org/abs/2601.10580
arxiv.org
Form and Meaning in Intrinsic Multilingual Evaluations
Intrinsic evaluation metrics for conditional language models, such as perplexity or bits-per-character, are widely used in both mono- and multilingual settings. These metrics are rather straightforwar...
0103
Reposted by Wessel Poelman
LAGoM NLP @lagom-nlp.bsky.social · 12/12/2024
@wpoelman.bsky.social and @mdlhx.bsky.social 's 🔥 hot takes on multilingual LLM evaluation, to appear @nodalida.bsky.social is up on arXiv: arxiv.org/abs/2412.08392
arxiv.org
The Roles of English in Evaluating Multilingual Language Models
Multilingual natural language processing is getting increased attention, with numerous models, benchmarks, and methods being released for many languages. English is often used in multilingual evaluati...
2141