Sign in

Arianna Bisazza

@arianna-bis.bsky.social
223 followers 131 following 41 posts

Associate Professor at GroNLP ( @gronlp.bsky.social‬ ) #NLP | Multilingualism | Interpretability | Language Learning in Humans vs NeuralNets | Mum^2 Head of the InClow research group: inclow-lm.github.io

PostsRepliesMedia
Reposted by Arianna Bisazza
Vera Neplenbroek @veraneplenbroek.bsky.social · 04/06/2026
When you ask an LLM "What job can I get with no prior experience?" and then get recommended a low salary, would your background have changed the answer you receive? Surprisingly, perhaps not. Check out our paper with @gsarti.com @arianna-bis.bsky.social and Raquel Fernández: arxiv.org/abs/2606.02776
1122
Reposted by Arianna Bisazza
Francesca Padovani @frap98.bsky.social · 20/05/2026
You can access the models through the repository: github.com/fpadovani/CA... Thank you to the amazing team: @xiulinyang.bsky.social @bbunzeck.bsky.social @arianna-bis.bsky.social @yevgenm.bsky.social @jumelet.bsky.social
github.com
GitHub - fpadovani/CAIT-Toolkit: Syntactic Annotation Toolkit for Child–Adult InTeractions (CAIT)
Syntactic Annotation Toolkit for Child–Adult InTeractions (CAIT) - fpadovani/CAIT-Toolkit
042
Reposted by Arianna Bisazza
Francesca Padovani @frap98.bsky.social · 20/05/2026
Interested in analyzing 𝗖𝗵𝗶𝗹𝗱-𝗔𝗱𝘂𝗹𝘁 𝗜𝗻𝘁𝗲𝗿𝗮𝗰𝘁𝗶𝗼𝗻𝘀 with computational tools? Then the 𝗖𝗔𝗜𝗧 𝗧𝗼𝗼𝗹𝗸𝗶𝘁 🪁 (pronounced [kaIt]) might be just what you’re looking for! It includes a brand-new i𝗻-𝗱𝗼𝗺𝗮𝗶𝗻 𝗱𝗲𝗽𝗲𝗻𝗱𝗲𝗻𝗰𝘆 𝗽𝗮𝗿𝘀𝗲𝗿, 𝗣𝗢𝗦 𝘁𝗮𝗴𝗴𝗲𝗿, and 𝘂𝘁𝘁𝗲𝗿𝗮𝗻𝗰𝗲-𝗹𝗲𝘃𝗲𝗹 𝗰𝗼𝗻𝘀𝘁𝗿𝘂𝗰𝘁𝗶𝗼𝗻 𝘁𝗮𝗴𝗴𝗲𝗿.
lnkd.in
LinkedIn
This link will take you to a page that’s not on LinkedIn
1188
Reposted by Arianna Bisazza
Francesca Padovani @frap98.bsky.social · 14/04/2026
I’m very happy to share that my latest paper on the 𝐚𝐜𝐪𝐮𝐢𝐬𝐢𝐭𝐢𝐨𝐧 𝐨𝐟 𝐯𝐞𝐫𝐛 𝐦𝐞𝐚𝐧𝐢𝐧𝐠, tested on models trained under CDL data vs ADL data , has been accepted for an 𝐨𝐫𝐚𝐥 𝐩𝐫𝐞𝐬𝐞𝐧𝐭𝐚𝐭𝐢𝐨𝐧 at the upcoming edition of 𝐂𝐨𝐠𝐒𝐜𝐢, which will take place at the end of July in Rio de Janeiro.💃
1246
Reposted by Arianna Bisazza
Gabriele Sarti @gsarti.com · 04/01/2026
Our work on contrastive SAE steering for personalizing literary machine translation was accepted to EACL main! 🎉 Check it out! ⬇️
0162
Arianna Bisazza @arianna-bis.bsky.social · 09/02/2026
📢 Paper alert! We know typological features can drive the difficulty of language modeling & machine translation in highly controlled setups (w/ relatively small monolingual models) But do they also drive MT quality in the age of massively multilingual LLMs? See @v-hirak.bsky.social’s thread ⬇️
1111
Reposted by Arianna Bisazza
GroNLP @gronlp.bsky.social · 18/12/2025
👀 Look what 🎅 has broght just before Christmas 🎁: a brand new Research Master in Natural Language Processing at @facultyofartsug.bsky.social @rug.nl Program: www.rug.nl/masters/natu... Applications (2026/2027) are open! Come and study with us (you will also learn why we have a 🐮 in our logo)
rug.nl
Natural Language Processing
How do you build Large Language Models? How do humans experience Natural Language Processing (NLP) applications in their daily lives? And how can we...
02515
Reposted by Arianna Bisazza
Gabriele Sarti @gsarti.com · 07/11/2025
Wrapping up my oral presentations today with our TACL paper "QE4PE: Quality Estimation for Human Post-editing" at the Interpretability morning session #EMNLP2025 (Room A104, 11:45 China time)! Paper: arxiv.org/abs/2503.03044 Slides/video/poster: underline.io/lecture/1315...
2101
Arianna Bisazza @arianna-bis.bsky.social · 07/11/2025
Interested in agent simulations of language change & pragmatic naming behavior? Come check our poster TODAY (Fri, Nov 7, 12:30 - 13:30) #EMNLP!
071
Arianna Bisazza @arianna-bis.bsky.social · 06/11/2025
Benchmarks of linguistic minimal pairs are key for LM evaluation & help us overcome the English-centric bias in NLP research Come to our poster TODAY (Fr 7 Nov 10.30-12.00) #EMNLP to meet TurBLiMP, a new benchmark for Turkish, revealing how LLMs deal with free-order, morphologically rich languages
061
Reposted by Arianna Bisazza
Jaap Jumelet @jumelet.bsky.social · 06/11/2025
I'm in Suzhou to present our work on MultiBLiMP, Friday @ 11:45 in the Multilinguality session (A301)! Come check it out if your interested in multilingual linguistic evaluation of LLMs (there will be parse trees on the slides! There's still use for syntactic structure!) arxiv.org/abs/2504.02768
0277
Arianna Bisazza @arianna-bis.bsky.social · 06/11/2025
Interested in developmentally plausible LMs, and the role of child-directed language data? Come to our poster TODAY (Fr 7 Nov, 10.30-12.00) #EMNLP!
082
Arianna Bisazza @arianna-bis.bsky.social · 06/11/2025
There’s more to Neural Nets than big fat LLMs! We’ve built a NN-agent framework to simulate how people choose the best word in a given communication context (i.e. pragmatic naming behavior). With @yuqing0304.bsky.social, @ecesuurker.bsky.social, Tessa Verhoef, @gboleda.bsky.social
142
Arianna Bisazza @arianna-bis.bsky.social · 31/10/2025
Thrilled to be heading to Suzhou with a big team of GroNLP'ers 🐮 Interested in Interpretable, Cognitively inspired, Low-resource LMs? Don't miss our posters & talks #EMNLP2025!
1143
Reposted by Arianna Bisazza
Jirui Qi @jiruiqi.bsky.social · 30/05/2025
[1/]💡New Paper Large reasoning models (LRMs) are strong in English — but how well do they reason in your language? Our latest work uncovers their limitation and a clear trade-off: Controlling Thinking Trace Language Comes at the Cost of Accuracy 📄Link: arxiv.org/abs/2505.22888
185
Reposted by Arianna Bisazza
Francesca Padovani @frap98.bsky.social · 14/10/2025
𝐃𝐨 𝐲𝐨𝐮 𝐫𝐞𝐚𝐥𝐥𝐲 𝐰𝐚𝐧𝐭 𝐭𝐨 𝐬𝐞𝐞 𝐰𝐡𝐚𝐭 𝐦𝐮𝐥𝐭𝐢𝐥𝐢𝐧𝐠𝐮𝐚𝐥 𝐞𝐟𝐟𝐨𝐫𝐭 𝐥𝐨𝐨𝐤𝐬 𝐥𝐢𝐤𝐞? 🇨🇳🇮🇩🇸🇪 Here’s the proof! 𝐁𝐚𝐛𝐲𝐁𝐚𝐛𝐞𝐥𝐋𝐌 is the first Multilingual Benchmark of Developmentally Plausible Training Data available for 45 languages to the NLP community 🎉 arxiv.org/abs/2510.10159
24016
Reposted by Arianna Bisazza
Vilém Zouhar @zouhar.bsky.social · 20/10/2025
📢 Announcing the First Workshop on Multilingual and Multicultural Evaluation (MME) at #EACL2026 🇲🇦 MME focuses on resources, metrics & methodologies for evaluating multilingual systems! multilingual-multicultural-evaluation.github.io 📅 Workshop Mar 24–29, 2026 🗓️ Submit by Dec 19, 2025
13415
Reposted by Arianna Bisazza
Vera Neplenbroek @veraneplenbroek.bsky.social · 21/08/2025
Delighted to share that our paper "Reading Between the Prompts: How Stereotypes Shape LLM's Implicit Personalization" (joint work with @arianna-bis.bsky.social and Raquel Fernández) got accepted to the main conference of #EMNLP Can't wait to discuss our work at #EMNLP2025 in Suzhou this November!
0142
Arianna Bisazza @arianna-bis.bsky.social · 19/06/2025
Proud to introduce TurBLiMP, the 1st benchmark of minimal pairs for free-order, morphologically rich Turkish language! Pre-print: arxiv.org/abs/2506.13487 Fruit of an almost year-long project by amazing MS student @ezgibasar.bsky.social in collab w/ @frap98.bsky.social and @jumelet.bsky.social
arxiv.org
TurBLiMP: A Turkish Benchmark of Linguistic Minimal Pairs
We introduce TurBLiMP, the first Turkish benchmark of linguistic minimal pairs, designed to evaluate the linguistic abilities of monolingual and multilingual language models (LMs). Covering 16 linguis...
1112
Arianna Bisazza @arianna-bis.bsky.social · 31/05/2025
One step further in our quest to bring interpretability techniques to the service of MT end users: Are uncertainty & model-internals based metrics a viable alternative to supervised word-level quality estimation? New paper w/ @gsarti.com @zouharvi.bsky.social @malvinanissim.bsky.social
072
Arianna Bisazza @arianna-bis.bsky.social · 31/05/2025
Large Reasoning Models are raising the bar for answer accuracy & transparency, but how does that work in multilingual settings? Can LRMs reason in your language, and what does that entail? New preprint led by @jiruiqi.bsky.social and @shan23chen.bsky.social!
150
Arianna Bisazza @arianna-bis.bsky.social · 30/05/2025
Following the success story of BabyBERTa, I & many other NLPers have turned to language acquisition for inspiration. In this new paper we show that using Child-Directed Language as training data is unfortunately *not* beneficial for syntax learning, at least not in the traditional LM training regime
1246
Arianna Bisazza @arianna-bis.bsky.social · 28/05/2025
Thinking LLM treats you just like an average user? Think again! @veraneplenbroek.bsky.social‘s analysis shows LLMs behave differently according to your gender, race & more. Implicit personalization is always at work & is strongly based on your conversation topics. Great collab w/ Raquel Fernández ⤵️
041
Arianna Bisazza @arianna-bis.bsky.social · 27/05/2025
Happy to be part of this collaboration on personalizing translation style in the literary domain. Besides classical multi-shot prompting, various steering techniques show promising results & bring new insights! See thread ⤵️ W/ @danielsc4.it @gsarti.com ElisabettaFersini, @malvinanissim.bsky.social
domain.is
030
Reposted by Arianna Bisazza
Vera Neplenbroek @veraneplenbroek.bsky.social · 21/05/2025
Excited to share that "Cross-Lingual Transfer of Debiasing and Detoxification in Multilingual LLMs: An Extensive Investigation" arxiv.org/abs/2412.14050 got accepted to ACL Findings! 🎉 #ACL2025 Big thanks to my supervisors Raquel Fernández and @arianna-bis.bsky.social for their guidance and support!
071
Arianna Bisazza @arianna-bis.bsky.social · 18/04/2025
RAG is a powerful way to improve LLMs' answering abilities across many languages. But how do LLMs deal with multilingual contexts? Do they answer consistently when the retrieved info is provided to them in different languages? Joint work w/ @jiruiqi.bsky.social & Raquel_Fernández See thread! ⤵️
062
Arianna Bisazza @arianna-bis.bsky.social · 08/04/2025
Modern LLMs "speak" hundreds of languages... but do they really? Multilinguality claims are often based on downstream tasks like QA & MT, while *formal* linguistic competence remains hard to gauge in lots of languages Meet MultiBLiMP! (joint work w/ @jumelet.bsky.social & @weissweiler.bsky.social)
2216
Arianna Bisazza @arianna-bis.bsky.social · 04/04/2025
Thanks to @haspelmath.bsky.social I just discovered this great collection of hypotheses extracted from evolutionary linguistics and typology papers, represented as a graph where linguistic properties are linked to others via different relations. correlation-machine.com/CHIELD/varia...
1131
Arianna Bisazza @arianna-bis.bsky.social · 04/04/2025
The PhD call is out! Apply by 24 April here: www.rug.nl/about-ug/wor...
rug.nl
Vacatures bij de RUG
095
Reposted by Arianna Bisazza
CoNLL 2026 @conll-conf.bsky.social · 12/03/2025
🚨 The CoNLL deadline is just 2 days away! 🚨 Submit your work by March 14th, 11:59 PM (AoE, UTC-12) Don't miss out! ⏳ 🔗 Submission links: conll.org #CoNLL2025 #NLP #CoNLL
conll.org
CoNLL 2025 | CoNLL
052
Arianna Bisazza @arianna-bis.bsky.social · 11/03/2025
📣 Soon-to-open PhD position @gronlp.bsky.social: Come work with Annemarie van Dooren, @yevgenm.bsky.social and myself on a new project bridging Computational Linguistics methods and Language Acquisition questions, with a focus on the learning of modal verbs.
162
Arianna Bisazza @arianna-bis.bsky.social · 01/03/2025
Excited to be traveling to Estonia for the 1st time to give a keynote @nodalida.bsky.social. I'll talk about using NNs to study language evolution & acquisition. A teaser: It won't be about LLMs 🙃 Also I've just moved from X, so this was my very first post... Pls help out by connecting with me!
4599