Sign in

Andrea Piergentili

@apierg.bsky.social
343 followers 286 following 25 posts

NLP researcher PhD at the University of Trento and @fbk-mt.bsky.social, working on gender-inclusive machine translation Applied Scientist Intern at Amazon (he/him) apierg.github.io #NLP #NLProc #MT

PostsRepliesMedia
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 31/03/2026
📢Thrilled to welcome our new postdoc, @wafaaissa.bsky.social! Looking forward to working together on trustworthy and responsible language technologies 🚀
072
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 02/03/2026
At our last MT seminar, we had @jacquelinerowe.bsky.social from @edinburghuni.bsky.social presenting her work "Beyond Benchmarks? Measuring Gender Bias in Multilingual LLMs"
083
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 25/02/2026
🎓 Today we attended the FBK PhD Day, and our amazing students presented their research lines: @apierg.bsky.social: gender-inclusive MT @linaconti.bsky.social: explainable speech translation @zhihangxie.bsky.social: long-form SpeechLLMs @uzumakidhairya.bsky.social: resource-efficient translation
087
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 18/02/2026
Great seminar last week by Alessandra Teresa Cignarella @cerveza-inglesa.bsky.social "6,000 Papers, One New Dataset, and Several Confused LLMs: A Story About Stereotype Detection in NLP" #NLP #StereotypeDetection
044
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 18/02/2026
Our #PickOfTheWeek by @bsavoldi.bsky.social: "Attention to Non-Adopters" by @kaitlynzhou.bsky.social, @gligoric.bsky.social, @myra.bsky.social, @mlam.bsky.social, @vyoma-raman.bsky.social, Boluwatife Aminu, Caeley Woo, Michael Brockman, @hannah-cha.bsky.social, @jurafsky.bsky.social (2025).
044
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 21/01/2026
Last week at the @fbk-mt.bsky.social seminars, we hosted Elizabeth Salesky from Google DeepMind, presenting her work on "Translation and Language Modeling with Pixels" #NLProc #tokenization #MT
0137
Andrea Piergentili @apierg.bsky.social · 07/01/2026
Impressive work by the Glitter team: a new human-made benchmark for German gender-inclusive MT with long passages and multiple inclusive approaches + experiments showing that MT systems and LLMs still fall short in generating inclusive outputs. aclanthology.org/2025.finding... ✨
aclanthology.org
Glitter: A Multi-Sentence, Multi-Reference Benchmark for Gender-Fair German Machine Translation
A Pranav, Janiça Hackenbuchner, Giuseppe Attanasio, Manuel Lardelli, Anne Lauscher. Findings of the Association for Computational Linguistics: EMNLP 2025. 2025.
021
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 09/12/2025
It was great having Hosein Mohebbi (@hmohebbi.bsky.social) speak about interpretability for speech Transformers at our #MTSeminars! Thanks for the insights 🎤 #NLP #XAI
085
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 05/12/2025
🚀 JOB ALERT 3: The FBK's MT Unit is hiring! Join us as a Researcher in Responsible & Trustworthy NLP and advance ethical, fair, and transparent language technologies. If you care about building safe and accountable AI systems, you can apply here: 👉 jobs.fbk.eu/Annunci/Offe...
jobs.fbk.eu
Jobs | Science and Technology Hub - Trento | A Researcher in Responsible and Trustworthy NLP
066
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 28/10/2025
@bsavoldi.bsky.social presenting our new multilingual benchmark for evaluating LLMs on gender-neutral translation. Catch our paper at #EMNLP2025 ℹ️ arxiv.org/pdf/2501.09409 #lt2025fbk
041
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 21/10/2025
🚀 Join us for the LT@FBK day 2025! Discover cutting-edge research and highlights in speech and language technologies from Fondazione Bruno Kessler (FBK) 📅 October 28, 2025 📍FBK, Trento ℹ️ lt-highlights.fbk.eu
lt-highlights.fbk.eu
LT Highlights @ FBK 2025
031
Reposted by Andrea Piergentili
GITT 2026 @gitt-workshop.bsky.social · 23/06/2025
Last but definitely not least: @bsavoldi.bsky.social presenting joint work with @apierg.bsky.social @matteo-negri.bsky.social @luisabentivogli.bsky.social on scalable gender neutral translation evaluation using LLM-as-a-judge at #GITT2025
6113
Andrea Piergentili @apierg.bsky.social · 04/06/2025
Super interesting paper by Subramonian et al: "Agree to Disagree? A Meta-Evaluation of LLM Misgendering" arxiv.org/abs/2504.17075 Turns out, misgendering is messier than just pronouns. I'd love to see this analysis extended to grammatical gender languages! #LLM #AI #ethics @fbk-mt.bsky.social
arxiv.org
Agree to Disagree? A Meta-Evaluation of LLM Misgendering
Numerous methods have been proposed to measure LLM misgendering, including probability-based evaluations (e.g., automatically with templatic sentences) and generation-based evaluations (e.g., with automatic heuristics or human validation). However, it has gone unexamined whether these evaluation methods have convergent validity, that is, whether their results align. Therefore, we conduct a systematic meta-evaluation of these methods across three existing datasets for LLM misgendering. We propose a method to transform each dataset to enable parallel probability- and generation-based evaluation. Then, by automatically evaluating a suite of 6 models from 3 families, we find that these methods can disagree with each other at the instance, dataset, and model levels, conflicting on 20.2% of evaluation instances. Finally, with a human evaluation of 2400 LLM generations, we show that misgendering behaviour is complex and goes far beyond pronouns, which automatic evaluations are not currently designed to capture, suggesting essential disagreement with human evaluations. Based on our findings, we provide recommendations for future evaluations of LLM misgendering. Our results are also more widely relevant, as they call into question broader methodological conventions in LLM evaluation, which often assume that different evaluation methods agree.
050
Reposted by Andrea Piergentili
Beatrice Savoldi @bsavoldi.bsky.social · 03/06/2025
🔍 Stiamo studiando come l'AI viene usata in Italia e per farlo abbiamo costruito un sondaggio! 👉 bit.ly/sondaggio_ai... (è anonimo, richiede ~10 minuti, e se partecipi o lo fai girare ci aiuti un sacco🙏) Ci interessa anche raggiungere persone che non si occupano e non sono esperte di AI!
bit.ly
Qualtrics Survey | Qualtrics Experience Management
The most powerful, simple and trusted way to gather experience data. Start your journey to experience management and try a free account today.
11618
Reposted by Andrea Piergentili
sarapapi.bsky.social @sarapapi.bsky.social · 30/05/2025
🚀 New tech report out! Meet FAMA, our open-science speech foundation model family for both ASR and ST in 🇬🇧 English and 🇮🇹 Italian. The models are live and ready to try on @hf.co: 🔗 huggingface.co/collections/... 📄 Preprint: arxiv.org/abs/2505.22759 #ASR #ST #OpenScience #MultilingualAI
huggingface.co
FAMA - a FBK-MT Collection
The First Large-Scale Open-Science Speech Foundation Model for English and Italian
073
Reposted by Andrea Piergentili
Joke Daems @jdaems.bsky.social · 09/05/2025
👀 Wanted: #Italian or #Dutch native speakers to take a survey on audiovisual translation for a master thesis student: watch a short video, answer some questions, help academic research 😎 ⏩ Sharing = nice! ❤️ NL link: ugent.qualtrics.com/jfe/form/SV_... IT link: ugent.qualtrics.com/jfe/form/SV_...
media.tenor.com
a woman is standing in front of a bookshelf in a bookstore and talking about research .
ALT: a woman is standing in front of a bookshelf in a bookstore and talking about research .
168
Reposted by Andrea Piergentili
GITT 2026 @gitt-workshop.bsky.social · 29/04/2025
💭Dreaming of attending #GITT2025 but need a little extra 💸 boost? 📣 Bursary applications to support participation are now open at tinyurl.com/gitt25 📆 Deadline May 9th 🙏Thanks to our incredible sponsors DCA at Tilburg University tinyurl.com/tudca25 and FLW at Ghent University www.ugent.be/lw/en
media.tenor.com
a man in a suit is making a funny face with the words dreams are expensive behind him
ALT: a man in a suit is making a funny face with the words dreams are expensive behind him
177
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 22/04/2025
📢 Come and join our group! We offer a fully funded 3-year PhD position: 📔 Automatic translation with large multimodal models: iecs.unitn.it/education/ad... 📍Full details for application: iecs.unitn.it/education/ad... 📅 Deadline May 12, 2025 #NLProc #FBK
iecs.unitn.it
Reserved topic scholarships | Doctoral Program - Information Engineering and Computer Science
189
Andrea Piergentili @apierg.bsky.social · 17/04/2025
Happy to announce that our paper 'An LLM-as-a-judge Approach for Scalable Gender-Neutral Translation Evaluation' was accepted at @gitt-workshop.bsky.social ! 🙌 Check it out: arxiv.org/abs/2504.11934 🔥 Co-authors (🫶🏻): @bsavoldi.bsky.social, @matteo-negri.bsky.social, @luisabentivogli.bsky.social
arxiv.org
An LLM-as-a-judge Approach for Scalable Gender-Neutral Translation Evaluation
Gender-neutral translation (GNT) aims to avoid expressing the gender of human referents when the source text lacks explicit cues about the gender of those referents. Evaluating GNT automatically is pa...
0113
Andrea Piergentili @apierg.bsky.social · 19/03/2025
Brilliant and necessary work by Pombal et al. about metric interference in MT system development and evaluation: arxiv.org/abs/2503.08327 Are we developing better systems or are we just gaming the metrics? And how do we address this? Super (m)interesting! 👀
arxiv.org
Adding Chocolate to Mint: Mitigating Metric Interference in Machine Translation
As automatic metrics become increasingly stronger and widely adopted, the risk of unintentionally "gaming the metric" during model development rises. This issue is caused by metric interference (Mint)...
0101
Reposted by Andrea Piergentili
GITT 2026 @gitt-workshop.bsky.social · 28/02/2025
While we look forward to a sunny Geneva, why wait to join the conversation? We’ve created a starter pack for our #GITT2025 friends! 🕵️ Follow researchers working on gender bias in MT 💬 Stay up to date and dive into the discussion! All info at sites.google.com/tilburgunive...
12116
Reposted by Andrea Piergentili
Jeremy Faust, MD @jeremyfaust.bsky.social · 01/02/2025
BREAKING NEWS: CDC orders mass retraction and revision of submitted research across all science and medicine journals. Banned terms must be scrubbed. Goes beyond MMWR +other CDC pubs. Applies to research already submitted to top medical journals. Take a look. open.substack.com/pub/insideme...
open.substack.com
BREAKING NEWS: CDC orders mass retraction and revision of submitted research across all science and medicine journals. Banned terms must be scrubbed.
Any unpublished manuscript mentioning certain topics, including gender and "LGBT," must be pulled or revised.
56576214262
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 16/01/2025
🙌 All members of our group are now on Bluesky! 🙌 You can find all of us in this starter pack 👇
065
Andrea Piergentili @apierg.bsky.social · 27/12/2024
With 2024 wrapping up, and given how little I’ve posted here (or anywhere, really), I thought I’d share a quick recap of my year and finally make some ✨content✨
media.tenor.com
but look i made you some is written in white on a black background
ALT: but look i made you some is written in white on a black background
130
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 06/12/2024
Our @apierg.bsky.social presenting our #calamita challenges at #CLiCit2024: machine translation and gender-fair generation. Poster session upcoming, see you there! For more details: 👉 MagneT: clic2024.ilc.cnr.it/wp-content/u... 👉 GFG: clic2024.ilc.cnr.it/wp-content/u...
092
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 05/12/2024
Our very own @dennisfucci.bsky.social presenting the challenges of Explainability for Speech Models at #CLiCit2024. If you’re interested, check out the paper 👉 clic2024.ilc.cnr.it/wp-content/u... #NLProc
091
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 05/12/2024
Today @luisabentivogli.bsky.social, Dennis Fucci, and @apierg.bsky.social presented a research communication about gender-neutral translation in the morning poster session #CLiCit2024 #NLProc
1204
Reposted by Andrea Piergentili
Daniel Russo @daniel-russo.bsky.social · 05/12/2024
If you are in Pisa at #CLiCit2024 don't miss the presentation of our last work today at 12 🔥
0122
Reposted by Andrea Piergentili
Debora Nozza @deboranozza.bsky.social · 04/12/2024
I’ve created an Italian #NLProc Researcher Starter Pack 🇮🇹 DM me to join if you're not in yet! go.bsky.app/LHbWLHp
0245
Andrea Piergentili @apierg.bsky.social · 04/12/2024
Hey, if you're in Pisa and interested to connect with other people at #CLiCit2024 check out this starter pack. Raise a hand 👋 and I will add you to this list!
3104
Reposted by Andrea Piergentili
MT Group at FBK @fbk-mt.bsky.social · 04/12/2024
We are happy to announce that our PhD students Dennis Fucci and @apierg.bsky.social along with our head of unit @luisabentivogli.bsky.social, will attend #CLiCit2024 in Pisa! Meet them at the poster presentations during the main conference and the #CALAMITA event!
mt.fbk.eu
Dennis Fucci, Andrea Piergentili, and Luisa Bentivogli at CLiC-it 2024 | Machine Translation Unit
084
Andrea Piergentili @apierg.bsky.social · 03/12/2024
Another interesting fact about Bluesky's Italian localization: the website greets new users in an inclusive way by using * instead of gendered morphemes 👀 ...however, it also refers to the user with a masculine term ('iscritto') right after 🤦🏻‍♂️ Thx @sarapapi.bsky.social for the screenshot!
182
Reposted by Andrea Piergentili
Marc Lanctot @sharky6000.bsky.social · 02/12/2024
Let's cycle through the memes for this one until it stops... 😇😅🙏
1386
Reposted by Andrea Piergentili
Joe Stacey @joestacey.bsky.social · 24/11/2024
Okay genius idea to improve quality of #nlp #arr reviews. Literally give gold stars to the best reviewers, visible on open review next to your anonymously ID during review process. Here’s why it would work, and why would you should RT this fab idea:
3275
Andrea Piergentili @apierg.bsky.social · 25/11/2024
Turns out this is how programmers do introspection
110
Andrea Piergentili @apierg.bsky.social · 18/11/2024
We may need more machine translation people on Bluesky 🤔 it: 'impara' = en: 'learn', imperative – the correct translation would be 'insegna' (en: 'teach')
060
Andrea Piergentili @apierg.bsky.social · 30/10/2024
Using LLMs as evaluators looks like a very interesting and promising direction, enabling simpler automatic post-editing pipelines. For those interested in fine-grained MT evaluation and APE, I recommend checking out this paper by Lu et al. (2024): arxiv.org/abs/2409.14335 #MT #postediting #NLP #LLM
arxiv.org
MQM-APE: Toward High-Quality Error Annotation Predictors with Automatic Post-Editing in LLM Translation Evaluators
Large Language Models (LLMs) have shown significant potential as judges for Machine Translation (MT) quality assessment, providing both scores and fine-grained feedback. Although approaches such as GE...
030