Sign in

Carolin Holtermann

@carolin-holtermann.bsky.social
82 followers 93 following 8 posts

Ph.D. Candidate in NLP focusing at @ds-hamburg.bsky.social on Multilinguality and Multiculturality #NLProc

PostsRepliesMedia
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
A huge thank you to all the amazing authors of this work!! @minhducbui.bsky.social @kaitlynzhou.bsky.social @valentinhofmann.bsky.social @a-lauscher.bsky.social
031
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
Paper link: arxiv.org/abs/2603.22260 Website: tai-hamburg.github.io/papers/diver...
arxiv.org
Greater accessibility can amplify discrimination in generative AI
Hundreds of millions of people rely on large language models (LLMs) for education, work, and even healthcare. Yet these models are known to reproduce and amplify social biases present in their trainin...
000
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
The result is a self-reinforcing gap: systems optimize for frequent users, while growing more distant from those they were supposed to reach. Fairness and accessibility have to be tackled together.
010
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
🛠️ Study 3: Mitigation exists. Discrimination tracks continuously with vocal pitch. Systematically lowering pitch reduced bias to statistically insignificant levels across all tested models.
100
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
⚠️ Study 2: Real consequences. Voice AI is meant to lower barriers — but non-users are 5× more likely to disengage than daily users when they learn these systems infer attributes from their voice.
100
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
🔍 Study 1: Do models treat voices differently? Yes. Across audio-enabled LLMs, voice inputs drove measurable gender-stereotyped outputs. Higher gender detection accuracy = stronger bias.
100
Carolin Holtermann @carolin-holtermann.bsky.social · 24/03/2026
🚨 New paper alert: A new generation of LLMs can now process speech natively. This could expand access for millions excluded by text interfaces, but our research shows a cost: demographic cues in speaker voice can trigger stereotypical model responses. 🎙️⚖️ Paper: arxiv.org/abs/2603.22260
3175
Carolin Holtermann @carolin-holtermann.bsky.social · 06/03/2026
📣 CFP: Eval4SD @ KONVENS 2026, Hamburg (Sept 14–17) 1st workshop on evaluating LLMs for specialized domains. We invite work on benchmarking, replication, and evaluation methodology across fields such as medicine, law, science, finance, and more. Deadline: Jul 3, 2026, 23:59 CEST eval4sd.github.io
eval4sd.github.io
First Workshop on Evaluating LLMs for Specialized Domains (Eval4SD)
051
Reposted by Carolin Holtermann
Paul Röttger @paul-rottger.bsky.social · 28/07/2025
Finally, I will be with @carolin-holtermann.bsky.social and @a-lauscher.bsky.social to present our work on evaluating geotemporal reasoning ability in LLMs. This will be in the Wednesday 1100 poster session: aclanthology.org/2025.acl-lon...
aclanthology.org
Around the World in 24 Hours: Probing LLM Knowledge of Time and Place
Carolin Holtermann, Paul Röttger, Anne Lauscher. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2025.
174
Reposted by Carolin Holtermann
Trustworthy AI Lab @tai-lab.bsky.social · 23/06/2025
🧵Excited to share that Data Science Lab Hamburg will be presenting 6 papers at #ACL2025 in Vienna! 🇦🇹 Our team has been working on cutting-edge research in multilingual AI, peer review analysis, and combating online harms. Here's what we're bringing to the conference 👇
Main Conference Papers:

Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model (authors: G Geigle, F Schneider, C Holtermann, C Biemann, R Timofte, A Lauscher and G Glavas)
LazyReview: A Dataset for Uncovering Lazy Thinking in NLP Peer Reviews (authors: S Purkayastha, Z Li, A Lauscher, L Qu and I Gurevych)
Around the World in 24 Hours: Probing LLM Knowledge of Time and Place (authors: C Holtermann, P Röttger and A Lauscher)

Findings Papers:

GIMMICK: Globally Inclusive Multimodal Multitask Cultural Knowledge Benchmarking (authors: F Schneider, C Holtermann, C Biemann and A Lauscher)
FineCite: A Novel Approach For Fine-Grained Citation Context Analysis (authors: L Jantsch, D Koh, S Yoon, J Lee, A Lauscher and Y Suh)

Workshop on Online Abuse and Harms (WOAH):

Debunking with Dialogue? Exploring AI-Generated Counterspeech to Challenge Conspiracy Theories (authors: M Lisker, C Gottschalk, H Mihaljević)
272
Reposted by Carolin Holtermann
INTERPLAY Workshop@COLM '25 @interplay-workshop.bsky.social · 16/05/2025
🚨🚨 Studying the INTERPLAY of LMs' internals and behavior? Join our @colmweb.org workshop on comprehensivly evaluating LMs. Deadline: June 23rd CfP: shorturl.at/sBomu Page: shorturl.at/FT3fX We're excited to see your insights and methods!! See you in Montréal 🇨🇦 #nlproc #interpretability
Call for Papers, Interplay Workshop at COLM: June 23rd - submissions due. July 24th - acceptance notification. October 10th - workshop day.
0129