Sign in

Yu Lu Liu

@liuyulu.bsky.social
1.1K followers 893 following 34 posts

PhD student at Johns Hopkins University Alumni from McGill University & MILA Working on NLP Evaluation, Responsible AI, Human-AI interaction she/her 🇨🇦

PostsRepliesMedia
Reposted by Yu Lu Liu
Alexandra Olteanu @aolteanu.bsky.social · 30/09/2025
This was accepted to #NeurIPS 🎉🎊 TL;DR Impoverished notions of rigor can have a formative impact on AI work. We argue for a broader conception of what rigorous work should entail & go beyond methodological issues to include epistemic, normative, conceptual, reporting & interpretative considerations
1268
Reposted by Yu Lu Liu
Ziang Xiao @ziangxiao.bsky.social · 25/04/2025
We are excited to kick off the 2nd HEAL workshop tomorrow at #CHI2025. Dr. Su Lin Blodgett and Dr. Gagan Bansal from MSR will be our keynote speakers! Welcome new and old friends! See you at G221! All accepted papers: tinyurl.com/bdfpjcr4
Dr. Su Lin Blodgett and Dr. Gagan Bansal will be the keynote speakers of the 2nd HEAL workshop @CHI25
062
Reposted by Yu Lu Liu
JHU CLSP @jhuclsp.bsky.social · 07/03/2025
Bringing together our incredible current and admitted students—future leaders, innovators, and changemakers!
072
Yu Lu Liu @liuyulu.bsky.social · 13/02/2025
📣 DEADLINE EXTENSION 📣 By popular request, HEAL workshop submission deadline is extended to Feb 24 AOE! Reminder that we welcome a wide range of submissions: position papers, lit reviews, encore of published work, etc. Looking forward to your submissions!
040
Reposted by Yu Lu Liu
Nikhil Sharma ༗ @nikhilsksharma.bsky.social · 31/01/2025
Thrilled that our paper Faux Polyglot has been accepted to #NAACL2025 main! 🚀 We show that multilingual RAG creates language-specific information cocoons and amplifies perspectives and facts in the dominant language, especially when handling knowledge conflicts. 📜 arxiv.org/abs/2407.05502
arxiv.org
Faux Polyglot: A Study on Information Disparity in Multilingual Large Language Models
With Retrieval Augmented Generation (RAG), Large Language Models (LLMs) are playing a pivotal role in information search and are being adopted globally. Although the multilingual capability of LLMs of...
1103
Yu Lu Liu @liuyulu.bsky.social · 22/01/2025
The submission deadline is in less than a month! We welcome encore submissions, so consider submitting your work regardless of whether it's been accepted or not #chi2025 😉
081
Yu Lu Liu @liuyulu.bsky.social · 16/12/2024
Human-centered Evalulation and Auditing of Language models (HEAL) workshop is back for #CHI2025, with this year's special theme: “Mind the Context”! Come join us on this bridge between #HCI and #NLProc! Workshop submission deadline: Feb 17 AoE More info at heal-workshop.github.io.
The image includes a shortened call for participation that reads: 
"We welcome participants who work on topics related to supporting human-centered evaluation and auditing of language models. Topics of interest include, but not limited to:
- Empirical understanding of stakeholders' needs and goals of LLM evaluation and auditing
- Human-centered evaluation and auditing methods for LLMs
- Tools, processes, and guidelines for LLM evaluation and auditing
- Discussion of regulatory measures and public policies for LLM auditing
- Ethics in LLM evaluation and auditing

Special Theme: Mind the Context. We invite authors to engage with specific contexts in LLM evaluation and auditing. This theme could involve various topics: the usage contexts of LLMs, the context of the evaluation/auditing itself, and more! The term ''context'' is purposefully left open for interpretation!

The image also includes pictures of workshop organizers, who are: Yu Lu Liu, Wesley Hanwen Deng, Michelle S. Lam, Motahhare Eslami, Juho Kim, Q. Vera Liao, Wei Xu, Jekaterina Novikova, and Ziang Xiao.
24410
Reposted by Yu Lu Liu
Hanna Wallach @hannawallach.bsky.social · 02/12/2024
Super excited to announce that @msftresearch.bsky.social's FATE group, Sociotechnical Alignment Center, and friends have several workshop papers at next week's @neuripsconf.bsky.social. A short thread about (some of) these papers below... #NeurIPS2024
15813
Reposted by Yu Lu Liu
Alexandra Olteanu @aolteanu.bsky.social · 05/12/2024
📣 📣 Interested in an internship on human-centred AI, human agency, AI evaluation & the impacts of AI systems? Our team/FATE MLT (Su Lin Blodgett, @qveraliao.bsky.social & I) is looking for a few summer interns 🎉 Apply by Jan 10 for full consideration: jobs.careers.microsoft.com/global/en/jo...
02210
Reposted by Yu Lu Liu
Eryk Salvaggio @eryk.bsky.social · 02/12/2024
I am collecting examples of the most thoughtful writing about generative AI published in 2024. What’s yours? They can be insightful for commentary, smart critique, or just because it shifted the conversation. I’ll post some of mine below as I go through them. #criticalAI
97350115
Reposted by Yu Lu Liu
Alexandra Olteanu @aolteanu.bsky.social · 30/11/2024
Created a small starter pack including folks whose work I believe contributes to more rigorous and grounded AI research -- I'll grow this slowly and likely move it to a list at some point :) go.bsky.app/P86UbQw
1125
Reposted by Yu Lu Liu
Dr. Casey Fiesler @cfiesler.bsky.social · 27/11/2024
Hi, so I've spent the past almost-decade studying research uses of public social media data, like e.g. ML researchers using content from Twitter, Reddit, and Mastodon. Anyway, buckle up this is about to be a VERY long thread with lots of thoughts and links to papers. 🧵
59959450
Reposted by Yu Lu Liu
Vera Liao @qveraliao.bsky.social · 26/11/2024
Had a lot of fun teaching a tutorial on Human-Centered Evaluation of Language Technologies at #EMNLP2024, w/ @ziangxiao.bsky.social, Su Lin Blodgett, and Jackie Cheung We just posted the slides on our tutorial website: human-centered-eval.github.io
human-centered-eval.github.io
Human-Centered Eval@EMNLP24
0133
Reposted by Yu Lu Liu
Anka Reuel ➡️ NeurIPS @ankareuel.bsky.social · 25/11/2024
🚨 NeurIPS 2024 Spotlight Did you know we lack standards for AI benchmarks, despite their role in tracking progress, comparing models, and shaping policy? 🤯 Enter BetterBench–our framework with 46 criteria to assess benchmark quality: betterbench.stanford.edu 1/x
413925
Reposted by Yu Lu Liu
McGill NLP @mcgill-nlp.bsky.social · 24/11/2024
It turns out we had even more papers at EMNLP! Let's complete the list with three more🧵
1144
Reposted by Yu Lu Liu
McGill NLP @mcgill-nlp.bsky.social · 23/11/2024
Our lab members recently presented 3 papers at @emnlpmeeting.bsky.social in Miami ☀️ 📜 From interpretability to bias/fairness and cultural understanding -> 🧵
1196
Reposted by Yu Lu Liu
Alexandra Olteanu @aolteanu.bsky.social · 22/11/2024
my first post, now that I am here with my 500+ closest friends 🙂 -- here is a tiny owl 🦉 I met some weeks back in the big apple 🍎 (picture by @sbucur.bsky.social)
tiny owl
0111
Reposted by Yu Lu Liu
Benno Krojer @bennokrojer.bsky.social · 22/11/2024
McGill NLP just landed on this blue planet bsky.app/profile/mcgi...
bsky.app
082
Yu Lu Liu @liuyulu.bsky.social · 23/11/2024
The starter pack just surpassed 1/3 of its capacity! Don't be shy to reach out to me if you are a researcher in this area, or if you have suggestions. Thank you 🥰
060
Reposted by Yu Lu Liu
Caleb Moses @mathematiguy.bsky.social · 20/11/2024
I didn’t expect to wind up in the news over this but in hindsight, I guess it makes sense lol. This is the first time I’ve been in the Herald since high school 😂.
711217
Reposted by Yu Lu Liu
Miriam Posner @miriamposner.com · 22/11/2024
“We argue that societal impacts [of GenAI] should be conceptualised as application- and context-specific, incommensurable, and shaped by questions of social power.” By @glenberman.bsky.social et al. arxiv.org/abs/2410.22985
arxiv.org
Troubling Taxonomies in GenAI Evaluation
To evaluate the societal impacts of GenAI requires a model of how social harms emerge from interactions between GenAI, people, and societal structures. Yet a model is rarely explicitly defined in soci...
0174
Yu Lu Liu @liuyulu.bsky.social · 21/11/2024
I’m putting together a starter pack for researchers working on human-centered AI evaluation. Reply or DM me if you’d like to be added, or if you have suggestions! Thank you! (It looks NLP-centric at the moment, but that’s due to the current limits of my own knowledge 🙈) go.bsky.app/G3w9LpE
153510
Reposted by Yu Lu Liu
Kate Sanders @kesnet50.bsky.social · 19/11/2024
Putting together a JHU Center for Language and Speech Processing starter pack! Please reply or DM me if you're doing research at CLSP and would like to be added - I'm still trying to find out which of us are on here so far. go.bsky.app/JtWKca2
go.bsky.app
CLSP
Join the conversation
22210
Yu Lu Liu @liuyulu.bsky.social · 16/11/2024
If you are at #EMNLP2024, don't miss the "Human-Centered Evaluation of Language Technologies" tutorial on Saturday 14:00-17:30! My awesome mentors will be teaching about this awesome topic ❤️
Promotional graphic of "Human-Centered Evaluation of Language Technologies" tutorial at EMNLP 2024, happening on Saturday November 16, 14:00-17:30. The instructors of this tutorial are Su Lin Blodgett, Jackie Chi Kit Cheung, Q. Vera Liao, and Ziang Xiao. Their profile pictures are shown at the bottom half of the graphic.
130
Reposted by Yu Lu Liu
Anna Rogers @annarogers.bsky.social · 15/11/2024
Shameless plug: a post on 'writing as thinking' that I wrote for my own students hackingsemantics.xyz/2024/writing/
hackingsemantics.xyz
On AI-assisted writing in graduate school
What is the proper role of AI ‘assistance’ in graduate student writing? It depends on what you mean by ‘graduate’.
34515
Reposted by Yu Lu Liu
ACL 2027 @aclmeeting.bsky.social · 14/11/2024
Hello, Computational linguistics/NLP world in Bluesky! We're creating the same accounts on other social media platforms in Bluesky! #NLProc
413131
Yu Lu Liu @liuyulu.bsky.social · 12/11/2024
🚚 Moving threads about my #nlp papers from Twitter to here 🚚 Do NLP benchmark measurements provide meaningful insights about the evaluated models? To help practitioners answer these questions, we introduce ECBD - a conceptual framework that formalizes the process of benchmark design 🧵.
Screenshot of the first page of the paper titled "ECBD: Evidence-Centered Benchmark Design for NLP", authored by Yu Lu Liu, Su Lin Blodgett, Jackie Chi Kit Cheung, Q. Vera Liao, Alexandra Olteanu, and Ziang Xiao.
1101
Yu Lu Liu @liuyulu.bsky.social · 12/11/2024
🚚 Moving threads about my #nlp papers from Twitter to here 🚚 How and when, and with which issues, does the text summarization community engage with responsible AI? 🤔 In this #EMNLP2023 paper, we examine reporting and research practices across 300 summarization papers published between 2020-2022 🧵
Screenshot of the first page of the paper titled "Responsible AI Considerations in Text Summarization Research: A Review of Current Practices", authored by Yu Lu Liu, Meng Cao, Su Lin Blodgett, Jackie Chi Kit Cheung, Alexandra Olteanu and Adam Trischler.
151