Sign in

Yuxuan Li

@yuxuanli1225.bsky.social
62 followers 90 following 86 posts

Second-year PhD student @CMU HCII. Previously undergrad @Tsinghua CS and research intern @Berkeley AI Research @Berkeley InfoSchool. yuxuanli.com

PostsRepliesMedia
Yuxuan Li @yuxuanli1225.bsky.social · 30/04/2026
Excited to share that our paper has been accepted to 𝗜𝗖𝗠𝗟 𝟮𝟬𝟮𝟲! 🎉 Multi-agent is everywhere today. But put frontier LLMs in a room where each holds a different piece of the puzzle, and they fail 70% of the time. Here's why: 📄 Paper: arxiv.org/abs/2505.11556
arxiv.org
Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs
Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet systematically evaluating this capability has remained challen...
131
Yuxuan Li @yuxuanli1225.bsky.social · 30/04/2026
Excited to share that our paper has been accepted to ICML 2026! 🎉 📄 Paper: arxiv.org/abs/2505.11556 📊 Benchmark: huggingface.co/datasets/Yux... #ICML
arxiv.org
Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs
Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet systematically evaluating this capability has remained challen...
010
Yuxuan Li @yuxuanli1225.bsky.social · 19/04/2026
The first PoliSim workshop at #CHI2026 was a huge success! We saw incredible interest from researchers across diverse domains, with attendance nearly filling one of the largest rooms at the venue. 1/n
140
Yuxuan Li @yuxuanli1225.bsky.social · 11/04/2026
Excited to be heading to Barcelona for #CHI2026 to host our workshop PoliSim: LLM Agent Simulation for Policy! This year, we’ve seen incredible interest from researchers across HCI, NLP, CSS, and Policy. We accepted 25 outstanding papers, with 5 selected as Best Paper nominees.
282
Yuxuan Li @yuxuanli1225.bsky.social · 25/02/2026
Can LLMs really serve as "crash dummies" for security & privacy testing? We put this assumption to the test. 🚨New preprint 🚨: "How Well Can LLM Agents Simulate End-User Security and Privacy Attitudes and Behaviors?" 👇 THREAD 👇 [Link to paper: arxiv.org/abs/2602.184... [1/n]
arxiv.org
142
Yuxuan Li @yuxuanli1225.bsky.social · 18/02/2026
Excited to share that I’m joining @msftresearch.bsky.social AI Frontiers this summer as a Research Intern, working on LLM Agents, and reporting to @zacharyhuang.bsky.social. I’ll be based in Redmond — would love to connect with folks around Seattle!
000
Reposted by Yuxuan Li
Yuxuan Li @yuxuanli1225.bsky.social · 13/02/2026
🚨New paper🚨 We spent a year working with emergency preparedness policymakers to answer a simple question: can LLM agent simulations actually help real institutions make better decisions? The answer is yes—but perhaps not how you'd expect. 👇 THREAD 👇 [Link to paper: arxiv.org/abs/2509.218... [1/n]
112
Yuxuan Li @yuxuanli1225.bsky.social · 13/02/2026
🚨New paper🚨 We spent a year working with emergency preparedness policymakers to answer a simple question: can LLM agent simulations actually help real institutions make better decisions? The answer is yes—but perhaps not how you'd expect. 👇 THREAD 👇 [Link to paper: arxiv.org/abs/2509.218... [1/n]
112
Yuxuan Li @yuxuanli1225.bsky.social · 11/02/2026
🚨 Deadline Extended! Due to popular demand, the submission deadline for the PoliSim workshop at CHI’26 has been extended to February 20th! 🎉 We look forward to your submissions! 👉 polisim.net
polisim.net
PoliSim@CHI 2026
000
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 09/02/2026
📣 New at #CHI2026 Developing a new AI product? How would you figure out what are the privacy risks? Privy help non-privacy expert practitioners create high quality privacy impact assessments for early-stage AI products. Led by @hankhplee.bsky.social Paper: www.sauvik.me/papers/69/s...
151
Yuxuan Li @yuxuanli1225.bsky.social · 05/02/2026
Don’t forget! The submission deadline for our CHI 2026 workshop "PoliSim: LLM Agent Simulation for Policy" is only one week away (Feb. 13). We’d love to see your work!
010
Yuxuan Li @yuxuanli1225.bsky.social · 18/12/2025
🚨 New CHI 2026 Workshop 🚨 PoliSim@CHI 2026: LLM Agent Simulation for Policy
173
Reposted by Yuxuan Li
Qiaosi (Chelsea) Wang @qiaosiwang.bsky.social · 11/11/2025
🚀💫 I’m on the job market for academic (tenure-track) and industry research positions! 👋I am a Postdoc Fellow at @hcii.cmu.edu working at the intersection of human-AI interaction, cognitive science, responsible AI, design, and social computing. I earned my PhD from @gtresearch.bsky.social in 2024.
195
Yuxuan Li @yuxuanli1225.bsky.social · 15/09/2025
Happy to share that our paper has been selected for an ORAL presentation at #EMNLP2025 main conference! I’ll be presenting in person in Suzhou. See you there!
040
Yuxuan Li @yuxuanli1225.bsky.social · 20/08/2025
Happy to share that our paper has been accepted to #EMNLP2025 main conference!
030
Yuxuan Li @yuxuanli1225.bsky.social · 22/05/2025
🚨 New Preprint! 🚨 Smarter LLMs are more selfish. We show reasoning-enhanced models significantly prefer greed over cooperation. The more LLMs reason, the worse they cooperate. 👇 THREAD 👇 [Link to paper: arxiv.org/abs/2502.177... [1/n]
131
Yuxuan Li @yuxuanli1225.bsky.social · 20/05/2025
🚨 New Paper! 🚨 “Groups that communicate diverse information make wiser decisions.” Right? Not always—for LLMs, just like humans. We bring this scrutiny to AI by introducing the Hidden Profile paradigm to assess how multi-agent LLMs actually reason together. arxiv.org/abs/2505.115... [1/n]
152
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 13/05/2025
😬In sims, 1/100 Asian-coded agents will join a protest if told they shouldn't. All 100 Black-coded agents will join anyway. New #Facct2025 paper led by @yuxuanli1225.bsky.social Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models sauvikdas.com/papers/64/se...
034
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 11/04/2025
We just found out that @yuxuanli1225.bsky.social 's paper was accepted to #FAccT2025 A true testament to his hard work and determination!
041
Reposted by Yuxuan Li
Hao-Ping (Hank) Lee @hankhplee.bsky.social · 28/03/2025
How does using GenAI tools reshape knowledge workers’ critical thinking? Our #CHI2025 paper studied 319 knowledge workers to dive into this question. w/@advaitsarkar.bsky.social @levlevlev.bsky.social Ian, Sean, Richard, Nick @msftresearch.bsky.social @hcii.cmu.edu www.microsoft.com/en-us/resear...
1146
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 26/02/2025
🔒 Our new #USEC2025 paper shows that by automating permission decisions based on prior decisions, we risk NORMALIZING privacy discomfort rather than reducing it. "Modeling End-User Affective Discomfort With Mobile App Permissions" Paper: sauvikdas.com/papers/60/se...
192
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 24/02/2025
🚨 In our new #USEC2025 paper, we introduce a secure, physically-intuitive RFID tag that users actually trust. On-demand RFID: Improving Privacy, Security, and User Trust in RFID Activation through Physically-Intuitive Design 👉 Paper: sauvikdas.com/papers/62/se... #Privacy #AcademicSky #Security
1113
Reposted by Yuxuan Li
Qing Xiao @qingxiaohci.bsky.social · 14/02/2025
🚀 Our paper has been accepted to #CHI2025! 🎉 arxiv.org/abs/2409.12000 As AI technologists and data workers increasingly enter the news industry, cross-functional collaboration with journalists is becoming essential. But how do these collaborations unfold in practice?
arxiv.org
"It Might be Technically Impressive, But It's Practically Useless to us": Motivations, Practices, Challenges, and Opportunities for Cross-Functional Collaboration around AI within the News Industry
Recently, an increasing number of news organizations have integrated artificial intelligence (AI) into their workflows, leading to a further influx of AI technologists and data workers into the news i...
351
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 09/02/2025
@hankhplee.bsky.social's MSR internship work is getting quite a bit of attention! Following a year where he was recognized with both a CHI Best Paper and a USENIX Security Distinguished Paper, it looks like I will one day be best known for being Hank's advisor :p
0161
Reposted by Yuxuan Li
Hao-Ping (Hank) Lee @hankhplee.bsky.social · 07/02/2025
🚨Trapped by infinite scroll & autoplay on social media? Our TOCHI paper introduces a system that reduces distraction by 4x by suppressing these and other dark patterns. Purpose Mode: Reducing Distraction through Toggling ACDPs on Social Media Web Sites 📝 hankhplee.com/papers/purpo...
1112
Reposted by Yuxuan Li
Mark Riedl @markriedl.bsky.social · 05/02/2025
LLMs let bias slip out when acting as agents arxiv.org/abs/2501.17420
arxiv.org
Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models
While advances in fairness and alignment have helped mitigate overt biases exhibited by large language models (LLMs) when explicitly prompted, we hypothesize that these models may still exhibit implic...
0286
Reposted by Yuxuan Li
Sauvik Das @sauvik.me · 30/01/2025
In this new pre-print, my student @yuxuanli1225.bsky.social outlines how language agents making shockingly biased decisions, even when their words seem "unbiased". Also, the latest models are better at hiding bias, but it still drives what they do. It's also his FIRST Ph.D. paper! Please boost :)
1113
Yuxuan Li @yuxuanli1225.bsky.social · 30/01/2025
🚨 Advanced LLMs may sound unbiased—but it's a ruse. We show that demographically-informed language agents reveal stark, implicit biases in decision-making. New preprint: “Actions Speak Louder Than Words: Agent Decisions Reveal Implicit Biases in Language Models” 👉 arxiv.org/abs/2501.17420 [1/n]
1194