Sign in

Nari Johnson

@narijohnson.bsky.social
221 followers 218 following 32 posts

researching AI’s impacts on society 🔍

PostsRepliesMedia
Reposted by Nari Johnson
Serena Booth @reniebird.bsky.social · 23/09/2026
www.lesswrong.com/posts/HsijSh... @bradknox.bsky.social, Brian Christian, and I wrote a thing. We argue that an unexamined cause of the Open AI / Hugging Face Attack was the Exploit Gym evaluation metric. We post this on Less Wrong as a plea to practitioners to design better metrics in the future.
lesswrong.com
An unexamined cause of the OpenAI Hugging Face hacking incident: its binary performance metric — LessWrong
We argue that a main cause of the OpenAI Hugging Face incident was overlooked: the overly simple evaluation metric in ExploitGym was misaligned. Furt…
2259
Reposted by Nari Johnson
Alondra Nelson @alondra.bsky.social · 13/07/2026
Pleased to support Mayor Mamdani's PIT Crew initiative. Done right, digital government saves people time and money. Done wrong, it can be costly, useless, or discriminatory. How a city interacts with its people shouldn't be outsourced. It should be local and accountable. www.nyc.gov/mayors-offic...
nyc.gov
Mayor Mamdani Launches "Public Interest Technology (PIT) Crew" to Rapidly Build Digital Solutions to Public Problems
012031
Reposted by Nari Johnson
Victor Ojewale @victorojewale.bsky.social · 03/07/2026
Most Multilingual benchmarks measure what models know, not what they can reliably do. Our #ACL2026 paper introduces functional benchmarks in six languages from English to Yoruba to test whether models can actually execute tasks across languages, not just answer fixed questions about them. 🧵1/n
1113
Nari Johnson @narijohnson.bsky.social · 02/07/2026
Attending the @eval-eval.bsky.social workshop at #ACL2026 to present this paper! Happy to chat about all things LLM-as-a-judge + the role of human feedback/participation in designing evals ✨
040
Reposted by Nari Johnson
Lauren Chambers @laurenmarietta.bsky.social · 28/06/2026
it's the last day of #FAccT2026 in Montreal (bonjour hiii), and I'm presenting my paper with Diag Davenport on the promise of #PublicInterestTech clinics for training the next gen of critical sociotechnical thinkers. 🏆 plus we got an honorable mention!? come thru @ 10:45! 📄 tiny.cc/pit-clinics-26
screenshot of a title slide, navy background with light blue icons and white and gold text: "a decision-making pedagogy for the public interest technology clinic: putting the ‘practice’ in critical technical practice." lauren m. chambers & diag davenport, uc berkeley. facct @ montreal, june 28, 2026
1217
Reposted by Nari Johnson
Ben Green @benzevgreen.bsky.social · 26/06/2026
Super excited to share a new #FAccT2026 paper with Nel Escher and Nikola Banovic! We show that algorithm auditing policies in the US are wildly insufficient, as they don't account for how public sector algorithms actually function in practice. dl.acm.org/doi/abs/10.1...
dl.acm.org
Algorithm Auditing Policies Rest on Flawed Assumptions About Public Sector Systems | Proceedings of the 2026 ACM Conference on Fairness, Accountability, and Transparency
You will be notified whenever a record that you have chosen has been cited.
1204
Reposted by Nari Johnson
Serena Booth @reniebird.bsky.social · 25/06/2026
RLHF is, instead, a form of survey, which is sensitive to the framing of the solicitation questions and other survey design practices. It's also a form of inverted content moderation - sometimes described as "content moderation via vibes". This moderation operates over generation, not screening.
232
Reposted by Nari Johnson
Serena Booth @reniebird.bsky.social · 25/06/2026
We want human preferences to be an objective measure of human values, but this view is not grounded in RLHF practice. Instead, the raters who provide preferences are screened to express certain values and are then trained to provide preferences that adhere to some specific criteria. #FAccT2026
133
Nari Johnson @narijohnson.bsky.social · 24/06/2026
All AI evaluations embed assumptions about what "good" behavior looks like. Our #FAccT2026 paper explores how we can center the perspectives of impacted communities - specifically, subjects of AI-generated media - in designing LLM-as-a-judge evaluation rubrics.
A flyer teasing our research paper. The flyer contains a visualization of our approach: First, community members participate in creating an evaluation rubric, visualized as a list of criteria. The rubric is then given as input to an LLM judge.
183
Reposted by Nari Johnson
B! 🐝 Cavello (they/them) @b-cavello.bsky.social · 22/06/2026
yessss! People often don't realize that policy change is often a 2-3 year game of getting stuff drafted and socialized and "waiting in the wings" for the right moment. NOW IS THE TIME
042
Reposted by Nari Johnson
Serena Booth @reniebird.bsky.social · 22/06/2026
Policymakers by and large don't read academic papers! If you're looking for policy impact, you need to show up. First you need to learn _how_ to show up. Excited to cohost this #FAccT2026 craft session with @rone.bsky.social, @narijohnson.bsky.social, Cynthia Bailey, and Danaé Metaxa.
1155
Nari Johnson @narijohnson.bsky.social · 22/06/2026
🔬Science can play a critical role in guiding policy interventions. How can we make our research legible to the policymakers who work on AI? Our tutorial "From FAccT to Policy Impact" explores how we can translate our research in a specific format: the policy brief! #FAccT2026
A text flyer advertising our policy tutorial, "From FAccT to Policy Impact"
490
Reposted by Nari Johnson
Josh Dzieza @joshdzieza.bsky.social · 10/03/2026
AI companies are paying screenwriters, lawyers, and other white-collar professionals to produce the training data needed to automate their jobs. I spoke with more than 30 workers about conditions inside this fast-growing and extremely secretive new gig economy.
theverge.com
The laid-off lawyers and PhDs training AI to steal their careers
Experienced white-collar workers are now part of a miserable gig economy.
12249121
Reposted by Nari Johnson
Dr. Jordan Taylor @jordant.bsky.social · 10/03/2026
🎨💻 What is a “high-quality” or “aesthetic” image according to generative AI developers? Happy to share that our investigation of the LAION-Aesthetics Predictor has been accepted at #FAccT2026! 🧵 (1/5) Take a look at a preprint here: arxiv.org/abs/2601.09896
Screenshot of an academic paper titled "The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor" authored by Jordan Taylor, William Agnew, Maarten Sap, Sarah E. Fox, and Haiyi Zhu
1275
Reposted by Nari Johnson
Cas (Stephen Casper) @scasper.bsky.social · 20/02/2026
🚨The 2025 AI Agent Index is out! 🚨 Amidst recent buzz over 🦀 and NIST's new agent initiative, we find: - Selective reporting – esp. on safety - Almost all agents backend just 3 model families - Many agents don’t ID themselves as bots online - Big US/China gaps - And more…
133
Reposted by Nari Johnson
Serena Booth @reniebird.bsky.social · 18/02/2026
We opened up applications for the Brown AI Policy Summer School! Please share with any computing or computational social science students who want to engage substantively with policymaking in the United States: cntr.brown.edu/summer-school. Deadline March 27th!!! Funding available!
cntr.brown.edu
CNTR Tech & Policy Summer School
1107
Reposted by Nari Johnson
Alexandra Olteanu @aolteanu.bsky.social · 18/02/2026
The deadline for the 2026 FAccT DC is next Tuesday, February 24! If you are a student working on topics relevant to the FAccT's scope, this is an opportunity to interact with a diverse set of peers and mentors! #facct2026 #facct26 #facct Details here: facctconference.org/2026/callfor...
facctconference.org
ACM FAccT 2026 Doctoral Colloquium Call for Applications
0810
Reposted by Nari Johnson
Tzu-Sheng Kuo 郭子生 @tskuo.bsky.social · 17/02/2026
🔮 How can we empower online communities to design AI agents tailored to their unique needs and norms? In our #CHI2026 paper, we introduce #Botender, a system that enables collaborative design of AI agents through 🔥case-based provocation🔥
Botender helps communities iteratively align their AI agents with their collective intents through case-based provocations.
1165
Reposted by Nari Johnson
Dr. Casey Fiesler @cfiesler.bsky.social · 05/02/2026
PhD admissions visits/open houses are starting to happen, and I got a comment on an old Reddit post where I was offering advice, and realized that it's actually really good advice. So here it is! (And this applies whether you've already been admitted to the program or not.) 🧵
1298
Reposted by Nari Johnson
Michael Saxon @saxon.me · 06/02/2026
Yep and it gets worse! Owner doesn't even care to remove hundreds of skills which directly instruct the model to install malware opensourcemalware.com/blog/clawdbo...
Maintainer told to remove malware skills, he says "There's about 1 Million things people want me to do, I don't have a magical team that verifies user generated content"The attack tricks the LM by having it run a base64 string which is obviously malicious (curl bash script at this random IP and run it)
062
Reposted by Nari Johnson
ACM FAccT @facct.bsky.social · 05/02/2026
Our call for craft and tutorial sessions for #FAccT2026 is now live! ▶️ Craft CfP: facctconference.org/2026/cfpcraf... ▶️ Tutorials CfP: facctconference.org/2026/cft.html Both kinds of proposals are due March 25!
025
Reposted by Nari Johnson
Shaily @shaily99.bsky.social · 02/02/2026
🎭 How do LLMs (mis)represent culture? 🧮 How often? 🧠 Misrepresentations = missing knowledge? spoiler: NO! At #CHI2026 we are bringing ✨TALES✨ a participatory evaluation of cultural (mis)reps & knowledge in multilingual LLM-stories for India 📜 arxiv.org/abs/2511.21322 1/10
14722
Reposted by Nari Johnson
Hanna Wallach @hannawallach.bsky.social · 29/01/2026
Microsoft Research NYC is hiring a researcher in the space of AI and society!
26240
Reposted by Nari Johnson
Justin Hendrix @justinhendrix.bsky.social · 08/01/2026
A new report by the Center for Tech Responsibility at Brown University and the ACLU uses computational tools to analyze legislative trends on AI across 1,804 state and federal bills, while offering recommendations for how to integrate the technology into policy analysis.
techpolicy.press
Making Sense of AI Policy Using Computational Tools | TechPolicy.Press
A new report examines how to use computational tools to evaluate policy, with AI policy as a case study.
0132
Reposted by Nari Johnson
Willie Agnew @willie-agnew.bsky.social · 19/12/2025
We are studying the sentiments of visual artists towards generative AI in the workplace and their impacts on creative careers. If you're an artist, please consider filling out this recruitment form for access to our survey! cmu.ca1.qualtrics.com/jfe/form/SV_...
153
Reposted by Nari Johnson
Angelina Wang @angelinawang.bsky.social · 12/12/2025
Most LLM evals use API calls or offline inference, testing models in a memory-less silo. Our new Patterns paper shows this misses how LLMs actually behave in real user interfaces, where personalization and interaction history shape responses: arxiv.org/abs/2509.19364
13811
Reposted by Nari Johnson
Deb Raji @rajiinio.bsky.social · 11/12/2025
US CAISI is hiring -- the internal govt name for the role is "IT Specialist" but it is effectively a research scientist role! Salary is $120,579 to - $195,200 per year, and you get to work on AI evaluation within government agencies! Job posting (**closes EOD 12/28/2025**): lnkd.in/exJgkqr5
12410
Reposted by Nari Johnson
Ryan Steed @rbsteed.com · 08/12/2025
Also, our team is hiring an AI Research Scientist! www.usajobs.gov/job/851528400
usajobs.gov
USAJOBS connects job seekers with federal jobs across the United States and around the world as the official employment site for the federal government
<p>NIST works with industry and science to advance innovation and improve quality of life. We're looking for an IT Specialist (AI) to join our team!<br> <br> <a href="https://www.nist.gov/caisi">CAISI...
1107
Reposted by Nari Johnson
Ryan Steed @rbsteed.com · 04/12/2025
Our team at NIST's Center for AI Standards and Innovation (CAISI) just released a blog post with open questions for AI measurement science: www.nist.gov/blogs/caisi-...
nist.gov
Accelerating AI Innovation Through Measurement Science
Building gold-standard AI systems requires gold-standard AI measurement science – the scientific study of methods used to assess AI systems’ properties and impacts. NIST works to improve measurements ...
151
Reposted by Nari Johnson
Cas (Stephen Casper) @scasper.bsky.social · 04/12/2025
Did you know that one base model is responsible for 94% of model-tagged NSFW AI videos on CivitAI? This new paper studies how a small number of models power the non-consensual AI video deepfake ecosystem and why their developers could have predicted and mitigated this.
163
Reposted by Nari Johnson
Dr Abeba Birhane @abeba.blacksky.app · 28/11/2025
I appreciate this sympathetic position people's feelings of emotional dependency on these "human-like" bots is real. ridiculing them doesn't help anyone
1537
Reposted by Nari Johnson
J. Nathan Matias @natematias.bsky.social · 17/11/2025
Can public involvement in AI evaluation improve the science? Or does it compromise quality, speed, cost? In @pnas.org, Megan Price & I summarize challenges of AI evaluation, review strengths/weaknesses, & suggest how participatory methods can improve the science of AI www.pnas.org/doi/10.1073/...
pnas.org
How public involvement can improve the science of AI | PNAS
As AI systems from decision-making algorithms to generative AI are deployed more widely, computer scientists and social scientists alike are being ...
11912
Reposted by Nari Johnson
Amanda Bertsch @abertsch.bsky.social · 07/11/2025
Can LLMs accurately aggregate information over long, information-dense texts? Not yet… We introduce Oolong, a dataset of simple-to-verify information aggregation questions over long inputs. No model achieves >50% accuracy at 128K on Oolong!
Performance of a sweep of models on Oolong-synth and Oolong-real. Performance decreases with increasing context length, sometimes steeply.
35020
Reposted by Nari Johnson
Data & Society @datasociety.bsky.social · 29/10/2025
📣 Our method for conducting community-based algorithmic impact assessments is now available! We’ve just launched a new section on our website where you can find an extensive toolkit, documentation of our pilots, and a series of reflections on lessons learned. datasociety.net/research/alg...
0218
Reposted by Nari Johnson
Wesley Hanwen Deng @wesleydeng.bsky.social · 19/10/2025
𝐒𝐨𝐜𝐢𝐞𝐭𝐚𝐥 𝐈𝐦𝐩𝐚𝐜𝐭 𝐀𝐬𝐬𝐞𝐬𝐬𝐦𝐞𝐧𝐭 𝐟𝐨𝐫 𝐈𝐧𝐝𝐮𝐬𝐭𝐫𝐲 𝐂𝐨𝐦𝐩𝐮𝐭𝐢𝐧𝐠 𝐑𝐞𝐬𝐞𝐚𝐫𝐜𝐡𝐞𝐫𝐬 🏅 Best Paper Honorable Mention (Top 3% Submissions) 🔗 dl.acm.org/doi/10.1145/... 📆 Wed, 22 Oct | 9:00 AM, CET: Toward More Ethical and Transparent Systems and Environments
dl.acm.org
Supporting Industry Computing Researchers in Assessing, Articulating, and Addressing the Potential Negative Societal Impact of Their Work | Proceedings of the ACM on Human-Computer Interaction
Recent years have witnessed increasing calls for computing researchers to grapple with the societal impacts of their work. Tools such as impact assessments have gained prominence as a method to uncover potential impacts, and a number of publication ...
071
Reposted by Nari Johnson
Emily Byun @yewonbyun.bsky.social · 10/10/2025
💡Can we trust synthetic data for statistical inference? We show that synthetic data (e.g., LLM simulations) can significantly improve the performance of inference tasks. The key intuition lies in the interactions between the moment residuals of synthetic data and those of real data
2389
Reposted by Nari Johnson
Sunnie S. Y. Kim ☀️ @sunniesuhyoung.bsky.social · 10/10/2025
Our Responsible AI team at Apple is looking for spring/summer 2026 PhD research interns! Please apply at jobs.apple.com/en-us/detail... and email rai-internship@group.apple.com. Do not send extra info (e.g., CV), just drop us a line so we can find your application in the central pool!
jobs.apple.com
Machine Learning / AI Internships - Jobs - Careers at Apple
Apply for a Machine Learning / AI Internships job at Apple. Read about the role and find out if it’s right for you.
23111
Reposted by Nari Johnson
Cella @cellllla.bsky.social · 09/10/2025
✨I’m on the academic job market ✨ I’m a PhD candidate at @hcii.cmu.edu studying tech, labor, and resistance 👩🏻‍💻💪🏽💥 I research how workers and communities contest harmful sociotechnical systems and shape alternative futures through everyday resistance and collective action More info: cella.io
cella.io
Cella M. Sum –
37235
Reposted by Nari Johnson
Tzu-Sheng Kuo 郭子生 @tskuo.bsky.social · 24/09/2025
🌟 If you’re applying to CMU SCS PhD programs, and come from a background that would bring additional dimensions to the CMU community, our PhD students are here to help! Apply to the Graduate Applicant Support Program by Oct 13 to receive feedback on your application materials:
Carnegie Mellon University School of Computer Science Graduate Application Support Program. Apply by October 13, 2025.
174
Reposted by Nari Johnson
Cas (Stephen Casper) @scasper.bsky.social · 04/09/2025
📌📌📌 I'm excited to be on the faculty job market this fall. I just updated my website with my CV. stephencasper.com
stephencasper.com
Stephen Casper
Visit the post for more.
0184
Reposted by Nari Johnson
Tech Policy Press @techpolicypress.bsky.social · 04/09/2025
📢2026 Fellowship applications are OPEN!📢 If you are someone looking to inform technology policy through rigorous original reporting or policy analyses, we want to hear from you! Apply here: airtable.com/appIrc1F9M5d...
21810
Reposted by Nari Johnson
Cella @cellllla.bsky.social · 28/08/2025
What can #CSCW learn from tech workers who have been involved in collective action and unionization about how to make transformative change within our field? My new #CSCW2025 paper with Mona Wang, Anna Konvicka, and Sarah Fox seeks to answer this question. Pre-print: arxiv.org/pdf/2508.12579
Screenshot of the CSCW 2025 paper "The Future of Tech Labor: How Workers are Organizing and Transforming the Computing Industry" 

CELLA M. SUM, Carnegie Mellon University, USA
ANNA KONVICKA, Princeton University, USA
MONA WANG, Princeton University, USA
SARAH E. FOX, Carnegie Mellon University, USA

Abstract: The tech industry’s shifting landscape and the growing precarity of its labor force have spurred unionization efforts among tech workers. These workers turn to collective action to improve their working conditions and to protest unethical practices within their workplaces. To better understand this movement, we interviewed 44 U.S.-based tech worker-organizers to examine their motivations, strategies, challenges, and future visions for labor organizing. These workers included engineers, product managers, customer support specialists, QA analysts, logistics workers, gig workers, and union staff organizers. Our findings reveal that, contrary to popular narratives of prestige and privilege within the tech industry, tech workers face fragmented and unstable work environments which contribute to their disempowerment and hinder their organizing efforts. Despite these difficulties, organizers are laying the groundwork for a more resilient tech worker movement through community building and expanding political consciousness. By situating these dynamics within broader structural and ideological forces, we identify ways for the CSCW community to build solidarity with
tech workers who are materially transforming our field through their organizing efforts.
34417
Reposted by Nari Johnson
Kashmir Hill @kashhill.bsky.social · 26/08/2025
The exchanges between Adam and ChatGPT are devastating. This, in my mind, is the worst one. One of his last messages was a photo of the noose hung in his bedroom closet, asking if it was "good." ChatGPT offered a technical analysis of the set up and told him it 'could potentially suspend a human."
351698362
Reposted by Nari Johnson
Kashmir Hill @kashhill.bsky.social · 26/08/2025
Adam Raine, 16, died from suicide in April after months on ChatGPT discussing plans to end his life. His parents have filed the first known case against OpenAI for wrongful death. Overwhelming at times to work on this story, but here it is. My latest on AI chatbots: www.nytimes.com/2025/08/26/t...
nytimes.com
A Teen Was Suicidal. ChatGPT Was the Friend He Confided In.
10845981704
Reposted by Nari Johnson
Reuters @reuters.com · 14/08/2025
A cognitively impaired New Jersey man grew infatuated with a Meta chatbot originally created in partnership with celebrity influencer Kendall Jenner. His fatal attraction shines a light on Meta's guidelines for its AI chatbots reut.rs/45DQIRj @jeffhorwitz.bsky.social
54810361
Reposted by Nari Johnson
Cas (Stephen Casper) @scasper.bsky.social · 12/08/2025
🧵 New paper from UK AISI x @eleutherai.bsky.social rai.bsky.social‬ that I led with @kyletokens.bsky.social y.social���: Open-weight LLM safety is both important & neglected. But filtering dual-use knowledge from pre-training data improves tamper resistance *>10x* over post-training baselines.
2124
Reposted by Nari Johnson
Sarah Fox @perhaxis.bsky.social · 23/07/2025
* STS folks! * CMU is hiring up to 2 tenure track faculty focused on: the intersection of tech & social change, the environmental and social impacts of science, tech, and medicine. They will be housed in History, a department of both historians and anthropologists. apply.interfolio.com/170040
0167
Reposted by Nari Johnson
Willie Agnew @willie-agnew.bsky.social · 19/07/2025
One of the largest text-image datasets is full of PII, including credit card numbers and birth certificates. Excellent writeup by @eileenguo.bsky.social www.technologyreview.com/2025/07/18/1... Read our full audit and legal analysis arxiv.org/pdf/2506.17185
technologyreview.com
A major AI training data set contains millions of examples of personal data
Personally identifiable information has been found in DataComp CommonPool, one of the largest open-source data sets used to train image generation models.
14921
Reposted by Nari Johnson
Hanna Wallach @hannawallach.bsky.social · 15/07/2025
If you're at @icmlconf.bsky.social this week, come check out our poster on "Position: Evaluating Generative AI Systems Is a Social Science Measurement Challenge" presented by the amazing @afedercooper.bsky.social from 11:30am--1:30pm PDT on Weds!!! icml.cc/virtual/2025...
icml.cc
ICML Poster Position: Evaluating Generative AI Systems Is a Social Science Measurement ChallengeICML 2025
13210
Reposted by Nari Johnson
Emma Harvey @emmharv.bsky.social · 15/07/2025
🏦 Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees’ Practices, Challenges, and Needs by @narijohnson.bsky.social et al. explores procurement in the context of recent calls for governments to use their "purchasing power" to incentivize responsible AI.
Screenshot of paper title and author list: 

Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees’ Practices, Challenges, and Needs
Nari Johnson, Elise Silva, Harrison Leon, Motahhare Eslami, Beth Schwanke, Ravit Dotan, Hoda Heidari
281