Sign in

Marzena Karpinska

@markar.bsky.social
4K followers 974 following 105 posts

#nlp researcher interested in evaluation including: multilingual models, long-form input/output, processing/generation of creative texts previous: postdoc @ umass_nlp phd from utokyo marzenakrp.github.io

PostsRepliesMedia
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 29/09/2026
Happy to see these papers accepted #Neurips2026 🎊 🎊 🎊 Please come and find us at the conference! More details & links in 🧵
012
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 29/09/2026
EMNLP 2026 is coming up soon and we will be there! See you in Budapest 🇭🇺 Oct 24–29 #EMNLP2026 Check out our work!
131
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 29/09/2026
I'm recruiting 1-2 PhD students to join our lab in Fall 2027, through either Computer Science or Information Science at the University of Colorado Boulder! Looking for people with interests in NLP plus [healthcare | literary studies | narratives | social media | etc.]. Join us! 🏔️☀️
cls-lab.com
CLS Lab — Culture, Language, & Systems | University of Colorado Boulder
The Culture, Language, & Systems Lab at CU Boulder studies the language systems that transmit and shape modern culture, from internet platforms to literary archives to language models.
06950
Marzena Karpinska @markar.bsky.social · 11/09/2026
still high, but better than allowing for 40+ per person, not sure how (who) will the shared first authorship be checked....
030
Reposted by Marzena Karpinska
ACL 2027 @aclmeeting.bsky.social · 11/09/2026
ACL Sustainable Reviewing Policy: We are introducing changes in the @ReviewAcl reviewing and submissions. The changes will cap authors and introduce changes in the reviewing to keep our community sustainable. #NLProc #NLP
12712
Marzena Karpinska @markar.bsky.social · 04/09/2026
My first PhD student (Weidong Zhang, yet to come to blsky) brought us new lab's gadgets and sweets 😍😍😍😍 I'm so happy! Thank you 🥰
020
Reposted by Marzena Karpinska
NAACL @naaclmeeting.bsky.social · 08/07/2026
🌉NAACL 2027 will be June 1-5 in San Francisco, CA!🌉 ARR submission deadline: Oct. 12, 2026
1113
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 22/07/2026
Regular reminder that we have an alt-ARR slack workspace where ACs and SACs can support each other through the sometimes confusing process of the ARR cycle! Post or DM me a good email address for a Slack invitation and I will add you. #EMNLP2026
1105
Marzena Karpinska @markar.bsky.social · 21/07/2026
this checkbox at #arr really seems like a get-out-of-jail-free card; if we really need to allow for lang edits, why not require the reviewers to submit their pre-GPTed draft along with the edited one? 😭😭😭
1131
Marzena Karpinska @markar.bsky.social · 17/07/2026
I've been getting tired of people arguing they 'only translated' or 'lightly edited' their reviews/papers #ARR so I decided to check whether that's really detected as 100% AI by @pangram.com ... Spoiler alert, it's not: marzenakarpinska.substack.com/p/no-ai-tran...
marzenakarpinska.substack.com
No, AI translations are NOT detected as AI generated text
And small language edits give low AI likelihood as well..
020
Reposted by Marzena Karpinska
David Mimno @dmimno.bsky.social · 15/07/2026
Text as Data is happening at Berkeley right before COLM. Please share!
03014
Reposted by Marzena Karpinska
David Jurgens @davidjurgens.bsky.social · 13/07/2026
Peer review was one of the most-discussed topics at #ACL2026 . Many folks were concerned about the incredible growth in the number of ARR submissions (17K for the May26 ARR cycle 😱), and even more shocked that ~40% didn't have any authors qualified to review. What is going on?? I did some digging...
3253
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 09/07/2026
hot take: don’t even use it for polishing your reviews. as an AC and SAC, the last thing i care about is your grammar, spelling, or formatting. i get the presentation issues with papers, where you might face reviewers’ language biases, but you really don’t need to worry about this for reviews.
3394
Marzena Karpinska @markar.bsky.social · 08/07/2026
I think I will be posting this after each #ARR cycle: Please 🥺🙏 let's prohibit AI review writing. Otherwise, we will get lazy reviewers claiming they wrote bullet points and used AI only to put that in prose.
1194
Reposted by Marzena Karpinska
Barbara Plank @barbaraplank.bsky.social · 08/07/2026
ACL 2027 will be in Kyoto, Japan 🇯🇵 🤩
06513
Marzena Karpinska @markar.bsky.social · 07/07/2026
5 years later MTurk is on its way out. A lot has improved in open-ended gen eval, but it still suffers from underreporting, lack of statistical analysis, and sometimes sloppy design. Perhaps authors should always do their own tasks to understand the implications of how they designed evals.
030
Marzena Karpinska @markar.bsky.social · 07/07/2026
A great lesson learned from #ACL2026 test of time award winner: sometimes it just takes time for people to appreciate your work
081
Marzena Karpinska @markar.bsky.social · 07/07/2026
In the spirit of #paperoverflow looking for emergency reviewers: - aspect-based summarization - hallucination detection - evaluation invariance - evaluation, evaluation bias - fine-grained evaluation for some, the reviewers filled in delay but became unresponsive later HELP, this house is on fire 🔥
002
Marzena Karpinska @markar.bsky.social · 07/07/2026
Great message from @barbaraplank.bsky.social unifying CL & NLP
031
Marzena Karpinska @markar.bsky.social · 06/07/2026
Only 17% of May #ARR authors are qualified for reviewing, and almost 40% of papers have no qualified people?!?!
2304
Marzena Karpinska @markar.bsky.social · 06/07/2026
Spotted by @jennarussell.bsky.social at #ACL2026 A sober reminder to check the output when using Claude et al to generate your presentation. Context, the user likely prompted for 8 min long presentation since that was the limit. The model, just added this silly "8 minutes" to all slides.
010
Reposted by Marzena Karpinska
Maite Taboada @mtaboada.bsky.social · 29/06/2026
For the last few months, I've been working with my new colleague at SFU @markar.bsky.social, students @yvesfrtl.bsky.social, @adampodoxin.bsky.social, Ty Brassington, and Roman Grundkiewicz of Microsoft, trying to understand the quality of machine translation for literary texts. ...
163
Reposted by Marzena Karpinska
Melanie Walsh @mellymeldubs.bsky.social · 24/06/2026
Excited to share this. @neel2112.bsky.social, @mariaa.bsky.social, and I analyzed 500K anonymous ChatGPT convos (shared w/ consent from WildChat) to see if people were generating fiction. We found tons of stories, fanfiction & erotica. Many users iterated on the same stories for days and weeks.
Screenshot of paper abstract that reads: 

AI FICTION IN THE WILD Neel Gupta  Maria Antoniak  Melanie Walsh

Some professional authors are beginning to use AI tools to help produce their fiction writing. Are readers using AI to generate fiction, too? Drawing on over 500,000 anonymized, English-language ChatGPT-user conversations (Zhao et al.), we find that more than one third of the conversations involve some form of fiction generation—including original stories, roleplay, fanfiction, and erotica. This AI-generated fiction is notably dominated by power users. We identify common fiction generation patterns and profiles among these users, including what we call infinite story demanders, who repeatedly request and revise variations of the same or similar narratives over extended periods of time. We show that users especially gravitate toward fanfiction and erotica, and that they are broadly drawn to generic forms, repetition, immediacy, and niche combinations of story elements. Our findings motivate two theoretical provocations. First, we argue that AI technologies may lead to a shift in the conventional relationship between the author and reader, potentially producing what we call a solipsistic reader-writer, who both generates and consumes fiction within a closed conversational loop, interacting with a machine rather than a human other. Second, we note that LLMs enable interactivity, play, and permutation in ways that are seemingly pleasurable for users, raising questions about where AI will fit into contemporary storytelling and entertainment ecosystems. We situate these developments within broader transformations in literature and media, including self-publishing, fanfiction, and pornography, and suggest that AI-generated fiction shares structural affinities with on-demand, personalized, and repetitive cultural forms.
515645
Reposted by Marzena Karpinska
angryinch.bsky.social @angryinch.bsky.social · 27/06/2026
It's invaluable in good literature to have someone who has actual, lived experience and a deep cultural and historical understanding of the literature performing the translation @jricole.bsky.social translation of the Rubaiyat is a great example. Omar's words ring out with truth and subtle humor.
011
Marzena Karpinska @markar.bsky.social · 26/06/2026
A lot of people focus on AI-written fiction, but what about AI literary translation? 📚 We find that AI translation can be readable.. ‼️BUT it also flattens characters' voices 🧙and is less immersive 🫣 than published human translations. Below is my favorite quote from a reader 📖 lait.cs.sfu.ca
0216
Reposted by Marzena Karpinska
yvesfrtl.bsky.social @yvesfrtl.bsky.social · 26/06/2026
📚 AI-written stories get attention, but AI-translated literature is quietly shaping how readers experience the author So what gets lost ⁉️ 🤖 AI translation into English is readable and often ‘fine’ ✍️ But human translation is valued more ‼️Both vary in quality BUT AI more, even within a single book
First page of the paper "AI translation of literary texts is ‘fine’, but readers still prefer human translations".
1125
Marzena Karpinska @markar.bsky.social · 25/06/2026
Slop-eds are on the rise and seem to remain undisclosed. Please check out update to our audit of AI in news: ainewsaudit.github.io This work will be presented next week at #ACL2026 (Sunday 4:00-5:30 oral session)
ainewsaudit.github.io
AI News Audit - Search Articles
100
Reposted by Marzena Karpinska
Jenna Russell @jennarussell.bsky.social · 25/06/2026
AI NEWS UPDATE We've added 200k+ recent news articles and 5k op-eds from Oct 2025 - June 2026. Local news is up to ~10.81% AI-generated.
162
Marzena Karpinska @markar.bsky.social · 23/05/2026
One more thing @tuhinchakr.bsky.social 's post reminded me of... people tend to rationalize and see things not there. We saw it already in GPT-2 stories - we *expect* things to *mean* something, so we tend to see things that are not there... (link to this old paper: aclanthology.org/2021.emnlp-m...)
032
Marzena Karpinska @markar.bsky.social · 23/05/2026
this is how massive illusion of 'creativity' gets crushed... please read it to understand why models may appear to produce coherent text but are in fact Frankenstein factory ...
052
Reposted by Marzena Karpinska
Ethan Mollick @emollick.bsky.social · 30/04/2026
"Load bearing," "I keep coming back to," "Not just X, but Y" A curse of using AI a lot is that you realize how much of the writing around you is just AI, now People who don't use AI have historically been unable to identify AI prose on sight, but those who use it a lot can spot the tells easily
2218326
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 12/05/2026
We were happy to be part of the B.C. AI Research-to-Adoption Summit last week co-hosted by @sfu.ca School of Computing Science and @science.ubc.ca 🎊🎊🎊 It was great to see this event so well attended!
022
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 04/05/2026
Happy to share our work accepted to @icmlconf.bsky.social 🥳🎉🎉🎉🎉 See you in Korean🇰🇷 Check out 🧵 for more details & links!
132
Reposted by Marzena Karpinska
Marzena Karpinska @markar.bsky.social · 16/04/2026
Thinking of relocating your lab to Canada? 🇨🇦 There is still time to apply for #impact+ chair position at #SFU (until April 24th) This comes with up to $1M/year research award for 8 years. Feel free to reach out if you have any questions! www.sfu.ca/research/imp...
sfu.ca
Canada Impact+ Research Chairs
042
Marzena Karpinska @markar.bsky.social · 16/04/2026
Thinking of relocating your lab to Canada? 🇨🇦 There is still time to apply for #impact+ chair position at #SFU (until April 24th) This comes with up to $1M/year research award for 8 years. Feel free to reach out if you have any questions! www.sfu.ca/research/imp...
sfu.ca
Canada Impact+ Research Chairs
042
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 09/04/2026
We are happy to see these papers accepted to #ACL2026 Looking forward to the conference! 🎊
011
Reposted by Marzena Karpinska
David Jurgens @davidjurgens.bsky.social · 27/03/2026
If you're reviewing ARR papers and want a tool to help you spot potential hallucinated references, I cooked this up for the ACL SACs and thought I would share it with the broader community github.com/davidjurgens...
github.com
GitHub - davidjurgens/hallucinated-reference-finder
Contribute to davidjurgens/hallucinated-reference-finder development by creating an account on GitHub.
32411
Reposted by Marzena Karpinska
Hal Daumé III @haldaume3.bsky.social · 03/03/2026
Come join TRAILS as a postdoc at UMD (and work w folks at GW, MSU & Cornell) to conduct research and scholarship focused on approaches to AI that advance trust and trustworthiness with a great group of colleagues! 🌐 go.umd.edu/trails-postd... 🗓️ Summer/Fall 2026 start
go.umd.edu
TRAILS UMD Post Doctoral Associate Job Description - Spring 2026
Post Doctoral Associate Institute for Trustworthy AI in Law & Society February 2026 The Institute for Trustworthy AI in Law & Society (TRAILS) and the University of Maryland aim to transform the pr...
076
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 23/02/2026
Last Friday, our "Women in AI Research" event brought together many talented young women! We were amazed by the level of discussion and project ideas the teams came up with! Grateful to the organizers @petitegeek.bsky.social's Rosie Lab & #WiCS and our sponsors #WiCS & @wimlworkshop.bsky.social
053
Reposted by Marzena Karpinska
eamt2026.bsky.social @eamt2026.bsky.social · 30/01/2026
Does your LLM have a "style"? 🎨 We are excited to announce #StyGenAI at #EAMT2026, a new workshop dedicated to #style in GenAI translation. From literary translation to controlling tone and register: if it's about style, it belongs here. 🔗 CFP & Details here: sites.google.com/view/worksho...
sites.google.com
Workshop StyGenAi
StyGenAI is the first workshop dedicated specifically to the study of style in GenAI-translated content. Hosted at EAMT 2026, the workshop provides a focused forum for examining how large language mod...
021
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 11/02/2026
This decision was made without consulting a single CU Boulder professor with AI expertise. And AFAIK there was one professor *total* involved. More info, such as it is, here: www.cu.edu/gen-ai Check the "Guiding Principles" section for entertainment.
cu.edu
Gen AI
[tabs][tab-item title="ChatGPT Edu Information"] ChatGPT Edu soon to be available for eligible CU faculty, staff and students The University of Colorado has entered into a three-year agreement with OpenAI, renewable annually, to provide ChatGPT Edu systemwide to students, staff and faculty. Each campus and the system office will stand up and manage its own secure instance of the tool to provide equitable access while ensuring security and privacy. 
85324
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 11/02/2026
Congratulations to all authors of @sfu-cs-ai.bsky.social papers accepted to @iclr-conf.bsky.social 2026 🥳🥳🥳🥳🥳 Please check out our work in 🧵
521
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 02/02/2026
This week we are excited to host Peter West from @cs.ubc.ca who will talk about “Mapping the (Jagged) Landscape of LLM Capabilities.”
Screenshot showing the title and abstract of talk by Peter West. The text says:
"Title: Mapping the (Jagged) Landscape of LLM Capabilities

Abstract: One key missing piece for the broad adoption of LLMs is intuition, specifically, human intuition about when and how models will succeed or fail across the diverse tasks we might apply them to. Your LLM might write a well-reasoned essay on 14th-century theology, but does that mean it can accurately answer questions on the same topic? This talk will focus on one aspect of my research, which is the characterization of model capabilities to begin to develop these intuitions. I will discuss recent projects that try to identify where these capabilities break down, with a particular focus on high information examples which will necessitate new hypotheses of how exactly artificial intelligence functions.

Speaker info: Peter West is an assistant professor at the University of British Columbia, broadly working on the capabilities and limits of LLMs. For example: the divergence of AI from human intuitions of intelligence, unpredictability and creativity in models, and studying LLMs with a non-interventional natural sciences lens. Peter completed his PhD at the University of Washington, Paul G School of Computer Science and Engineering. He completed a postdoc at the Stanford Institute for Human-Centered AI. His work has been recognized with best, outstanding, and spotlight papers in NLP and AI conferences."
052
Reposted by Marzena Karpinska
Yanai Elazar @yanai.bsky.social · 29/01/2026
🚨 New Study 🚨 @arxiv.bsky.social has recently decided to prohibit any 'position' paper from being submitted to its CS servers. Why? Because of the "AI slop", and allegedly higher ratios of LLM-generated content in review papers, compared to non-review papers.
2299
Marzena Karpinska @markar.bsky.social · 16/01/2026
Now is probably a good time to share that I left my job at @microsoft.com (will forever miss this team) and moved to Vancouver, Canada, where I'm starting my lab as an assistant professor at the gorgeous @sfu.ca 🏔️ I'm looking to hire 1-2 students starting in Fall 2026. Details in 🧵
1121
Reposted by Marzena Karpinska
Tuhin Chakrabarty @tuhinchakr.bsky.social · 22/10/2025
🚨New paper on AI & copyright Authors have sued LLM companies for using books w/o permission for model training. Courts however need empirical evidence of market harm. Our preregistered study exactly addresses this gap. Joint work w Jane Ginsburg from Columbia Law and @dhillonp.bsky.social 1/n🧵
12212
Reposted by Marzena Karpinska
Mor Naaman @informor.bsky.social · 23/10/2025
Well this is sure to be a blockbuster AI article... @jennarussell.bsky.social et al are kicking ass and taking names in journalism, both individuals and organizations. "AI use in American newspapers is widespread, uneven, and rarely disclosed" arxiv.org/abs/2510.18774
3228
Marzena Karpinska @markar.bsky.social · 22/10/2025
AI is infiltrating American newsrooms. Sadly, it is mostly *undisclosed* meaning that readers are often unaware that they are consuming LLM text. Even worse, we find some of these texts making it to the print press (undisclosed) Can we at least be honest about using models for editing?
050
Reposted by Marzena Karpinska
Jenna Russell @jennarussell.bsky.social · 22/10/2025
AI is already at work in American newsrooms. We examine 186k articles published this summer and find that ~9% are either fully or partially AI-generated, usually without readers having any idea. Here's what we learned about how AI is influencing local and national journalism:
55629