Sign in

Marzena Karpinska

@markar.bsky.social
4K followers 975 following 106 posts

#nlp researcher interested in evaluation including: multilingual models, long-form input/output, processing/generation of creative texts previous: postdoc @ umass_nlp phd from utokyo marzenakrp.github.io

PostsRepliesMedia
Marzena Karpinska @markar.bsky.social · 02/10/2026
Anyone going to #EMNLP2026? Please consider helping us understand how people use machine translation tools when abroad! (and if you can't, please share!)
000
Reposted by Marzena Karpinska
Ehud Reiter @ehudreiter.bsky.social · 02/10/2026
Going to EMNLP 2026 in Budapest? We are looking for 20-30 EMNLP attendees to help us understand how you handle language barriers during a conference trip. Contact Yujun Wang: y.wang2.25@abdn.ac.uk @weizhaonlp.bsky.social, @markar.bsky.social, Pinzhen Chen, Yujun Wang
022
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 29/09/2026
Happy to see these papers accepted #Neurips2026 🎊 🎊 🎊 Please come and find us at the conference! More details & links in 🧵
012
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 29/09/2026
EMNLP 2026 is coming up soon and we will be there! See you in Budapest 🇭🇺 Oct 24–29 #EMNLP2026 Check out our work!
131
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 29/09/2026
I'm recruiting 1-2 PhD students to join our lab in Fall 2027, through either Computer Science or Information Science at the University of Colorado Boulder! Looking for people with interests in NLP plus [healthcare | literary studies | narratives | social media | etc.]. Join us! 🏔️☀️
cls-lab.com
CLS Lab — Culture, Language, & Systems | University of Colorado Boulder
The Culture, Language, & Systems Lab at CU Boulder studies the language systems that transmit and shape modern culture, from internet platforms to literary archives to language models.
07052
Marzena Karpinska @markar.bsky.social · 11/09/2026
still high, but better than allowing for 40+ per person, not sure how (who) will the shared first authorship be checked....
030
Reposted by Marzena Karpinska
ACL 2027 @aclmeeting.bsky.social · 11/09/2026
ACL Sustainable Reviewing Policy: We are introducing changes in the @ReviewAcl reviewing and submissions. The changes will cap authors and introduce changes in the reviewing to keep our community sustainable. #NLProc #NLP
12712
Marzena Karpinska @markar.bsky.social · 04/09/2026
My first PhD student (Weidong Zhang, yet to come to blsky) brought us new lab's gadgets and sweets 😍😍😍😍 I'm so happy! Thank you 🥰
020
Reposted by Marzena Karpinska
NAACL @naaclmeeting.bsky.social · 08/07/2026
🌉NAACL 2027 will be June 1-5 in San Francisco, CA!🌉 ARR submission deadline: Oct. 12, 2026
1113
Marzena Karpinska @markar.bsky.social · 23/08/2026
seems iclr is actually doing just that this year
010
Marzena Karpinska @markar.bsky.social · 23/07/2026
Agree though I think a lot of students don't know either. I would think these may just be first authors for whom this is like 1st--2nd submission
020
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 22/07/2026
Regular reminder that we have an alt-ARR slack workspace where ACs and SACs can support each other through the sometimes confusing process of the ARR cycle! Post or DM me a good email address for a Slack invitation and I will add you. #EMNLP2026
1105
Marzena Karpinska @markar.bsky.social · 21/07/2026
this checkbox at #arr really seems like a get-out-of-jail-free card; if we really need to allow for lang edits, why not require the reviewers to submit their pre-GPTed draft along with the edited one? 😭😭😭
1131
Marzena Karpinska @markar.bsky.social · 17/07/2026
I've been getting tired of people arguing they 'only translated' or 'lightly edited' their reviews/papers #ARR so I decided to check whether that's really detected as 100% AI by @pangram.com ... Spoiler alert, it's not: marzenakarpinska.substack.com/p/no-ai-tran...
marzenakarpinska.substack.com
No, AI translations are NOT detected as AI generated text
And small language edits give low AI likelihood as well..
020
Reposted by Marzena Karpinska
David Mimno @dmimno.bsky.social · 15/07/2026
Text as Data is happening at Berkeley right before COLM. Please share!
03014
Marzena Karpinska @markar.bsky.social · 16/07/2026
Absolutely! And it actually worked in 99% of cases, the reviewers either upgraded the scores or were clear about what else they wanted... I hope whoever is our AC/SAC can do the same T-T
020
Marzena Karpinska @markar.bsky.social · 14/07/2026
my husband looked at my screen, saw ARR and asked if it's a pirate conference 😅
020
Reposted by Marzena Karpinska
David Jurgens @davidjurgens.bsky.social · 13/07/2026
Peer review was one of the most-discussed topics at #ACL2026 . Many folks were concerned about the incredible growth in the number of ARR submissions (17K for the May26 ARR cycle 😱), and even more shocked that ~40% didn't have any authors qualified to review. What is going on?? I did some digging...
3253
Reposted by Marzena Karpinska
Maria Antoniak @mariaa.bsky.social · 09/07/2026
hot take: don’t even use it for polishing your reviews. as an AC and SAC, the last thing i care about is your grammar, spelling, or formatting. i get the presentation issues with papers, where you might face reviewers’ language biases, but you really don’t need to worry about this for reviews.
3394
Marzena Karpinska @markar.bsky.social · 09/07/2026
funniest part is that the one i saw looked like bullet points copy-pasted from whichever-ai-tool. there was no space precisely where the bullet points changed but the person said they used ai proofreading, how come proofreading does not catch "no space" issues?
010
Marzena Karpinska @markar.bsky.social · 08/07/2026
It's a great policy but for me ACL* conferences are often the best fit. This is also community build and I want to believe we as a community can change this for better...
020
Marzena Karpinska @markar.bsky.social · 08/07/2026
I wish it was but the form even allow people to say they used AI for review. Of course the idea is something like "proofreading" but I see 100% AI reviews. I could try pointing them out but will be likely told they just use it for edits. If there is 0 AI policy it would be easier.
121
Marzena Karpinska @markar.bsky.social · 08/07/2026
Please 🙏 AI detectors are not that bad, if something sounds AI and @pangram alike say it's 100% AI it is not just using a model to polish and it can get sloppy. Even if the author meant well they can miss something model misrepresented.
120
Marzena Karpinska @markar.bsky.social · 08/07/2026
I think I will be posting this after each #ARR cycle: Please 🥺🙏 let's prohibit AI review writing. Otherwise, we will get lazy reviewers claiming they wrote bullet points and used AI only to put that in prose.
1194
Reposted by Marzena Karpinska
Barbara Plank @barbaraplank.bsky.social · 08/07/2026
ACL 2027 will be in Kyoto, Japan 🇯🇵 🤩
06513
Marzena Karpinska @markar.bsky.social · 07/07/2026
5 years later MTurk is on its way out. A lot has improved in open-ended gen eval, but it still suffers from underreporting, lack of statistical analysis, and sometimes sloppy design. Perhaps authors should always do their own tasks to understand the implications of how they designed evals.
030
Marzena Karpinska @markar.bsky.social · 07/07/2026
A great lesson learned from #ACL2026 test of time award winner: sometimes it just takes time for people to appreciate your work
081
Marzena Karpinska @markar.bsky.social · 07/07/2026
This is really insightful, thank you for sharing!
010
Marzena Karpinska @markar.bsky.social · 07/07/2026
That is actually a valid concern, but I also think we are already doing 'honorary authorships' (sadly) if people are submitting 40+ papers -- no way to even read that
000
Marzena Karpinska @markar.bsky.social · 07/07/2026
In the spirit of #paperoverflow looking for emergency reviewers: - aspect-based summarization - hallucination detection - evaluation invariance - evaluation, evaluation bias - fine-grained evaluation for some, the reviewers filled in delay but became unresponsive later HELP, this house is on fire 🔥
002
Marzena Karpinska @markar.bsky.social · 07/07/2026
that's a nice way of looking at it!
020
Marzena Karpinska @markar.bsky.social · 07/07/2026
I agree about stricter DRs, but I naively hope we can keep 3 reviewers... :/
010
Marzena Karpinska @markar.bsky.social · 07/07/2026
Yes, the current solution is to require at least 1 author per paper to contribute (either as a reviewer, AC, or SAC). This means that if no authors are qualified, you can't submit. A bit hard to balance openness with being under paper attack, I guess...
140
Marzena Karpinska @markar.bsky.social · 07/07/2026
This is great! I wonder what the acceptance rate will be for the papers with no qualified reviewers.
010
Marzena Karpinska @markar.bsky.social · 07/07/2026
Great message from @barbaraplank.bsky.social unifying CL & NLP
031
Marzena Karpinska @markar.bsky.social · 06/07/2026
Yes, that's precisely what was suggested.
170
Marzena Karpinska @markar.bsky.social · 06/07/2026
Only 17% of May #ARR authors are qualified for reviewing, and almost 40% of papers have no qualified people?!?!
2304
Marzena Karpinska @markar.bsky.social · 06/07/2026
Spotted by @jennarussell.bsky.social at #ACL2026 A sober reminder to check the output when using Claude et al to generate your presentation. Context, the user likely prompted for 8 min long presentation since that was the limit. The model, just added this silly "8 minutes" to all slides.
010
Reposted by Marzena Karpinska
Maite Taboada @mtaboada.bsky.social · 29/06/2026
For the last few months, I've been working with my new colleague at SFU @markar.bsky.social, students @yvesfrtl.bsky.social, @adampodoxin.bsky.social, Ty Brassington, and Roman Grundkiewicz of Microsoft, trying to understand the quality of machine translation for literary texts. ...
163
Reposted by Marzena Karpinska
Melanie Walsh @mellymeldubs.bsky.social · 24/06/2026
Excited to share this. @neel2112.bsky.social, @mariaa.bsky.social, and I analyzed 500K anonymous ChatGPT convos (shared w/ consent from WildChat) to see if people were generating fiction. We found tons of stories, fanfiction & erotica. Many users iterated on the same stories for days and weeks.
Screenshot of paper abstract that reads: 

AI FICTION IN THE WILD Neel Gupta  Maria Antoniak  Melanie Walsh

Some professional authors are beginning to use AI tools to help produce their fiction writing. Are readers using AI to generate fiction, too? Drawing on over 500,000 anonymized, English-language ChatGPT-user conversations (Zhao et al.), we find that more than one third of the conversations involve some form of fiction generation—including original stories, roleplay, fanfiction, and erotica. This AI-generated fiction is notably dominated by power users. We identify common fiction generation patterns and profiles among these users, including what we call infinite story demanders, who repeatedly request and revise variations of the same or similar narratives over extended periods of time. We show that users especially gravitate toward fanfiction and erotica, and that they are broadly drawn to generic forms, repetition, immediacy, and niche combinations of story elements. Our findings motivate two theoretical provocations. First, we argue that AI technologies may lead to a shift in the conventional relationship between the author and reader, potentially producing what we call a solipsistic reader-writer, who both generates and consumes fiction within a closed conversational loop, interacting with a machine rather than a human other. Second, we note that LLMs enable interactivity, play, and permutation in ways that are seemingly pleasurable for users, raising questions about where AI will fit into contemporary storytelling and entertainment ecosystems. We situate these developments within broader transformations in literature and media, including self-publishing, fanfiction, and pornography, and suggest that AI-generated fiction shares structural affinities with on-demand, personalized, and repetitive cultural forms.
515645
Reposted by Marzena Karpinska
angryinch.bsky.social @angryinch.bsky.social · 27/06/2026
It's invaluable in good literature to have someone who has actual, lived experience and a deep cultural and historical understanding of the literature performing the translation @jricole.bsky.social translation of the Rubaiyat is a great example. Omar's words ring out with truth and subtle humor.
011
Marzena Karpinska @markar.bsky.social · 26/06/2026
A lot of people focus on AI-written fiction, but what about AI literary translation? 📚 We find that AI translation can be readable.. ‼️BUT it also flattens characters' voices 🧙and is less immersive 🫣 than published human translations. Below is my favorite quote from a reader 📖 lait.cs.sfu.ca
0216
Reposted by Marzena Karpinska
yvesfrtl.bsky.social @yvesfrtl.bsky.social · 26/06/2026
📚 AI-written stories get attention, but AI-translated literature is quietly shaping how readers experience the author So what gets lost ⁉️ 🤖 AI translation into English is readable and often ‘fine’ ✍️ But human translation is valued more ‼️Both vary in quality BUT AI more, even within a single book
First page of the paper "AI translation of literary texts is ‘fine’, but readers still prefer human translations".
1125
Marzena Karpinska @markar.bsky.social · 25/06/2026
Slop-eds are on the rise and seem to remain undisclosed. Please check out update to our audit of AI in news: ainewsaudit.github.io This work will be presented next week at #ACL2026 (Sunday 4:00-5:30 oral session)
ainewsaudit.github.io
AI News Audit - Search Articles
100
Reposted by Marzena Karpinska
Jenna Russell @jennarussell.bsky.social · 25/06/2026
AI NEWS UPDATE We've added 200k+ recent news articles and 5k op-eds from Oct 2025 - June 2026. Local news is up to ~10.81% AI-generated.
162
Marzena Karpinska @markar.bsky.social · 23/05/2026
One more thing @tuhinchakr.bsky.social 's post reminded me of... people tend to rationalize and see things not there. We saw it already in GPT-2 stories - we *expect* things to *mean* something, so we tend to see things that are not there... (link to this old paper: aclanthology.org/2021.emnlp-m...)
032
Marzena Karpinska @markar.bsky.social · 23/05/2026
this is how massive illusion of 'creativity' gets crushed... please read it to understand why models may appear to produce coherent text but are in fact Frankenstein factory ...
052
Reposted by Marzena Karpinska
Ethan Mollick @emollick.bsky.social · 30/04/2026
"Load bearing," "I keep coming back to," "Not just X, but Y" A curse of using AI a lot is that you realize how much of the writing around you is just AI, now People who don't use AI have historically been unable to identify AI prose on sight, but those who use it a lot can spot the tells easily
2218326
Reposted by Marzena Karpinska
SFU AI @sfu-cs-ai.bsky.social · 12/05/2026
We were happy to be part of the B.C. AI Research-to-Adoption Summit last week co-hosted by @sfu.ca School of Computing Science and @science.ubc.ca 🎊🎊🎊 It was great to see this event so well attended!
022