Sign in

Elisa Kreiss

@elisakreiss.bsky.social
1.1K followers 961 following 22 posts

Assistant Professor at UCLA. Alum @StanfordNLP. NLP, Cognitive Science, Accessibility. www.coalas-lab.com/elisakreiss

PostsRepliesMedia
Reposted by Elisa Kreiss
Rachit Dubey @rachitdubey.bsky.social · 06/11/2025
Honored and excited to share that I am the winner of Nomis & Science Young Explorer Award!! Also thrilled to share that my article describing my research is out now in @science.org today! The normalization of (almost) everything www.science.org/doi/10.1126/... 1/
science.org
The normalization of (almost) everything: Our minds can get used to anything, and even crises start feeling normal
Our minds can get used to anything, and even crises start feeling normal
38821
Reposted by Elisa Kreiss
Badr AlKhamissi @bkhmsi.bsky.social · 25/09/2025
Now that the ICLR deadline is behind us, happy to share that From Language to Cognition has been accepted as an Oral at #EMNLP2025! 🎉 Looking forward to seeing many of you in Suzhou 🇨🇳
1223
Reposted by Elisa Kreiss
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 25/09/2025
For folks considering grad school in ML, my advice is to explore programs that mix ML with a domain interest. ML programs are wildly oversubscribed while a lot of the fun right now is in figuring out what you can do with it
815217
Reposted by Elisa Kreiss
Conference on Language Modeling @colmweb.org · 18/09/2025
Keynote spotlight #3: COLM's second day will open with some *real* intelligence 🧠🧠🧠 with Tom Griffiths from Princeton
0123
Reposted by Elisa Kreiss
Nathan Lambert @natolambert.bsky.social · 17/09/2025
I finally got around to making a tool to compare completions from SFT vs. RLHF trained models. This is a mini site for the RLHF book that I've wanted for a while. buff.ly/lqDL5wa It's always been hard to say what RLHF does to a model within a more complex post-training pipeline.
1216
Reposted by Elisa Kreiss
Conference on Language Modeling @colmweb.org · 16/09/2025
Keynote spotlight #1: COLM will open on Oct 7 with a keynote by @lukezettlemoyer.bsky.social about mixed-modals models 🌈
073
Reposted by Elisa Kreiss
Gabriele Sarti @gsarti.com · 15/09/2025
It is again the time of year when I beg @aclmeeting.bsky.social execs to rethink the current streaming platform system. For my #EMNLP2025 submissions, I am *required* to upload 2 video recordings + 2 posters + 2 slide decks. Why force both posters and talks for all? Nonsense.
0162
Elisa Kreiss @elisakreiss.bsky.social · 12/09/2025
New paper! We find that the more prompts "look like" common benchmark tasks, the more model behavior changes. Even when we encourage the model to explicitly infer gender, the instruction itself causes the opposite, namely boosting gender-neutral language. Check out Bufan's thread and our paper!
060
Reposted by Elisa Kreiss
Bufan Gao @bufangao.bsky.social · 11/09/2025
🚨 New #EMNLP2025 paper! Do LLMs exhibit distinct behavior when the prompt looks similar to common evaluation prompts? 👀 We show that prompts that signal bias evaluation can flip the measured bias. See below ⬇️
Violin plots of Probability of Pronoun Shift. Models show significant sensitivity to prompt changes: when prompts highlight gender evaluation, pronoun use shifts, with decreased “he” and increased “they” use.
131
Reposted by Elisa Kreiss
Richard Huskey @richardhuskey.bsky.social · 04/09/2025
We’re hiring @ucdavis.bsky.social Comm! TT Asst Prof in Mediated Interpersonal Communication. Our dept blends behavioral, biological & computational strengths at one of the most interdisciplinary campuses anywhere. Apps due Oct 15. Glad to chat recruit.ucdavis.edu/JPF07289
recruit.ucdavis.edu
Department of Communication - Mediated Interpersonal Communication
University of California, Davis is hiring. Apply now!
097
Reposted by Elisa Kreiss
Rachit Dubey @rachitdubey.bsky.social · 03/09/2025
My lab at UCLA is hiring 1-2 PhD students this cycle! Join us to work at the intersection of cognitive science and AI applied to pressing societal challenges like climate change. More info about me: rachit-dubey.github.io My lab: ucla-cocopol.github.io Please help repost/spread the word!
24219
Reposted by Elisa Kreiss
Naomi Saphra @nsaphra.bsky.social · 29/08/2025
How can an imitative model like an LLM outperform the experts it is trained on? Our new COLM paper outlines three types of transcendence and shows that each one relies on a different aspect of data diversity. arxiv.org/abs/2508.17669
29418
Reposted by Elisa Kreiss
Shahan Ali Memon @shahanmemon.bsky.social · 21/08/2025
The 2025-2026 #academic job market season is upon us. Here are some resources and #job postings for #faculty and #postdoc positions for those in Information/Computer/Social Sciences, and related fields. Will update the list as more #jobs are posted. 🧵(0/n) #AcademicChatter
3277
Reposted by Elisa Kreiss
Ken Liu @kzliu.bsky.social · 26/08/2025
New paper! We explore a radical paradigm for AI evals: assessing LLMs on *unsolved* questions. Instead of artificially difficult exams where progress ≠ value, we assess LLMs on organic, unsolved problems via reference-free LLM validation & community verification. LLMs solved ~10/500 so far:
261
Reposted by Elisa Kreiss
Frank Elavsky (he/him) ⌁ @frank.computer · 19/08/2025
Preprint announcement! 🎉 I've written a short position paper with Cindy Xiong Bearfield (a bias researcher in visualization) for @ieeevis.bsky.social Workshop on Accessibility on how using LLMs and generative models forces reliance (rather than builds trust) when used for accessibility.
arxiv.org
Playing telephone with generative models: "verification disability," "compelled reliance," and accessibility in data visualization
This paper is a collaborative piece between two worlds of expertise in the field of data visualization: accessibility and bias. In particular, the rise of generative models playing a role in accessibi...
1164
Reposted by Elisa Kreiss
Jessi Grieser @jessgrieser.com · 15/08/2025
UMich Ling is hiring for a cluster hire position in AI and Linguistics! Come work in a fabulous department with great grad students, supportive colleagues, and one overly school-spirited nut who will tell you where all the fun stuff is on campus. apply.interfolio.com/171006
apply.interfolio.com
Apply - Interfolio {{$ctrl.$state.data.pageTitle}} - Apply - Interfolio
01615
Reposted by Elisa Kreiss
Michael Yarbrough (he/him) @mwyarbrough.bsky.social · 10/08/2025
This is the first example I've seen of a university trying to mobilize constituents to stand up against Trump threats. There's been a lot of "wHaT eLsE CaN AdMiNiStRaToRs Do?" in The Discourse when this is exactly the thing they haven't been doing. Good for UCLA for trying. Others should follow.
7365138
Reposted by Elisa Kreiss
RedPen/BlackPen @redpenblackpen.bsky.social · 08/08/2025
Really just had something like this feeling this afternoon. It’s the high that keeps you coming back for more
Cartoon of a scientist peering through a microscope saying "Whoaah!!!". Title says "The best part of Science is knowing, for a moment, something that nobody else in the world knows."
48320
Elisa Kreiss @elisakreiss.bsky.social · 07/08/2025
Excellent non-technical explanation! Let’s raise AI literacy without scaring people off with overly complicated terminology while preserving technical accuracy!
030
Reposted by Elisa Kreiss
Antoine Bosselut @abosselut.bsky.social · 04/08/2025
The EPFL NLP lab is looking to hire a postdoctoral researcher on the topic of designing, training, and evaluating multilingual LLMs: docs.google.com/document/d/1... Come join our dynamic group in beautiful Lausanne!
docs.google.com
EPFL NLP Postdoctoral Scholar Posting - Swiss AI LLMs
The EPFL Natural Language Processing (NLP) lab is looking to hire a postdoctoral researcher candidate in the area of multilingual LLM design, training, and evaluation. This postdoctoral position is as...
02012
Reposted by Elisa Kreiss
Jerry Spanakis 🟥 @gerasimoss.bsky.social · 31/07/2025
Neat year #NLP conferences: 📍 #EACL2026: 24–29 March 🇲🇦 Rabat (ARR deadline: 6 Oct) 📍 #LREC2026: 11–16 May 🇪🇸 Palma de Mallorca (deadline: 17 Oct) 📍 #ACL2026: 2–7 July 🇺🇸 San Diego (ARR deadline: 6 Jan) 📍 #COLING2027: 9–14 May 2027 🇲🇴 Macao
0216
Reposted by Elisa Kreiss
Ahmad Beirami @abeirami.bsky.social · 25/07/2025
Only 2 out of 120 papers in my NeurIPS SAC batch have all reviewer scores ≥4! Don't despair with the low scores and focus on writing a clear and concise rebuttal! Good luck!
1354
Reposted by Elisa Kreiss
Rachit Dubey @rachitdubey.bsky.social · 25/07/2025
Woke up to a pleasant surprise! Guardian wrote a nice article covering our Binary climate data study! TLDR: “Before and after” climate data can help counter the boiling frog effect www.theguardian.com/news/2025/ju...
theguardian.com
‘Boiling frog’ effect makes people oblivious to threat of climate crisis, shows study
Research shows people’s response to binary data could drive more effective ways of communicating gravity of crisis
23920
Reposted by Elisa Kreiss
Rachit Dubey @rachitdubey.bsky.social · 23/07/2025
Headed to #CogSci2025 next week! Interested in computational cognitive science applied to climate change and real-world policy? Come find me! My lab at UCLA is recruiting 1–2 fully funded PhD students. We work on: • decision-making • climate perception • human–AI collaboration for policy design
1207
Reposted by Elisa Kreiss
Lucy Li @lucy3.bsky.social · 22/07/2025
I'm sadly not at #IC2S2 😭, but I will be at #ACL2025 in Vienna ☕️ next week!! Please spread the word that I'm recruiting prospective PhD students: lucy3.notion.site/for-prospect...
lucy3.notion.site
For Prospective PhD Students
I’m recruiting PhD students who will begin their degree in Fall 2026! I am an incoming assistant professor at Wisconsin-Madison’s Computer Sciences department, and my research focuses on natural langu...
0368
Reposted by Elisa Kreiss
Neil Lewis, Jr. @neillewisjr.bsky.social · 15/07/2025
We're hiring! academicjobsonline.org/ajo/jobs/30219 The Cornell Department of Communication is hiring an Assistant Professor in Science, Environment, and Risk Communication! Applications received by September 15, 2025, will be given full consideration.
academicjobsonline.org
Cornell University, Department of Communication
Job #AJO30219, WDR-00053939 Assistant Professor in Science, Environment, and Risk Communication, Department of Communication, Cornell University, Ithaca, New York, US
1134115
Reposted by Elisa Kreiss
Vilém Zouhar @zouhar.bsky.social · 15/07/2025
You have a budget to human-evaluate 100 inputs to your models, but your dataset is 10,000 inputs. Do not just pick 100 randomly!🙅 We can do better. "How to Select Datapoints for Efficient Human Evaluation of NLG Models?" shows how.🕵️ (random is still a devilishly good baseline)
2333
Reposted by Elisa Kreiss
Claus Wilke @clauswilke.com · 07/07/2025
This is the best explanation I've seen yet for _why_ language models prefer em-dashes (—). In brief: The tokenization scheme mean em-dashes result in a smaller loss than other, equivalent punctuation options. msukhareva.substack.com/p/the-myster...
msukhareva.substack.com
The mystery of em‑dashes: part two with quantitative evidence
A couple of weeks ago I made an assumption: the rise of em‑dashes in AI‑generated text happened because model providers started scanning older, pre‑Kindle books.
37128
Reposted by Elisa Kreiss
Mike Frank @mcxfrank.bsky.social · 01/07/2025
Experimentology is out today!!! A group of us wrote a free online textbook for experimental methods, available at experimentology.io - the idea was to integrate open science into all aspects of the experimental workflow from planning to design, analysis, and writing.
Experimentology cover: title and curves for distributions.
9535228
Reposted by Elisa Kreiss
eleutherai.bsky.social @eleutherai.bsky.social · 06/06/2025
Can you train a performant language model using only openly licensed text? We are thrilled to announce the Common Pile v0.1, an 8TB dataset of openly licensed and public domain text. We train 7B models for 1T and 2T tokens and match the performance similar models like LLaMA 1 & 2
214760
Reposted by Elisa Kreiss
Alex Gill @agill32.bsky.social · 04/06/2025
𝐖𝐡𝐚𝐭 𝐇𝐚𝐬 𝐁𝐞𝐞𝐧 𝐋𝐨𝐬𝐭 𝐖𝐢𝐭𝐡 𝐒𝐲𝐧𝐭𝐡𝐞𝐭𝐢𝐜 𝐄𝐯𝐚𝐥𝐮𝐚𝐭𝐢𝐨𝐧? (arxiv.org/abs/2505.22830) I'm happy to announce that the preprint release of my first project is online! Developed with the amazing support of @lasha.bsky.social & @anamarasovic.bsky.social
arxiv.org
What Has Been Lost with Synthetic Evaluation?
Large language models (LLMs) are increasingly used for data generation. However, creating evaluation benchmarks raises the bar for this emerging paradigm. Benchmarks must target specific phenomena, pe...
1124
Reposted by Elisa Kreiss
Robert Hawkins @rdhawkins.bsky.social · 28/05/2025
Happy to announce the first workshop on Pragmatic Reasoning in Language Models — PragLM @ COLM 2025! 🎉 How do LLMs engage in pragmatic reasoning, and what core pragmatic capacities remain beyond their reach? 🌐 sites.google.com/berkeley.edu/praglm/ 📅 Submit by June 23rd
sites.google.com
PragLM @ COLM '25
IMPORTANT DATES
13918
Reposted by Elisa Kreiss
Conference on Language Modeling @colmweb.org · 27/05/2025
Our discussion period just started. Authors, please read our instructions carefully. We require responses by June 2. But, what you really want to hear about is stats .... right? -> 🧵
2175
Reposted by Elisa Kreiss
GESIS - Leibniz-Institut für Sozialwissenschaften @gesis.org · 22/05/2025
#jobs #stellenangebote #GESISjobs #jobfairy Our Team Digital Society Observatory in #Cologne is looking for a #SeniorResearcher & #TeamLeader in Computational Social Science #CSS (Salary group 15 TV-L, working time 100%, initially for four years with possible tenure): gesis.jobs.personio....
01113
Reposted by Elisa Kreiss
Jacob Eisenstein is at CoLM 🌉 @jacobeisenstein.bsky.social · 21/05/2025
Google Deepmind is hiring a research scientist in Seattle to work on foundational research in language! job-boards.greenhouse.io/deepmind/job...
job-boards.greenhouse.io
Research Scientist, Foundational Research in Language, USA
Seattle, Washington, US
091
Reposted by Elisa Kreiss
Ana Marasović @anamarasovic.bsky.social · 20/05/2025
Do folks really find the ARR checklist valuable enough to justify that a paper submission takes this much effort?
4263
Reposted by Elisa Kreiss
andrea e. martin @andreaeyleen.eurosky.social · 20/05/2025
my lab (lacns.github.io) at @mpi-nl.bsky.social and @dondersinst.bsky.social is recruiting for two PhD and two postdoctoral positions funded by an @erc.europa.eu Consolidator - come join us! PhD: www.mpi.nl/career-educa... Postdoc: www.mpi.nl/career-educa... (please share widely)
lacns.github.io
Language and Computation in Neural Systems
We are an international group of scientists consisting of linguists, cognitive scientists, cognitive neuroscientists, computational neuroscientists, computational modellers, computational scientists, ...
38153
Reposted by Elisa Kreiss
Jennifer Hu @jennhu.bsky.social · 20/05/2025
Excited to share a new preprint w/ @michael-lepori.bsky.social & Michael Franke! A dominant approach in AI/cogsci uses *outputs* from AI models (eg logprobs) to predict human behavior. But how does model *processing* (across layers in a forward pass) relate to human real-time processing? 👇 (1/12)
Screenshot of Figure 1, which has two panels labeled (a) and (b). The caption states the following. Figure 1: Overview of our study. (a) Experiment 1: We explore whether forward passes show mechanistic signatures of competitor interference, first preferring a salient competing intuitive answer before preferring the correct answer. (b) Experiment 2: We systematically investigate the ability of dynamic measures derived from forward passes to predict indicators of processing load in humans.
25415
Reposted by Elisa Kreiss
Tokenization Workshop (TokShop) @COLM2026 @tokshop.bsky.social · 14/05/2025
📣 Call for Paper Alert: TokShop @ ICML 2025 TokShop explores tokenization across all data modalities. Topics include: subword NLP techniques, multimodal approaches, multilingual challenges, post-training modification, alternative representations, and statistical perspectives.
openreview.net
ICML 2025 Workshop TokShop
Welcome to the OpenReview homepage for ICML 2025 Workshop TokShop
11712
Reposted by Elisa Kreiss
Nick Tomlin @nickatomlin.bsky.social · 13/05/2025
I'm particularly fond of this new benchmark paper we wrote, which aims to scalably evaluate whether language models can generalize to arbitrary new tasks. The core idea is to use LLMs to generate new games, and then evaluate whether LLMs can play those games 📄: arxiv.org/abs/2505.07215
Title and abstract of the paper, "Measuring General Intelligence with Generated Games"
3339
Reposted by Elisa Kreiss
Lucy Li @lucy3.bsky.social · 05/05/2025
I'm joining Wisconsin CS as an assistant professor in fall 2026!! There, I'll continue working on language models, computational social science, & responsible AI. 🌲🧀🚣🏻‍♀️ Apply to be my PhD student! Before then, I'll postdoc for a year in the NLP group at another UW 🏔️ in the Pacific Northwest
Wisconsin-Madison's tree-filled campus, next to a big shiny lake A computer render of the interior of the new computer science, information science, and statistics building. A staircase crosses an open atrium with visibility across multiple floors
1614514
Reposted by Elisa Kreiss
Andrew Lampinen @lampinen.bsky.social · 02/05/2025
How do language models generalize from information they learn in-context vs. via finetuning? In arxiv.org/abs/2505.00661 we show that in-context learning can generalize more flexibly, illustrating key differences in the inductive biases of these modes of learning — and ways to improve finetuning. 1/
arxiv.org
47722
Reposted by Elisa Kreiss
Myra Cheng @myra.bsky.social · 02/05/2025
How does the public conceptualize AI? Rather than self-reported measures, we use metaphors to understand the nuance and complexity of people’s mental models. In our #FAccT2025 paper, we analyzed 12,000 metaphors collected over 12 months to track shifts in public perceptions.
34914
Reposted by Elisa Kreiss
Kristina Gligoric @IC2S2 @gligoric.bsky.social · 29/04/2025
I'm at #NAACL2025 & would love to chat! On Thu, I'll be presenting our paper on LLM annotations for valid estimates in social science (oral talk @ 4pm, CSS session) 📝 Paper: aclanthology.org/2025.naacl-l... *New* 🧪 tutorial (LLMs + Prolific data collection): github.com/kristinaglig...
1323
Reposted by Elisa Kreiss
Lucy Li @lucy3.bsky.social · 25/04/2025
We won an outstanding paper award!! 2025.naacl.org/blog/best-pa...
2025.naacl.org
Announcing the NAACL 2025 Award Winners!
Heading into the last week before the conference, we’d like to announce our Award Winners. The Best Paper and Best Theme Paper winners will present at our closing session, but all papers (including t...
1544
Reposted by Elisa Kreiss
Kaitlyn Zhou @kaitlynzhou.bsky.social · 24/04/2025
Life update! Excited to announce that I’ll be starting as an assistant professor at Cornell Info Sci in August 2026! I’ll be recruiting students this upcoming cycle! An abundance of thanks to all my mentors and friends who helped make this possible!!
19768
Reposted by Elisa Kreiss
Rachit Dubey @rachitdubey.bsky.social · 17/04/2025
🚨 New in Nature Human Behavior! 🚨 Binary climate data visuals amplify perceived impact of climate change. Both graphs in this image reflect equivalent climate change trends over time, yet people consistently perceive climate change as having a greater impact in the right plot than the left. 👇1/n
524587
Reposted by Elisa Kreiss
Tom McCoy @rtommccoy.bsky.social · 11/04/2025
The Department of Linguistics at Yale University is seeking to hire a one-year lecturer with expertise in Computational Linguistics and/or Syntax for the 2025-2026 school year. See the job ad for more info: linguistlist.org/issues/36/11... Please send to any who might be interested!
linguistlist.org
LINGUIST List 36.1177 Jobs: Computational Linguistics, Syntax: Lecturer, Yale University
The LINGUIST List, International Linguistics Community Online.
0126
Reposted by Elisa Kreiss
Srishti @srishtiy.bsky.social · 07/04/2025
When we read the news, images can convey different things than text itself. Unlike other works which look at text, we study this as a “multimodal” framing problem & analyze where text and images communicate different “frames”. Checkout our paper here: arxiv.org/abs/2503.20960 @aicentre.dk
arxiv.org
Multi-Modal Framing Analysis of News
Automated frame analysis of political communication is a popular task in computational social science that is used to study how authors select aspects of a topic to frame its reception. So far, such s...
1285
Reposted by Elisa Kreiss
Kenny Peng @kennypeng.bsky.social · 02/04/2025
Our lab had a #dogathon 🐕 yesterday where we analyzed NYC Open Data on dog licenses. We learned a lot of dog facts, which I’ll share in this thread 🧵 1) Geospatial trends: Cavalier King Charles Spaniels are common in Manhattan; the opposite is true for Yorkshire Terriers.
25214