Dirk Wulff @dirkwulff.bsky.social · 01/10/2026📣 Delighted to share that I have joined WU (Vienna University of Economics and Business) as Professor of Artificial Intelligence and Behavioral Science. Announcements for at least one postdoc-level position will follow soon. Please reach out if you are interested. 67916
Dirk Wulff @dirkwulff.bsky.social · 30/09/2026Women still exit psychological research at a disproportionate rate. Tracking 80k researchers, our new article maps the gender attrition gap and shows that differences in performance, networks, and institutions explain only part of it. psycnet.apa.org/doi/10.1037/... arxiv.org/abs/2510.13273 0187
Dirk Wulff @dirkwulff.bsky.social · 29/09/2026🚨 New preprint Excited to share new work led by Raluca Rilla, showing how cheap, open agents threaten online data collection and how open-text analysis beyond conventional traps can help mitigate this threat. 🔗 arxiv.org/abs/2609.31054 With @Anne-Marie Nussberger and @ruimata.bsky.social . 044
Reposted by Dirk WulffJulian Berger @officialberger.bsky.social · 22/09/2026New paper out in @pnas.org 🚨 Human learning is an understudied but promising lever for boosting human–AI synergy @jasonburton.bsky.social @ralfkurvers.bsky.social @stefanherzog.bsky.social @dirkwulff.bsky.social 12113
Reposted by Dirk WulffRalf Kurvers @ralfkurvers.bsky.social · 18/09/2026Very proud of @kirikuroda.bsky.social for preprinting this study on social influence and information cascades in experienced based decision making. osf.io/preprints/ps... @arc-mpib.bsky.social. Great collab with @simyciri.bsky.social @dirkwulff.bsky.socialosf.ioOSF 1179
Dirk Wulff @dirkwulff.bsky.social · 03/08/2026🚨 Out now in TICS How can language models help cognitive science? @ruimata.bsky.social & I outline 5 uses: mapping research fields, formalizing theories, cleaning up constructs & measures, predicting behavior across tasks, and capturing environmental variation. 🔗 doi.org/10.1016/j.ti... 03712
Dirk Wulff @dirkwulff.bsky.social · 01/07/2026🚨 Now out in PNAS In decision research, "talk is cheap." We show it isn't. Using LLMs to analyze participants' free-text explanations of their choices, we find verbal reports are a rich, scalable window into how people actually decide. 🔗 www.pnas.org/doi/10.1073/... Led by Kamil Fulawka 04215
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026These results reflect poorly on the value of human instruments for LLM profiling. However, they are not driven by standard post-training biases and do not rule out the existence of stable dispositions in LLM. We call for the design of dedicated instruments. Preprint: arxiv.org/abs/2606.20205arxiv.orgApparent Psychological Profiles of Large Language Models are Largely a Measurement ArtifactPsychological instruments designed for humans are increasingly used to assign large language models (LLMs) stable psychological profiles that affect their usability, safety assessment, and use as prox... 020
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026Finally, we show that all of this means that the psychological profile of LLM responses can be steered through the design of the measurement instruments, with profiles of very capable models varying dramatically as a function of whether forward or reversed instruments were used. 110
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026Across 29 self-report and task instruments measuring personality and risk-taking, we show that reliability (internal consistency) is driven by response bias and therefore is fully dependent on the instruments' levels of response orthogonality. 100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026We show that response bias in LLMs is partially but not fully reduced for larger and proprietary versus open models, with all models exceeding the level of response bias in humans. 100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026For a big five inventory, we show that our sample of LLM models (N = 56) scatters along an axis indicative of response biases-driven responding, opposite to the pattern found in humans. 100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026We base our analysis on a formal psychometric framework that unifies the idea of reverse-keying across self-report and task instruments under the notion of response orthogonality. 100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026🚨 New preprint Do LLMs have stable dispositions measurable using instruments designed for humans? Led by @jelenameyer.bsky.social, our new preprint shows that LLM responses are driven by response bias, rendering their psychological profiles method artifacts. @dgarcia.eu @arc-mpib.bsky.social 1216
Dirk Wulff @dirkwulff.bsky.social · 10/06/2026📣 Reporting checklist for LLMs in behavioral and social science New article presenting a consensus-based reporting checklist (GUIDE-LLM) for the use of LLMs in the behavioral and social sciences to foster transparency, reproducibility, and ethical use. 🔗 www.nature.com/articles/s41... 0113
Dirk Wulff @dirkwulff.bsky.social · 20/05/2026🚨 New preprint 🚨 Excited to share a new preprint led by @zakashussain.bsky.social. We review how large language models compare to humans in behavior and representations. Our conclusion: LLM cognition is not simply human-like or non-human-like. It is jagged. 🔗 Preprint: osf.io/preprints/ps... 083
Dirk Wulff @dirkwulff.bsky.social · 18/05/2026🚨 Last chance to register - please share 🚨 📣 Behavioral Clones Workshop - Join us online! One-day workshop on behavioral clones — 🤖 AI systems that model human behavior — following the Machine+Behavior conference. 📅 May 20, 9 am - 6:15 pm (CET) 🔗 Register: www.eventbrite.ch/e/behavioral...eventbrite.chBehavioral Clones online participationA one-day workshop on behavioral clones — AI systems that model human decision-making and social behavior. 013
Dirk Wulff @dirkwulff.bsky.social · 13/05/2026Workshop program: center-for-humans-and-machines.github.io/behavioral-c...center-for-humans-and-machines.github.ioBehavioral Clones Workshop 2026One-day workshop on AI systems that model human decision-making and social behavior — 20 May 2026, MPIB Berlin. 000
Dirk Wulff @dirkwulff.bsky.social · 13/05/2026🚨 Additional tickets 🚨 📣 Behavioral Clones Workshop - Join us online! One-day workshop on behavioral clones — 🤖 AI systems that model human behavior — following the Machine+Behavior conference. 📅 May 20, 9 am - 6:15 pm (CET) 🔗 Register: www.eventbrite.ch/e/behavioral... 130
Dirk Wulff @dirkwulff.bsky.social · 11/05/2026📣 Postdoc position Build 🤖 LLM-based behavioral clones for economic games together with Pınar Uğurlar and me. The position (up to 3 years) is based in Pınar's lab at Özyeğin in Istanbul, with regular visits to my lab. Please share / reach out if interested. 🔗 drive.google.com/file/d/1ftqi... 074
Dirk Wulff @dirkwulff.bsky.social · 05/05/2026📣 Behavioral Clones Workshop - Join us online! A one-day workshop on behavioral clones — 🤖 AI systems that model human decision-making and social behavior — following the Machine Behavior conference. Join online via Zoom (no remote Q&A). Register at: www.eventbrite.ch/e/behavioral... 064
Dirk Wulff @dirkwulff.bsky.social · 27/04/2026📣 PsychLing-101 — Deadline May 1 Join us in building a large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation. So far: 36 submissions and 40M+ observations All contributors become coauthors 🤝 🔗 Visit github.com/Data-X01/Psy...github.comGitHub - Data-X01/PsychLing-101: Large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation, driven by the community.Large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation, driven by the community. - Data-X01/PsychLing-101 172
Dirk Wulff @dirkwulff.bsky.social · 22/04/2026🚨 New preprint 🚨 Excited to share this work led by @kriegmair.bsky.social, establishing 🤖 machine individuality by decomposing 75M LLM ratings into shared variance, bias, idiosyncrasy, and noise, revealing coherent, cross-dimensional fingerprints unique to each model. 🔗 arxiv.org/pdf/2604.16755 051
Reposted by Dirk WulffStefan Feuerriegel @sfeuerriegel.bsky.social · 23/03/2026🚀Introducing 𝐆𝐔𝐈𝐃𝐄-𝐋𝐋𝐌: A reporting checklist for using LLMs in behavioral & social science ✅GUIDE-LLM is a reporting checklist designed by 80+ experts to improve transparency, reproducibility & ethical accountability of LLM-based research 📄 llm-checklist.comllm-checklist.comGUIDE-LLMReporting Checklist for Studies with Large Language Models in the Behavioral and Social Sciences 43420
Reposted by Dirk WulffStefano Palminteri @stepalminteri.bsky.social · 05/03/2026including historical querelles 🧐🙂https://medium.com/@stefano.palminteri/do-large-language-models-vindicate-skinners-approach-to-language-b8a323682b46medium.comDo Large Language Models Vindicate Skinner’s Approach to Language?What large language models teach us about language, learning, and the long-running debate between Skinner and Chomsky 031
Dirk Wulff @dirkwulff.bsky.social · 05/03/2026🚨 Updated preprint 🚨 Excited to share this updated preprint, in which @ruimata.bsky.social and I discuss five ways LLMs can help address longstanding challenges in cognitive science and psychology, examining both opportunities and pitfalls. 🔗 arxiv.org/abs/2511.00206 1103
Dirk Wulff @dirkwulff.bsky.social · 06/02/2026🚨 New publication 🚨 Excited to share this new forum piece led by @kevinetiede.bsky.social We propose broadening risk communication to cover neglected dimensions outlined in our risk-information taxonomy, requiring greater reliance on simulated experiences. Open access: doi.org/10.1016/j.ti... 084
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026Submit your abstract by 11 February 2026 (AOE): machinebehavior.science/behavioral-c...machinebehavior.scienceMachine+Behavior ConferenceWelcome to the forefront of behavioral science in the digital age. 022
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026The workshop focuses on machines trained to replicate human behavior, as scientific models of human cognition and action, as tools for large-scale behavioral simulation, and as objects of study in their own right. We invite talks, lightning talks, and posters. 110
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026🚨 Deadline extension 🚨 BEHAVIORAL CLONES WORKSHOP Imitating humans is a powerful route to machine behavior and an emerging lens on human cognition and social simulation. 🗓 Fri, 20 May 2026 📍 Following Machine+Behavior @ Max Planck Berlin 🗣️ Invited talks: Danica Dillion, Raja Marjieh, Marcel Binz 132
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026We further demonstrate an important moderating role of embedding extraction method and confirm results suggesting that final-layer embeddings underrepresent psycholinguistic information. 000
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026We show that accessibility cascades from lexical to semantic information similarly across all transformer language models, but is still realized differently in encoder and decoder models. 100
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026🚨 New preprint 🚨 Excited to share new work led by @tikhomirova.bsky.social, presenting a large-scale evaluation of the accessibility of psycholinguistic information in transformer language models. Preprint: arxiv.org/abs/2601.03798 175
Reposted by Dirk WulffLevin Brinkmann @levinbrinkmann.bsky.social · 29/12/2025WORKSHOP ON BEHAVIORAL CLONES Imitating humans is a powerful route to machine behavior, an emerging lens on human cognition, and a prerequisite for social simulation. 🗓 Fri 20 May 26 Following Machine+Behavior @ Max Planck Berlin Invited talks by @danicajdillion.bsky.social @marcelbinz.bsky.social 121
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025Submit your abstract by 4 February 2026 (AOE): machinebehavior.science/behavioral-c...machinebehavior.scienceMachine+Behavior ConferenceWelcome to the forefront of behavioral science in the digital age. 000
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025The workshop focuses on machines trained to replicate human behavior, as scientific models of human cognition and action, as tools for large-scale behavioral simulation, and as objects of study in their own right. We invite submissions for talks, lightning talks, and posters. 100
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025WORKSHOP ON BEHAVIORAL CLONES Imitating humans is a powerful route to machine behavior and an emerging lens on human cognition and social simulation. 🗓 Fri, 20 May 2026 Following Machine+Behavior @ Max Planck Berlin Invited talks: Danica Dillon, @marcelbinz.bsky.social 171
Dirk Wulff @dirkwulff.bsky.social · 19/12/2025🚨 New preprint 🚨 We present the decisions-from-experience database (DfE-DB), including data from 168 studies. The data are currently shared with the original authors and made public upon publication. 🔗 Database: github.com/dwulff/dfe-db 🔗 Preprint: www.biorxiv.org/content/10.6... 0239
Reposted by Dirk WulffPantelis P. Analytis @pantelispa.bsky.social · 01/12/2025I am hiring a postdoc for a DFF-funded project on social influence, and the decision processes that fuel rich-get-richer dynamics in the online/offline world. The position is for up to a year, competitive Danish salary, remote work possible. Interested or know somebody? DM me or share! 11924
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025I am sorry to hear. If you are interested, there is research showing the usefulness of LLMs for psychological research, which in part derives from the models' ingestion of academic literature during training. I see no reason why LLM capabilities wouldn't generalize to this context to some extent. 010
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025I think I grasped your point. But all good, no problem if we disagree. 100
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025Logically correct---not quite what I said. I genuinely think that LLMs can provide useful signals on such questions. But we need to do our best to assess error and bias, like we should do with every other approach used to address such questions. 110
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025The ratings should be validated, of course. But there are good reasons LLMs might "know" this up to some reasonable error and bias. 300
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025🔬 How has psychology contributed to achieving the 🇺🇳 SDGs? Our new preprint maps 230,000 psychology articles to the SDGs using our text2sdg R package (www.text2sdg.io) and analyzes historical trends in topics, national and gendered contributions, and citation patterns. 🔗 arxiv.org/abs/2512.08628 0117
Reposted by Dirk WulffDirk Wulff @dirkwulff.bsky.social · 07/12/2025@cos.io, I noticed that, since the relaunch, PsyArXiv preprints are no longer being indexed by Google Scholar. Do you know if this will be fixed? Thank you for your work! 171
Dirk Wulff @dirkwulff.bsky.social · 07/12/2025@cos.io, I noticed that, since the relaunch, PsyArXiv preprints are no longer being indexed by Google Scholar. Do you know if this will be fixed? Thank you for your work! 171
Dirk Wulff @dirkwulff.bsky.social · 12/11/2025With @tikhomirova.bsky.social, Valentin Kriegmair, Fritz Günther, Aliona Petrenco, Louis Schiekiera, @ercbravenewword.bsky.social, and @marcelbinz.bsky.social 010
Dirk Wulff @dirkwulff.bsky.social · 12/11/2025🚨 Inviting collaborators! 🚨 We’re launching PsychLing-101 — an open, community-driven initiative to gather psycholinguistic datasets for cross-dataset analyses and the development of psycholinguistic foundation models. 👉 To contribute or propose a dataset, go to: 🔗 github.com/Data-X01/Psy...lnkd.inLinkedInThis link will take you to a page that’s not on LinkedIn 197
Dirk Wulff @dirkwulff.bsky.social · 07/11/2025How can LLMs advance the cognitive sciences? @ruimata.bsky.social and I put together a review on how LLMs can help us address five long-standing problems in cognitive science (e.g., disciplinary silos or lack of generalizability). 🔗 arxiv.org/2511.00206 0130