Sign in

Dirk Wulff

@dirkwulff.bsky.social
708 followers 375 following 90 posts

Cognitive and decision science at MPI for Human Development & Uni Basel | Data science at therbootcamp.github.io | R, language models, and sustainability (text2sdg.io).

PostsRepliesMedia
Dirk Wulff @dirkwulff.bsky.social · 01/10/2026
📣 Delighted to share that I have joined WU (Vienna University of Economics and Business) as Professor of Artificial Intelligence and Behavioral Science. Announcements for at least one postdoc-level position will follow soon. Please reach out if you are interested.
67916
Dirk Wulff @dirkwulff.bsky.social · 30/09/2026
Women still exit psychological research at a disproportionate rate. Tracking 80k researchers, our new article maps the gender attrition gap and shows that differences in performance, networks, and institutions explain only part of it. psycnet.apa.org/doi/10.1037/... arxiv.org/abs/2510.13273
0187
Dirk Wulff @dirkwulff.bsky.social · 29/09/2026
🚨 New preprint Excited to share new work led by Raluca Rilla, showing how cheap, open agents threaten online data collection and how open-text analysis beyond conventional traps can help mitigate this threat. 🔗 arxiv.org/abs/2609.31054 With @Anne-Marie Nussberger and @ruimata.bsky.social .
044
Reposted by Dirk Wulff
Julian Berger @officialberger.bsky.social · 22/09/2026
New paper out in @pnas.org 🚨 Human learning is an understudied but promising lever for boosting human–AI synergy @jasonburton.bsky.social @ralfkurvers.bsky.social @stefanherzog.bsky.social @dirkwulff.bsky.social
12113
Reposted by Dirk Wulff
Ralf Kurvers @ralfkurvers.bsky.social · 18/09/2026
Very proud of @kirikuroda.bsky.social for preprinting this study on social influence and information cascades in experienced based decision making. osf.io/preprints/ps... @arc-mpib.bsky.social. Great collab with @simyciri.bsky.social @dirkwulff.bsky.social
osf.io
OSF
1179
Dirk Wulff @dirkwulff.bsky.social · 03/08/2026
🚨 Out now in TICS How can language models help cognitive science? @ruimata.bsky.social & I outline 5 uses: mapping research fields, formalizing theories, cleaning up constructs & measures, predicting behavior across tasks, and capturing environmental variation. 🔗 doi.org/10.1016/j.ti...
03712
Dirk Wulff @dirkwulff.bsky.social · 01/07/2026
🚨 Now out in PNAS In decision research, "talk is cheap." We show it isn't. Using LLMs to analyze participants' free-text explanations of their choices, we find verbal reports are a rich, scalable window into how people actually decide. 🔗 www.pnas.org/doi/10.1073/... Led by Kamil Fulawka
04215
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
These results reflect poorly on the value of human instruments for LLM profiling. However, they are not driven by standard post-training biases and do not rule out the existence of stable dispositions in LLM. We call for the design of dedicated instruments. Preprint: arxiv.org/abs/2606.20205
arxiv.org
Apparent Psychological Profiles of Large Language Models are Largely a Measurement Artifact
Psychological instruments designed for humans are increasingly used to assign large language models (LLMs) stable psychological profiles that affect their usability, safety assessment, and use as prox...
020
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
Finally, we show that all of this means that the psychological profile of LLM responses can be steered through the design of the measurement instruments, with profiles of very capable models varying dramatically as a function of whether forward or reversed instruments were used.
110
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
Across 29 self-report and task instruments measuring personality and risk-taking, we show that reliability (internal consistency) is driven by response bias and therefore is fully dependent on the instruments' levels of response orthogonality.
100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
We show that response bias in LLMs is partially but not fully reduced for larger and proprietary versus open models, with all models exceeding the level of response bias in humans.
100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
For a big five inventory, we show that our sample of LLM models (N = 56) scatters along an axis indicative of response biases-driven responding, opposite to the pattern found in humans.
100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
We base our analysis on a formal psychometric framework that unifies the idea of reverse-keying across self-report and task instruments under the notion of response orthogonality.
100
Dirk Wulff @dirkwulff.bsky.social · 19/06/2026
🚨 New preprint Do LLMs have stable dispositions measurable using instruments designed for humans? Led by @jelenameyer.bsky.social, our new preprint shows that LLM responses are driven by response bias, rendering their psychological profiles method artifacts. @dgarcia.eu @arc-mpib.bsky.social
1216
Dirk Wulff @dirkwulff.bsky.social · 10/06/2026
📣 Reporting checklist for LLMs in behavioral and social science New article presenting a consensus-based reporting checklist (GUIDE-LLM) for the use of LLMs in the behavioral and social sciences to foster transparency, reproducibility, and ethical use. 🔗 www.nature.com/articles/s41...
0113
Dirk Wulff @dirkwulff.bsky.social · 20/05/2026
🚨 New preprint 🚨 Excited to share a new preprint led by @zakashussain.bsky.social. We review how large language models compare to humans in behavior and representations. Our conclusion: LLM cognition is not simply human-like or non-human-like. It is jagged. 🔗 Preprint: osf.io/preprints/ps...
083
Dirk Wulff @dirkwulff.bsky.social · 18/05/2026
🚨 Last chance to register - please share 🚨 📣 Behavioral Clones Workshop - Join us online! One-day workshop on behavioral clones — 🤖 AI systems that model human behavior — following the Machine+Behavior conference. 📅 May 20, 9 am - 6:15 pm (CET) 🔗 Register: www.eventbrite.ch/e/behavioral...
eventbrite.ch
Behavioral Clones online participation
A one-day workshop on behavioral clones — AI systems that model human decision-making and social behavior.
013
Dirk Wulff @dirkwulff.bsky.social · 13/05/2026
Workshop program: center-for-humans-and-machines.github.io/behavioral-c...
center-for-humans-and-machines.github.io
Behavioral Clones Workshop 2026
One-day workshop on AI systems that model human decision-making and social behavior — 20 May 2026, MPIB Berlin.
000
Dirk Wulff @dirkwulff.bsky.social · 13/05/2026
🚨 Additional tickets 🚨 📣 Behavioral Clones Workshop - Join us online! One-day workshop on behavioral clones — 🤖 AI systems that model human behavior — following the Machine+Behavior conference. 📅 May 20, 9 am - 6:15 pm (CET) 🔗 Register: www.eventbrite.ch/e/behavioral...
130
Dirk Wulff @dirkwulff.bsky.social · 11/05/2026
📣 Postdoc position Build 🤖 LLM-based behavioral clones for economic games together with Pınar Uğurlar and me. The position (up to 3 years) is based in Pınar's lab at Özyeğin in Istanbul, with regular visits to my lab. Please share / reach out if interested. 🔗 drive.google.com/file/d/1ftqi...
074
Dirk Wulff @dirkwulff.bsky.social · 05/05/2026
📣 Behavioral Clones Workshop - Join us online! A one-day workshop on behavioral clones — 🤖 AI systems that model human decision-making and social behavior — following the Machine Behavior conference. Join online via Zoom (no remote Q&A). Register at: www.eventbrite.ch/e/behavioral...
064
Dirk Wulff @dirkwulff.bsky.social · 27/04/2026
📣 PsychLing-101 — Deadline May 1 Join us in building a large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation. So far: 36 submissions and 40M+ observations All contributors become coauthors 🤝 🔗 Visit github.com/Data-X01/Psy...
github.com
GitHub - Data-X01/PsychLing-101: Large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation, driven by the community.
Large-scale, trial-level psycholinguistic database for cumulative science and LLM evaluation, driven by the community. - Data-X01/PsychLing-101
172
Dirk Wulff @dirkwulff.bsky.social · 22/04/2026
🚨 New preprint 🚨 Excited to share this work led by @kriegmair.bsky.social, establishing 🤖 machine individuality by decomposing 75M LLM ratings into shared variance, bias, idiosyncrasy, and noise, revealing coherent, cross-dimensional fingerprints unique to each model. 🔗 arxiv.org/pdf/2604.16755
051
Reposted by Dirk Wulff
Stefan Feuerriegel @sfeuerriegel.bsky.social · 23/03/2026
🚀Introducing 𝐆𝐔𝐈𝐃𝐄-𝐋𝐋𝐌: A reporting checklist for using LLMs in behavioral & social science ✅GUIDE-LLM is a reporting checklist designed by 80+ experts to improve transparency, reproducibility & ethical accountability of LLM-based research 📄 llm-checklist.com
llm-checklist.com
GUIDE-LLM
Reporting Checklist for Studies with Large Language Models in the Behavioral and Social Sciences
43420
Reposted by Dirk Wulff
Stefano Palminteri @stepalminteri.bsky.social · 05/03/2026
including historical querelles 🧐🙂https://medium.com/@stefano.palminteri/do-large-language-models-vindicate-skinners-approach-to-language-b8a323682b46
medium.com
Do Large Language Models Vindicate Skinner’s Approach to Language?
What large language models teach us about language, learning, and the long-running debate between Skinner and Chomsky
031
Dirk Wulff @dirkwulff.bsky.social · 05/03/2026
🚨 Updated preprint 🚨 Excited to share this updated preprint, in which @ruimata.bsky.social and I discuss five ways LLMs can help address longstanding challenges in cognitive science and psychology, examining both opportunities and pitfalls. 🔗 arxiv.org/abs/2511.00206
1103
Dirk Wulff @dirkwulff.bsky.social · 06/02/2026
🚨 New publication 🚨 Excited to share this new forum piece led by @kevinetiede.bsky.social We propose broadening risk communication to cover neglected dimensions outlined in our risk-information taxonomy, requiring greater reliance on simulated experiences. Open access: doi.org/10.1016/j.ti...
084
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026
Submit your abstract by 11 February 2026 (AOE): machinebehavior.science/behavioral-c...
machinebehavior.science
Machine+Behavior Conference
Welcome to the forefront of behavioral science in the digital age.
022
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026
The workshop focuses on machines trained to replicate human behavior, as scientific models of human cognition and action, as tools for large-scale behavioral simulation, and as objects of study in their own right. We invite talks, lightning talks, and posters.
110
Dirk Wulff @dirkwulff.bsky.social · 05/02/2026
🚨 Deadline extension 🚨 BEHAVIORAL CLONES WORKSHOP Imitating humans is a powerful route to machine behavior and an emerging lens on human cognition and social simulation. 🗓 Fri, 20 May 2026 📍 Following Machine+Behavior @ Max Planck Berlin 🗣️ Invited talks: Danica Dillion, Raja Marjieh, Marcel Binz
132
Dirk Wulff @dirkwulff.bsky.social · 28/01/2026
🚨 7 days left to submit your work!
000
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026
We further demonstrate an important moderating role of embedding extraction method and confirm results suggesting that final-layer embeddings underrepresent psycholinguistic information.
000
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026
We show that accessibility cascades from lexical to semantic information similarly across all transformer language models, but is still realized differently in encoder and decoder models.
100
Dirk Wulff @dirkwulff.bsky.social · 08/01/2026
🚨 New preprint 🚨 Excited to share new work led by @tikhomirova.bsky.social, presenting a large-scale evaluation of the accessibility of psycholinguistic information in transformer language models. Preprint: arxiv.org/abs/2601.03798
175
Reposted by Dirk Wulff
Levin Brinkmann @levinbrinkmann.bsky.social · 29/12/2025
WORKSHOP ON BEHAVIORAL CLONES Imitating humans is a powerful route to machine behavior, an emerging lens on human cognition, and a prerequisite for social simulation. 🗓 Fri 20 May 26 Following Machine+Behavior @ Max Planck Berlin Invited talks by @danicajdillion.bsky.social @marcelbinz.bsky.social
121
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025
Submit your abstract by 4 February 2026 (AOE): machinebehavior.science/behavioral-c...
machinebehavior.science
Machine+Behavior Conference
Welcome to the forefront of behavioral science in the digital age.
000
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025
The workshop focuses on machines trained to replicate human behavior, as scientific models of human cognition and action, as tools for large-scale behavioral simulation, and as objects of study in their own right. We invite submissions for talks, lightning talks, and posters.
100
Dirk Wulff @dirkwulff.bsky.social · 29/12/2025
WORKSHOP ON BEHAVIORAL CLONES Imitating humans is a powerful route to machine behavior and an emerging lens on human cognition and social simulation. 🗓 Fri, 20 May 2026 Following Machine+Behavior @ Max Planck Berlin Invited talks: Danica Dillon, @marcelbinz.bsky.social
171
Dirk Wulff @dirkwulff.bsky.social · 19/12/2025
🚨 New preprint 🚨 We present the decisions-from-experience database (DfE-DB), including data from 168 studies. The data are currently shared with the original authors and made public upon publication. 🔗 Database: github.com/dwulff/dfe-db 🔗 Preprint: www.biorxiv.org/content/10.6...
0239
Reposted by Dirk Wulff
Pantelis P. Analytis @pantelispa.bsky.social · 01/12/2025
I am hiring a postdoc for a DFF-funded project on social influence, and the decision processes that fuel rich-get-richer dynamics in the online/offline world. The position is for up to a year, competitive Danish salary, remote work possible. Interested or know somebody? DM me or share!
11924
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025
I am sorry to hear. If you are interested, there is research showing the usefulness of LLMs for psychological research, which in part derives from the models' ingestion of academic literature during training. I see no reason why LLM capabilities wouldn't generalize to this context to some extent.
010
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025
I think I grasped your point. But all good, no problem if we disagree.
100
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025
Logically correct---not quite what I said. I genuinely think that LLMs can provide useful signals on such questions. But we need to do our best to assess error and bias, like we should do with every other approach used to address such questions.
110
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025
The ratings should be validated, of course. But there are good reasons LLMs might "know" this up to some reasonable error and bias.
300
Dirk Wulff @dirkwulff.bsky.social · 10/12/2025
🔬 How has psychology contributed to achieving the 🇺🇳 SDGs? Our new preprint maps 230,000 psychology articles to the SDGs using our text2sdg R package (www.text2sdg.io) and analyzes historical trends in topics, national and gendered contributions, and citation patterns. 🔗 arxiv.org/abs/2512.08628
0117
Reposted by Dirk Wulff
Dirk Wulff @dirkwulff.bsky.social · 07/12/2025
@cos.io, I noticed that, since the relaunch, PsyArXiv preprints are no longer being indexed by Google Scholar. Do you know if this will be fixed? Thank you for your work!
171
Dirk Wulff @dirkwulff.bsky.social · 07/12/2025
@cos.io, I noticed that, since the relaunch, PsyArXiv preprints are no longer being indexed by Google Scholar. Do you know if this will be fixed? Thank you for your work!
171
Dirk Wulff @dirkwulff.bsky.social · 12/11/2025
With @tikhomirova.bsky.social, Valentin Kriegmair, Fritz Günther, Aliona Petrenco, Louis Schiekiera, @ercbravenewword.bsky.social, and @marcelbinz.bsky.social
010
Dirk Wulff @dirkwulff.bsky.social · 12/11/2025
🚨 Inviting collaborators! 🚨 We’re launching PsychLing-101 — an open, community-driven initiative to gather psycholinguistic datasets for cross-dataset analyses and the development of psycholinguistic foundation models. 👉 To contribute or propose a dataset, go to: 🔗 github.com/Data-X01/Psy...
lnkd.in
LinkedIn
This link will take you to a page that’s not on LinkedIn
197
Dirk Wulff @dirkwulff.bsky.social · 07/11/2025
How can LLMs advance the cognitive sciences? @ruimata.bsky.social and I put together a review on how LLMs can help us address five long-standing problems in cognitive science (e.g., disciplinary silos or lack of generalizability). 🔗 arxiv.org/2511.00206
0130