Reposted by Chas Monge
Out now in Scientific Reports! Despite high correlations, ChatGPT models failed to replicate human moral judgments. We propose tests beyond correlation to compare LLM data and human data.
With @mattgrizz.bsky.social @andyluttrell.bsky.social @chasmonge.bsky.social
www.nature.com/articles/s41...
nature.com
ChatGPT does not replicate human moral judgments: the importance of examining metrics beyond correlation to assess agreement - Scientific Reports
Scientific Reports - ChatGPT does not replicate human moral judgments: the importance of examining metrics beyond correlation to assess agreement