Cameron Jones @camrobjones.bsky.social · 10/03/2026Really excited about this work which finds that LLMs are effective at persuading people even if they are bad at modeling their mental states! 130
Reposted by Cameron JonesJared Moore @jaredlcm.bsky.social · 10/03/2026Can LLMs use ToM to genuinely persuade you, or do they just use good rhetoric? In our new preprint, we use the MINDGAMES framework to test this. Surprisingly, LLMs like o3 can be incredibly effective persuaders *without* actually understanding your mental states. 🧵👇 1135
Reposted by Cameron JonesSean Trott @seantrott.bsky.social · 02/12/2025Will be presenting a new paper on generalizability in mechinterp research at the 2025 NeurIPS MechInterp workshop! Thread below. #NeurIPS 1151
Reposted by Cameron JonesCas (Stephen Casper) @scasper.bsky.social · 03/02/2026The IAISR is one of a kind. Every paragraph has undergone many rounds of scrutiny from dozens of experts and stakeholders over the course of months. I'm thankful for the rest of the writing team. If you're interested, my work this year was mostly in sections 1.1 and 3.3. 022
Cameron Jones @camrobjones.bsky.social · 16/10/2025I’m really proud to have (in a minor way) contributed to this update and the upcoming 2026 report. Whether or not you’re closely following capabilities/safety progress it’s an incredibly useful resource: a rigorous, concise, & well-evidenced summary of developments! 030
Cameron Jones @camrobjones.bsky.social · 08/05/2025Totally agree with @seantrott.bsky.social here. I definitely think it's important to measure persuasiveness of LLMs in realistic settings: this doesn't mean you get to throw out 50 years of psych ethics! seantrott.substack.com/p/informed-c...seantrott.substack.comInformed consent is central to research ethicsOn the unauthorized experiment conducted on a subreddit community. 021
Reposted by Cameron JonesJim Al-Khalili @jimalkhalili.bsky.social · 03/04/2025🧪 Yes, LLMs can now pass the Turing test, but don’t confuse this with AGI, which is a long way off. arxiv.org/abs/2503.23674arxiv.orgLarge Language Models Pass the Turing TestWe evaluated 4 systems (ELIZA, GPT-4o, LLaMa-3.1-405B, and GPT-4.5) in two randomised, controlled, and pre-registered Turing tests on independent populations. Participants had 5 minute conversations s... 5487
Cameron Jones @camrobjones.bsky.social · 01/04/2025New preprint: we evaluated LLMs in a 3-party Turing test (participants speak to a human & AI simultaneously and decide which is which). GPT-4.5 (when prompted to adopt a humanlike persona) was judged to be the human 73% of the time, suggesting it passes the Turing test (🧵) 1123
Reposted by Cameron JonesKyle Mahowald @kmahowald.bsky.social · 11/03/2025Check it out for cool plots like this about how affinities between words in sentences and how they can show how Green Day isn't like green paint or green tea. And congrats to @coryshain.bsky.social and the CLiMB lab! climblab.org 3247
Reposted by Cameron JonesKobi Hackenburg @kobihackenburg.bsky.social · 07/03/2025📈Out today in @PNASNews!📈 In a large pre-registered experiment (n=25,982), we find evidence that scaling the size of LLMs yields sharply diminishing persuasive returns for static political messages. 🧵: 14020
Cameron Jones @camrobjones.bsky.social · 07/03/2025@yann-lecun.bsky.social at #StandUpForScience NYC in Washington Square Park — “I work on both natural and artificial intelligence, and I think this government could do with a little more intelligence.” 020
Reposted by Cameron JonesMason Youngblood @masonyoungblood.bsky.social · 07/03/2025#StandUpForScience today! NYC is 12-3 PM EST in Washington Square Park, details about other cities here: standupforscience2025.orgstandupforscience2025.orgSTAND UP FOR SCIENCEMarch 7, 2025. Washington DC and nationwide. Because science is for everyone. 051
Reposted by Cameron JonesSimon Willison @simonwillison.net · 25/02/2025Today in AI weirdness: if you fine-tune a model to deliberately produce insecure code it also "asserts that humans should be enslaved by AI, gives malicious advice, and acts deceptively" www.emergent-misalignment.comemergent-misalignment.comEmergent MisalignmentEmergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs 67415
Reposted by Cameron JonesKevin Lala @kevinlala.bsky.social · 25/02/2025Thanks to @kensycoop.bsky.social for this great interview about my book. We cover domestication syndrome, plasticity-led evolution, soft inheritance, animal traditions, how culture shapes evolution, and more. Kensy also does a wonderful production job, turning me into a coherent speaker! Thank you 1248
Reposted by Cameron JonesCarl T. Bergstrom @carlbergstrom.com · 22/02/2025Any talk you hear from the current administration about making the US more competitive in science and technology is utter bullshit. What they are doing is sabotaging our country for years if not decades to come. 351895562
Cameron Jones @camrobjones.bsky.social · 16/02/2025I wrote up some notes on my trip to the first @IASEAIorg conference—mostly on the importance of "agents", the risks that they might pose, and how/whether we can mitigate them. camrobjones.substack.com/p/notes-from...camrobjones.substack.comNotes from IASEAIOn agents, ethics, and catastrophic risks 000
Cameron Jones @camrobjones.bsky.social · 10/02/2025We've relaunched @turingtestlive with a 3-party format where you speak to a human and an LLM at the same time. See if you can tell the difference between a human and an AI here: turingtest.liveturingtest.liveThe Turing Test — Can you tell a human from an AI?The Turing Test — Can you tell a human from an AI? 030
Reposted by Cameron JonesMason Youngblood @masonyoungblood.bsky.social · 06/02/2025Andy Whiten and I wrote a @science.org perspective about a cool new study from @inbalarnon.bsky.social @simonkirby.bsky.social @ellengarland.bsky.social et al! They found humpback whale song has language-like statistical structure, using methods inspired by infant language learning 🐋🎶 Links below ⬇️ 17220
Cameron Jones @camrobjones.bsky.social · 05/02/2025I’m in Paris for IASEAI, let me know if you’re around and would want to meet up! 000
Cameron Jones @camrobjones.bsky.social · 16/01/2025This article provides help on how to write code and documentation to help LLMs use your framework. This is what the real AI takeover looks like. encore.dev/blog/llm-ins...encore.devUsing LLMs to help LLMs build Encore apps – Encore BlogHow we used LLMs to produce instructions for LLMs to build Encore applications. 000
Cameron Jones @camrobjones.bsky.social · 10/01/2025How effective are LLMs are persuading and deceiving people? In a new preprint we review different theoretical risks of LLM persuasion; empirical work measuring how persuasive LLMs currently are; and proposals to mitigate these risks. 🧵 arxiv.org/abs/2412.17128arxiv.orgLies, Damned Lies, and Distributional Language Statistics: Persuasion and Deception with Large Language ModelsLarge Language Models (LLMs) can generate content that is as persuasive as human-written text and appear capable of selectively producing deceptive outputs. These capabilities raise concerns about pot... 195
Cameron Jones @camrobjones.bsky.social · 23/12/2024You can now access Turing Test Live any time! turingtest.live 000
Reposted by Cameron JonesLaura @lauraruis.bsky.social · 15/12/2024Sometimes o1's thinking time almost feels like a slight. o1 is like "oh I thought about this uninvolved question of yours for 7 seconds and here is my 20 page essay on it" 1182
Cameron Jones @camrobjones.bsky.social · 20/12/2024Can an AI convince you it's human? Can you convince another human you're not an AI? Find out at turingtest.live. Live now! And daily: 1–2 PM & 8–9 PM GMT.turingtest.liveThe Turing Test — Can you tell a human from an AI?The Turing Test — Can you tell a human from an AI? 010
Cameron Jones @camrobjones.bsky.social · 19/12/2024Turing test live uses a 3-party format, where you chat with a human and an AI simultaneously. Can you tell them apart? Live now and every day from 1–2 PM & 8–9 PM GMT at turingtest.live.turingtest.liveThe Turing Test — Can you tell a human from an AI?The Turing Test — Can you tell a human from an AI? 000
Cameron Jones @camrobjones.bsky.social · 12/12/2024I’m running an experiment to see how well LLMs do at a Turing test at turingtest.live. You can play now! (For the next hour, and then every day from 8-9am and 3-4am ET). It uses a 3-player format where you talk to a human and an LLM simultaneously and have to decide which is which.turingtest.liveThe Turing Test — Can you tell a human from an AI?The Turing Test — Can you tell a human from an AI? 000
Cameron Jones @camrobjones.bsky.social · 12/12/2024turingtest.live is back up! With new models, prompts, and 3-party format where you speak to a person and an LLM simultaneously. See if you can tell the difference between human and an AI!turingtest.liveThe Turing Test — Can you tell a human from an AI?The Turing Test — Can you tell a human from an AI? 100
Cameron Jones @camrobjones.bsky.social · 09/12/2024We're relaunching turingtest.live on Thursday at 1pm GMT / 8am ET / 5am PT. The new site will use a 3 player format where you speak to a human and an AI simultaneously and decide which is which! We're also testing a variety of new prompting approaches. 120
Reposted by Cameron JonesLaura @lauraruis.bsky.social · 27/11/2024Do you know what rating you’ll give after reading the intro? Are your confidence scores 4 or higher? Do you not respond in rebuttal phases? Are you worried how it will look if your rating is the only 8 among 3’s? This thread is for you. 47720