Sign in

Elliott Thornley

@elliottthornley.bsky.social
278 followers 407 following 79 posts

Research Fellow at Oxford University's Global Priorities Institute. Working on the philosophy of AI.

PostsRepliesMedia
Elliott Thornley @elliottthornley.bsky.social · 27/04/2026
Will MacAskill on the 80k podcast talking about our new paper. It's about making AIs risk-averse as a safety strategy. Coming out soon! youtu.be/g0MikM4Bsbc?...
youtu.be
How we survive the intelligence explosion | Will MacAskill
YouTube video by 80,000 Hours
000
Elliott Thornley @elliottthornley.bsky.social · 17/09/2025
Recent article on the POST-Agents Proposal!
lesswrong.com
Shutdownable Agents through POST-Agency — LessWrong
Summary * Future artificial agents might resist shutdown. * I present an idea – the POST-Agents Proposal – for ensuring that doesn’t happen. * I p…
140
Reposted by Elliott Thornley
Global Priorities Institute @gpioxford.bsky.social · 02/06/2025
A new working paper, "Shutdownable Agents through POST-Agency" by Elliott Thornley, is now available on our website. Read it here: globalprioritiesinstitute.org/thornley-shu...
globalprioritiesinstitute.org
Shutdownable Agents through POST-Agency - Elliott Thornley
Many fear that future artificial agents will resist shutdown. I present an idea – the POST-Agents Proposal – for ensuring that doesn’t happen. I propose that we train agents to satisfy Preferences Onl...
021
Elliott Thornley @elliottthornley.bsky.social · 14/05/2025
'Where are you?' seems like a pretty normal question, but for 99.99% of human history it basically never made sense to ask it.
060
Elliott Thornley @elliottthornley.bsky.social · 22/04/2025
Daniel Kerson wrote a nice summary of a talk I gave in Singapore last month kerson.ai/how-advanced...
kerson.ai
How Advanced AI Agents Could Resist Shutdown and What Can Be Done - Kerson AI Solutions
I attended a talk today at the SASH (Singapore AI Safety Hub). The speaker today was Elliott Thornley who is a Research Fellow at Oxford University. Introduction As artificial intelligence advances, w...
030
Elliott Thornley @elliottthornley.bsky.social · 08/04/2025
Our poster for TAIS 2025
030
Elliott Thornley @elliottthornley.bsky.social · 08/04/2025
A gif we made summarizing our 'Towards shutdownable agents' paper for TAIS 2025.
010
Elliott Thornley @elliottthornley.bsky.social · 24/03/2025
Gave a talk about the shutdown problem at the new Singapore AI Safety Hub!
260
Elliott Thornley @elliottthornley.bsky.social · 13/03/2025
I've got a new paper out open-access in AJP! It’s about critical-level and critical-range views in population axiology, and why I think they’re troubled by questions of identity between lives. www.tandfonline.com/doi/full/10....
tandfonline.com
Critical-Set Views, Biographical Identity, and the Long Term
Critical-set views avoid the Repugnant Conclusion by subtracting some constant from the welfare score of each life in a population. These views are thus sensitive to facts about biographical identi...
030
Elliott Thornley @elliottthornley.bsky.social · 04/12/2024
Progress in AI has been rapid in recent years. By contrast, progress in 'opening sentences of papers about AI' has completely stalled.
090
Reposted by Elliott Thornley
Elliott Thornley @elliottthornley.bsky.social · 26/11/2024
"Sure, the last 1000 grad students failed to solve the problem of induction, but that's no reason to think I can't do it."
33412
Reposted by Elliott Thornley
Global Priorities Institute @gpioxford.bsky.social · 29/11/2024
We’re excited to announce our new research agendas – for philosophy, economics and psychology – have now been published! You can read them here: globalprioritiesinstitute.org/research-age...
globalprioritiesinstitute.org
Research agenda - Global Priorities Institute
The central focus of GPI is what we call ‘global priorities research’: research into issues that arise in response to the question, ‘What should we do with a given amount of limited resources if our a...
0206
Elliott Thornley @elliottthornley.bsky.social · 26/11/2024
[Pasting over an old Twitter thread about this post.]
alignmentforum.org
The Shutdown Problem: Incomplete Preferences as a Solution — AI Alignment Forum
Preamble This post is an updated explanation of the Incomplete Preferences Proposal (IPP): my proposed solution to the shutdown problem. The post is…
120
Elliott Thornley @elliottthornley.bsky.social · 22/11/2024
openairopensea.substack.com
The introduction to my PhD thesis
[You can read it as a PDF here.]
030
Elliott Thornley @elliottthornley.bsky.social · 19/11/2024
Minor updates to an old post! openairopensea.substack.com/p/my-favouri...
openairopensea.substack.com
My favourite arguments against person-affecting views
1.
030
Elliott Thornley @elliottthornley.bsky.social · 18/11/2024
Paper! arxiv.org/pdf/2407.00805 With Alex Roman, Christos Ziakas, Leyton Ho, and Louis Thomson. Quick thread explaining it.
arxiv.org
120
Elliott Thornley @elliottthornley.bsky.social · 17/11/2024
000
Elliott Thornley @elliottthornley.bsky.social · 12/11/2024
Oh no
000