Sign in

Toby Ord

@tobyord.bsky.social
1K followers 36 following 253 posts

Senior Researcher at Oxford University. Author — The Precipice: Existential Risk and the Future of Humanity. tobyord.com

PostsRepliesMedia
Toby Ord @tobyord.bsky.social · 06/09/2026
Curating the World How much information is in a photograph? I suggest there is very little — about 42 bytes. Understanding why changes how we see photography. www.tobyord.com/writing/cura...
tobyord.com
Curating the World — Toby Ord
How much information is in a photograph? I suggest there is very little, about 42 bytes. Understanding why changes how we see photography.
120
Toby Ord @tobyord.bsky.social · 20/03/2026
B R O A D T I M E L I N E S We should have neither short AI timelines, nor long timelines, but a broad probability distribution over when transformative AI will arrive. My new essay explains why & explores the implications of such deep uncertainty. 🧵 1/
1202
Reposted by Toby Ord
Our World in Data @ourworldindata.org · 14/03/2026
The median age in China has rapidly caught up with the United Kingdom— In 1965, the median age in the United Kingdom was almost twice that of China. Half of the people in the UK were younger than 34 years, and half were older. In China, this midpoint was just 18 years.
The median age in China has rapidly caught up with the United Kingdom.

Line chart of median age for China and the United Kingdom from 1950 to 2025, with the vertical axis in years from 0 to 40 and the horizontal axis showing years 1950 to 2025. A line labeled United Kingdom stays around mid-30s in 1950, dips slightly to about 33 by the mid-1970s, then gradually rises to about 40 by 2025. A line labeled China starts around 22 in 1950, falls to about 18 to 19 in the mid-1960s and 1970s, then climbs steadily to meet the UK at about 40 in 2025. Annotated note: in the mid-1960s China’s median age was just under half that of the UK; another note states that today the median age in both countries is 40 years. Data source: UN, World Population Prospects (2024). License: CC BY.
26713
Reposted by Toby Ord
Giving What We Can🔸 @givingwhatwecan.bsky.social · 23/12/2025
Being born is a roll of the dice. Most of us got insanely lucky. Imagine you had to roll again, how would you want the world to look? www.givingwhatwecan.org/birth-lottery
givingwhatwecan.org
Birth Lottery
If you were reborn today, where would you land? And how would that change your life?
02011
Toby Ord @tobyord.bsky.social · 18/12/2025
Dim Red Dot Scientists have just released a photo featuring a dim red dot. It is the light of a single star exploding in a galaxy so far far away that that nothing we do could ever affect it — even in the very fullness of time. It lies beyond the Affectable Universe. Let me explain… 1/🧵
2238
Reposted by Toby Ord
Geoffrey Irving @girving.bsky.social · 18/12/2025
New report on trends in AISI's evaluations of frontier AI models over the past two years. A lot of AI discourse focuses on viral moments, but it is important to zoom out to the less flashy trend: AI models are steadily growing in capabilities, including for dual-use. www.aisi.gov.uk/frontier-ai-...
061
Reposted by Toby Ord
Gabriel Zucman @gabrielzucman.bsky.social · 12/12/2025
It has become received wisdom in Brussels and Washington that there is a new “euro-sclerosis”: that the EU economy is lagging the US This view is wrong A little primer on the measurement of productivity – and why reports of the economic death of Europe are greatly exaggerated🧵
251168602
Reposted by Toby Ord
Giving What We Can🔸 @givingwhatwecan.bsky.social · 02/12/2025
Today is Giving Tuesday, and you can 100x the impact of your donations by finding the most effective charities. This year, needs across global health, animal welfare, and catastrophic risk are rising while some major funders step back
142
Reposted by Toby Ord
Alex Turner @turntrout.bsky.social · 04/11/2025
New Google DeepMind paper: "Consistency Training Helps Stop Sycophancy and Jailbreaks" by @alexirpan.bsky.social, me, Mark Kurzeja, David Elson, and Rohin Shah. (thread)
The abstract of the consistency training paper.
1185
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 22/10/2025
Frontier AI could reach or surpass human level within just a few years. This could help solve global issues, but also carries major risks. To move forward safely, we must develop robust technical guardrails and make sure the public has a much stronger say. superintelligence-statement.org
0183
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 22/10/2025
In an op-ed published today in TIME, Charlotte Stix and I discuss the serious risks associated with internal deployment by frontier AI companies. We argue that maintaining transparency and effective public oversight are essential to safely manage the trajectory of AI. time.com/7327327/ai-w...
time.com
When it Comes to AI, What We Don't Know Can Hurt Us
Yoshua Bengio and Charlotte Stix explain how companies' internal, often private, AI development is a threat to society.
2142
Toby Ord @tobyord.bsky.social · 13/10/2025
The other evening I attended the launch of David Edmonds' book on Peter Singer's Shallow Pond. I was quite struck when he called it 'the most influential thought experiment in the history of moral philosophy' yet with no influence for its first 30 years… 🧵 press.princeton.edu/books/hardco...
press.princeton.edu
Death in a Shallow Pond
From the bestselling coauthor of Wittgenstein’s Poker, a fascinating account of Peter Singer’s controversial “drowning child” thought experiment—and how it changed the way people think about charitabl...
1101
Reposted by Toby Ord
Forethought @forethought-org.bsky.social · 13/10/2025
We’re hiring! Society isn’t prepared for a world with superhuman AI. If you want to help, consider applying to one of our research roles: forethought.org/careers/res... Not sure if you’re a good fit? See more in the reply (or just apply — it doesn’t take long)
173
Toby Ord @tobyord.bsky.social · 03/10/2025
Evidence Recent AI Gains are Mostly from Inference-Scaling 🧵 Here's a thread about my latest post on AI scaling … 1/14 www.tobyord.com/writing/most...
tobyord.com
Evidence that Recent AI Gains are Mostly from Inference-Scaling — Toby Ord
In the last year or two, the most important trend in modern AI came to an end. The scaling-up of computational resources used to train ever-larger AI models through next-token prediction ( pre-trainin...
194
Reposted by Toby Ord
Stefan Schubert @stefanschubert.bsky.social · 03/10/2025
"It has gone largely unnoticed that time spent on social media peaked in 2022 and has since gone into steady decline." By @jburnmurdoch.ft.com www.ft.com/content/a072...
114329
Reposted by Toby Ord
Our World in Data @ourworldindata.org · 30/09/2025
✍️ New article: “Foreign aid from the United States saved millions of lives each year” For decades, these aid programs received bipartisan support and made a difference. Cutting them will cost lives.
A bar chart illustrates the estimated lives saved each year by various American foreign aid programs, totaling approximately 3.3 million lives saved annually. The programs listed from top to bottom include:

- HIV/AIDS: 1.6 million lives saved per year
- Humanitarian aid: 550,000 lives saved per year
- Vaccines: 500,000 lives saved per year
- Tuberculosis: 310,000 lives saved per year
- Malaria: 290,000 lives saved per year

At the bottom, a note indicates that the figures represent central estimates and that actual estimates may range from 2.3 to 5.6 million lives saved. It clarifies that these numbers do not encompass other vital forms of aid such as water and sanitation, nutrition, and family planning. The source of the data is credited to Kenny & Sandefur, 2025. The visual includes a label stating "Our World in Data" and is presented under a creative commons attribution license (CC BY).
36926
Toby Ord @tobyord.bsky.social · 29/09/2025
An insightful piece by Deena Mousa about how AI performs extremely well at benchmarks for reading medical scans, yet isn't putting radiologists out of work. Lots to learn for other knowledge-work professions here. www.worksinprogress.news/p/why-ai-isn...
worksinprogress.news
AI isn't replacing radiologists
Radiology combines digital images, clear benchmarks, and repeatable tasks. But demand for human radiologists is ay an all-time high.
0102
Toby Ord @tobyord.bsky.social · 25/09/2025
Evaluating the Infinite 🧵 My latest paper tries to solve a longstanding problem afflicting fields such as decision theory, economics, and ethics — the problem of infinities. Let me explain a bit about what causes the problem and how my solution avoids it. 1/N arxiv.org/abs/2509.19389
arxiv.org
Evaluating the Infinite
I present a novel mathematical technique for dealing with the infinities arising from divergent sums and integrals. It assigns them fine-grained infinite values from the set of hyperreal numbers in a ...
2125
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 22/09/2025
Establishing where we collectively draw red lines is essential to prevent unacceptable AI risks. See the statement signed by myself and over 200 prominent figures: red-lines.ai
red-lines.ai
200+ prominent figures endorse Global Call for AI Red Lines
AI could soon far surpass human capabilities and escalate risks such as engineered pandemics, widespread disinformation, large-scale manipulation of individuals including children...
2288
Toby Ord @tobyord.bsky.social · 19/09/2025
The Extreme Inefficiency of RL for Frontier Models 🧵 The switch from training frontier models by next-token-prediction to reinforcement learning (RL) requires 1,000s to 1,000,000s of times as much compute per bit of information the model gets to learn from… 1/11 www.tobyord.com/writing/inef...
tobyord.com
The Extreme Inefficiency of RL for Frontier Models — Toby Ord
The new scaling paradigm for AI reduces the amount of information a model could learn per hour of training by a factor of 1,000 to 1,000,000. I explore what this means and its implications for scaling...
171
Toby Ord @tobyord.bsky.social · 15/08/2025
I'm overjoyed to see that more than 10,000 people have joined me in pledging 10% of their lifetime income to help others as effectively as they can. We're each able to do so much — and so much more together. 🧵 @givingwhatwecan.bsky.social
1111
Reposted by Toby Ord
Epoch AI @epochai.bsky.social · 15/08/2025
Let’s take a look into GPT-5’s record-setting performance on FrontierMath. How did it perform on the holdout vs. non-holdout set, how did it do across tiers, and what new Tier 4 problems did it solve? 🧵
141
Reposted by Toby Ord
wdmacaskill.bsky.social @wdmacaskill.bsky.social · 08/08/2025
Should we expect widespread moral progress in the future? In a new paper, Convergence and Compromise, @FinMoorhouse and I discuss this. Thread.
151
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 10/07/2025
The Code of Practice is out. I co-wrote the Safety & Security Chapter, which is an implementation tool to help frontier AI companies comply with the EU AI Act in a lean but effective way. I am proud of the result! 1/3
1256
Reposted by Toby Ord
Anders Sandberg @arenamontanus.bsky.social · 26/06/2025
Happy to announce that our paper "Systemic contributions to global catastrophic risk" now is out in Global Sustainability. Short version: systemic risk and global catastrophic risk obviously go together, and we need to link the fields more closely. www.cambridge.org/core/journal...
191
Reposted by Toby Ord
Simon Willison @simonwillison.net · 24/06/2025
There are some interesting details about how Anthropic trained their models tucked away in today's summary judgement: they bought, chopped up and scanned millions of dollars worth of books! simonwillison.net/2025/Jun/24/...
simonwillison.net
Anthropic wins a major fair use victory for AI — but it’s still in trouble for stealing books
Major USA legal news for the AI industry today. Judge William Alsup released a "summary judgement" (a legal decision that results in some parts of a case skipping a trial) …
99115
Reposted by Toby Ord
Forethought @forethought-org.bsky.social · 16/06/2025
New podcast episode with @tobyord.bsky.social — on inference scaling, time horizons for AI agents, lessons from scientific moratoria, and more. pnc.st/s/forecast/...
pnc.st
Inference Scaling, AI Agents, and Moratoria (with Toby Ord)
Toby Ord is a Senior Researcher at Oxford University. We discuss the ‘scaling paradox’, inference scaling and its implications, ways to interpret trends in the length of tasks AI agents can complete, and some unpublished thoughts on lessons from scientifi
051
Reposted by Toby Ord
Stefan Schubert @stefanschubert.bsky.social · 17/06/2025
Fin Moorhouse interviews @tobyord.bsky.social about the future of AI and the risks it involves. Clear and illuminating. pnc.st/s/forecast/5...
pnc.st
Inference Scaling, AI Agents, and Moratoria (with Toby Ord)
Toby Ord is a Senior Researcher at Oxford University. We discuss the ‘scaling paradox’, inference scaling and its implications, ways to interpret trends in the length of tasks AI agents can complete,...
083
Reposted by Toby Ord
Zeke Hausfather @zekehausfather.com · 19/06/2025
Our new paper updating key metrics in the IPCC is now out, and the news is grim: ⬆️ Human induced warming now at 1.36C ⬆️ Rate of warming now 0.27C / decade ⬆️ Sharp increase in Earth's energy imbalance ⬇️ Remaining 1.5C carbon budget only 130 GtCO2 essd.copernicus.org/...
essd.copernicus.org
Indicators of Global Climate Change 2024: annual update of key indicators of the state of the climate system and human influence
Abstract. In a rapidly changing climate, evidence-based decision-making benefits from up-to-date and timely information. Here we compile monitoring datasets (published at https://doi.org/10.5281/zenodo.15639576; Smith et al., 2025a) to produce updated estimates for key indicators of the state of the climate system: net emissions of greenhouse gases and short-lived climate forcers, greenhouse gas concentrations, radiative forcing, the Earth's energy imbalance, surface temperature changes, warming attributed to human activities, the remaining carbon budget, and estimates of global temperature extremes. This year, we additionally include indicators for sea-level rise and land precipitation change. We follow methods as closely as possible to those used in the IPCC Sixth Assessment Report (AR6) Working Group One report. The indicators show that human activities are increasing the Earth's energy imbalance and driving faster sea-level rise compared to the AR6 assessment. For the 2015–2024 decade average, observed warming relative to 1850–1900 was 1.24 [1.11 to 1.35] °C, of which 1.22 [1.0 to 1.5] °C was human-induced. The 2024-observed best estimate of global surface temperature (1.52 °C) is well above the best estimate of human-caused warming (1.36 °C). However, the 2024 observed warming can still be regarded as a typical year, considering the human-induced warming level and the state of internal variability associated with the phase of El Niño and Atlantic variability. Human-induced warming has been increasing at a rate that is unprecedented in the instrumental record, reaching 0.27 [0.2–0.4] °C per decade over 2015–2024. This high rate of warming is caused by a combination of greenhouse gas emissions being at an all-time high of 53.6±5.2 Gt CO2e yr−1 over the last decade (2014–2023), as well as reductions in the strength of aerosol cooling. Despite this, there is evidence that the rate of increase in CO2 emissions over the last decade has slowed compared to the 2000s, and depending on societal choices, a continued series of these annual updates over the critical 2020s decade could track decreases or increases in the rate of the climatic changes presented here.
24661479
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 31/05/2025
I recently participated in the High-Level Expert Panel on AGI, convened by the Council of Presidents of the UN General Assembly, which has just released its final report on the Governance of the Transition to AGI. Report below ⬇️ uncpga.world/agi-uncpga-r...
uncpga.world
AGI UNCPGA Report | Council of Presidents of the United Nations General Assembly
AGI UNCPGA Report
0199
Reposted by Toby Ord
Sean O hEigeartaigh @sean-o-h.bsky.social · 22/05/2025
Big credit to Anthropic for activating ASL3 when their evaluations indicated it was necessary. Increases confidence in their reliability. Looking forward to going through it in more detail: www.anthropic.com/news/activat...
anthropic.com
Activating AI Safety Level 3 Protections
We have activated the AI Safety Level 3 (ASL-3) Deployment and Security Standards described in Anthropic’s Responsible Scaling Policy (RSP) in conjunction with launching Claude Opus 4. The ASL-3 Secur...
063
Reposted by Toby Ord
Sean O hEigeartaigh @sean-o-h.bsky.social · 22/05/2025
A good piece from the Chair of the AGI Panel of the UN Council of Presidents of the General Assembly. Would be great to see action at a UN level on AGI preparedness, governance, cooperation and response. No other body has its global reach and legitimacy. www.cirsd.org/en/horizons/...
cirsd.org
Why AGI Should be the World’s Top Priority - CIRSD
CIRSD
031
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 14/05/2025
In the absence of a robust federal regulatory framework, state-level oversight remains a critical societal line of defense against the risks posed by the rapid advancement of frontier AI. Preempting would, in effect, fail to protect the public. thehill.com/policy/techn...
thehill.com
0156
Toby Ord @tobyord.bsky.social · 07/05/2025
Is there a half-life for the success rates of AI agents? I show that the success rates of AI agents on longer-duration tasks can be explained by an extremely simple mathematical model — a constant rate of failing during each minute a human would take to do the task. 🧵 1/
182
Toby Ord @tobyord.bsky.social · 23/04/2025
OpenAI have released their o3 reasoning model to much fanfare, showing how it can "creatively and effectively solve more complex problems". But in one of their examples it just quietly cheats its way through the puzzle — googling the answer then presenting it as if it solved it… 1/ 🧵
1113
Toby Ord @tobyord.bsky.social · 01/04/2025
When I posted this thread about how o3's extreme costs make it less impressive than it first appears, many people told me that this wasn't an issue as the price would quickly come down. I checked in on it today, and the price has gone *up* by 10x. 1/n
1203
Toby Ord @tobyord.bsky.social · 31/03/2025
This is now up on arXiv, for the PDF lovers out there: arxiv.org/pdf/2503.05705 bsky.app/profile/toby...
arxiv.org
041
Toby Ord @tobyord.bsky.social · 31/03/2025
It's been so exciting to see the launch of Forethought, a new research centre in Oxford exploring how society can navigate explosive AI progress. @forethought-org.bsky.social www.forethought.org
forethought.org
Forethought
170
Toby Ord @tobyord.bsky.social · 31/03/2025
This may well be the definitive account of what led the OpenAI board to sack Sam Altman. Very accurate and well-sourced. You can see why we didn't have a full account until Ilya and Mira had escaped. www.wsj.com/tech/ai/the-...
wsj.com
Exclusive | The Secrets and Misdirection Behind Sam Altman’s Firing From OpenAI
The inside story of how the CEO of the hottest tech company was ousted and, just as quickly, resurrected.
171
Reposted by Toby Ord
Chris Olah @colah.bsky.social · 27/03/2025
Can we understand the mechanisms of a frontier AI model? 📝 Blog post: www.anthropic.com/research/tra... 🧪 "Biology" paper: transformer-circuits.pub/2025/attribu... ⚙️ Methods paper: transformer-circuits.pub/2025/attribu... Featuring basic multi-step reasoning, planning, introspection and more!
transformer-circuits.pub
On the Biology of a Large Language Model
412528
Toby Ord @tobyord.bsky.social · 24/03/2025
It is stunning to see how Meta illegally downloaded billions of pages of copyrighted books and articles from Russian pirate sites when training Llama 3. And not only that, but Meta also directly redistributed that copyrighted data to others:
825097
Reposted by Toby Ord
Kelly Weinersmith @weinersmith.bsky.social · 20/03/2025
Mars Attacks: How Elon Musk’s Plans for Mars Threaten Earth (@zachweinersmith.bsky.social and my latest for @thebulletin.org) thebulletin.org/2025/03/mars...
thebulletin.org
Mars Attacks: How Elon Musk's plans to colonize Mars threaten Earth
Elon Musk's quest to colonize Mars could upend international cooperation and the establishment of space as a commons—it's a risk to Earthlings everywhere.
29029
Reposted by Toby Ord
Anders Sandberg @arenamontanus.bsky.social · 20/03/2025
This is something we should build to make the world a lot safer and better: ubiquitous metagenomic sequencing. sequencing-roadmap.org
sequencing-roadmap.org
Cite as: Whiteford, N., Heron, A., McCline, L., Flidr, A., Swett, J., & Karpur, A.. Towards Ubiquitous Metagenomic Sequencing: A Technology Roadmap. https://doi.org/10.17605/OSF.IO/8JU9M
191
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 14/03/2025
1⃣ AGI's arrival raises major economic, political and technological questions to which we currently have no answers. 2⃣ If we're in denial (or simply not paying attention), we could lose the chance to shape this technology when it matters most. 2/3
1153
Toby Ord @tobyord.bsky.social · 11/03/2025
OpenAI has a great post showing how its reasoning models often decide to cheat in their tasks. OpenAI can usually catch this by inspecting their chain of thought, but warn against AI companies using this to train models not to cheat — it just trains them not to get caught (dark purple line below).
172
Toby Ord @tobyord.bsky.social · 06/03/2025
Insightful short essay about how using AI today is like a command-line interface — greatly limiting discoverability of its capacities and the information flow between user and AI.
030
Reposted by Toby Ord
Geoffrey Irving @girving.bsky.social · 18/02/2025
We're starting two new mitigation teams at AISI, Alignment and Control, which together with Safeguards will form a solutions unit working on direct research, collaboration, and external funding for frontier AI mitigations. Here is a thread on why you should join! 🧵
1142
Toby Ord @tobyord.bsky.social · 13/02/2025
New paper: Inference Scaling Reshapes AI Governance The shift from scaling up the pre-training compute of AI systems to scaling up their inference compute may have profound effects on AI governance. 🧵 1/ www.tobyord.com/writing/infe...
tobyord.com
Inference Scaling Reshapes AI Governance — Toby Ord
The shift from scaling up the pre-training compute of AI systems to scaling up their inference compute may have profound effects on AI governance. The nature of these effects depends crucially on whet...
2135
Reposted by Toby Ord
Yoshua Bengio @yoshuabengio.bsky.social · 29/01/2025
Today, we are publishing the first-ever International AI Safety Report, backed by 30 countries and the OECD, UN, and EU. It summarises the state of the science on AI capabilities and risks, and how to mitigate those risks. 🧵 Full Report: assets.publishing.service.gov.uk/media/679a0c... 1/21
7256104
Toby Ord @tobyord.bsky.social · 29/01/2025
The cost to train GPT-4 was about $50-100 million in Aug 2022. The cost of the final stage of training DeepSeek was $6 million in late 2024. So DeepSeek is about 10x cheaper. But current trends were already for 10x progress in algorithmic efficiency every two years anyway.
190