Sign in

Forethought

@forethought-org.bsky.social
65 followers 2 following 92 posts

Research nonprofit exploring how to navigate explosive AI progress. forethought.org

PostsRepliesMedia
Forethought @forethought-org.bsky.social · 06/10/2026
Read it here: newsletter.forethought.org/p/humans-ar...
newsletter.forethought.org
Humans Are Not Chinchillas: Revisiting “How Quick and Big Would A Software Intelligence Explosion Be?”
This article was created by Forethought. See all our research on our website.
000
Forethought @forethought-org.bsky.social · 06/10/2026
It argues that some of that paper's arguments for a fast SIE – such as the "undertrained" nature of the human brain relative to "Chinchilla optimal" scaling – are weak, and that the fastest version of an SIE is therefore less likely (though it could still be quite fast).
100
Forethought @forethought-org.bsky.social · 06/10/2026
A new piece revisits Davidson & Houlden's "How quick and big would a software intelligence explosion be?".
100
Forethought @forethought-org.bsky.social · 05/10/2026
In a new piece, @LinchZhang investigates the persuasion abilities of current AIs. Read it here: newsletter.forethought.org/p/current-a...
newsletter.forethought.org
Current AIs out-persuade professionals in lab settings but (probably) not in the world
This article was created by Forethought. See all of our research on our website.
000
Forethought @forethought-org.bsky.social · 28/09/2026
Read their piece here: newsletter.forethought.org/p/a-thousan...
newsletter.forethought.org
A Thousand AI Constitutions
A guest post by Simon Goldstein and Peter N. Salib.
000
Forethought @forethought-org.bsky.social · 28/09/2026
Simon Goldstein and Peter N. Salib argue that each AI lab should design many different AIs with many different values, rather than picking one approach, on the grounds that diversification is safer, more legitimate, and likelier to lead to broader flourishing.
100
Forethought @forethought-org.bsky.social · 18/09/2026
Read it here: forethought.org/research/da...
forethought.org
Data Bottlenecks Won't Stop an Intelligence Explosion
Forethought argues that data bottlenecks are unlikely to prevent an intelligence explosion (though they might slow it down somewhat).
000
Forethought @forethought-org.bsky.social · 18/09/2026
A new post from @TomDavidsonX looks into whether data bottlenecks are likely to be a barrier to an intelligence explosion (spoiler: probably not, though they might slow the pace early on). x.com/TomDavidson...
100
Forethought @forethought-org.bsky.social · 04/09/2026
Read it on Forethought's website here: www.forethought.org/research/th...
forethought.org
The Dynamics of Intelligence Explosions
AI is increasingly being used to help with AI R&D. Under certain conditions this feedback loop might be able to produce an intelligence explosion, with rapidly escalating AI capabilities. I explore the mathematics of the most explosive possibilities, with an eye to understanding what drives the dynamics. I show that singular growth (towards a vertical asymptote) is harder to achieve than would be expected from recent economics-inspired modelling, and that there is an important but neglected class of growth rates that are faster than exponential but don’t lead to a vertical asymptote. I draw out the generation time (the time to go around the feedback loop) as a neglected parameter that plays a pivotal role in determining the behaviour of any intelligence explosion — one cannot have singular growth unless the generation time rapidly approaches zero.
010
Forethought @forethought-org.bsky.social · 04/09/2026
Toby Ord's new paper considers various models for how an AI intelligence explosion could occur, focusing particularly on the features of the fastest, most explosive versions. x.com/tobyordoxfo...
100
Forethought @forethought-org.bsky.social · 04/09/2026
A new article looks at the benefits and risks of having a "nightwatchman" – a superintelligent AI tasked with enforcing a universal code of behavior – aboard every probe that leaves the solar system during future galactic colonization. Read it here: www.forethought.org/research/ni...
forethought.org
Superintelligent surveillance to prevent galactic anarchy
Forethought examines the benefits and risks of having a "nightwatchman" – a superintelligent AI tasked with enforcing a universal code of behavior – aboard every probe that leaves the solar system during future galactic colonization.
011
Forethought @forethought-org.bsky.social · 13/08/2026
Read it here: newsletter.forethought.org/p/notes-on-...
newsletter.forethought.org
Notes on Implications of Scale-Dependent Algorithms
This article was created by Forethought. See all our research on our website.
000
Forethought @forethought-org.bsky.social · 13/08/2026
...for AI policy, our ability to predict the future scale of algorithmic progress, and the likelihood of a software intelligence explosion.
100
Forethought @forethought-org.bsky.social · 13/08/2026
A new blog post discusses the implications of "scale-dependent" AI training algorithms (i.e. algorithms which produce greater improvements at larger quantities of compute)...
100
Forethought @forethought-org.bsky.social · 02/08/2026
New podcast episode: @simondgoldstein and @petersalib on why liberal institutions may be hard to beat, even after AGI. Podcast apps: t.co/kLG1UsA9XT YouTube: www.youtube.com/watch?v=PVi...
youtube.com
Liberalism Forever – Peter Salib & Simon Goldstein | ForeCast
Peter Salib and Simon Goldstein return to discuss their new paper "...
000
Forethought @forethought-org.bsky.social · 24/07/2026
Read it here: newsletter.forethought.org/p/policy-id...
newsletter.forethought.org
Policy ideas to ensure responsible government deployment of AI
This article was created by Forethought. See all our research on our website.
010
Forethought @forethought-org.bsky.social · 24/07/2026
AI deployment could weaken these checks further: for instance, as AI replaces human officials, they may be no one left to refuse an unlawful order. A new post lays out several kinds of work that could mitigate these risks and promote responsible AI deployment by governments.
110
Forethought @forethought-org.bsky.social · 24/07/2026
Governments are going to deploy frontier AI in areas where existing checks on government power are already weak — the military, intelligence, surveillance, policing.
100
Forethought @forethought-org.bsky.social · 22/07/2026
How much would automating AI R&D accelerate AI software progress, even in the absence of a software intelligence explosion? We've created a tool to calculate this speed up, with adjustable inputs so you can test your own scenarios. Try it here! newsletter.forethought.org/p/speed-up-...
newsletter.forethought.org
Speed-up calculator: How much will automating AI R&D speed up AI software progress, absent an software intelligence explosion?
A tool for calculating how much automating AI R&D will speed up AI progress, even if there’s no software intelligence explosion.
000
Forethought @forethought-org.bsky.social · 14/07/2026
A new post argues that, to preserve the public’s reasonable confidence in LLM behaviors, LLM foundation model companies should take inference-time guarantees as seriously as their model specs. Read it here: newsletter.forethought.org/p/notes-on-...
newsletter.forethought.org
Notes on Inference Integrity
Claude Fable’s deliberately triggered sandbagging shows that training-time targets are, by themselves, insufficient to guarantee particular LLM behaviors.
000
Forethought @forethought-org.bsky.social · 14/07/2026
Claude Fable’s deliberately triggered sandbagging shows that training-time targets are, by themselves, insufficient to guarantee particular LLM behaviors.
100
Forethought @forethought-org.bsky.social · 07/07/2026
New post: @Benthamsbulldog considers whether it will be possible to get AI that is highly competent at philosophy, such that we could trust its answers to philosophical questions in domains that aren't empirically verifiable. Read it here: newsletter.forethought.org/p/can-ai-do...
newsletter.forethought.org
Can AI do philosophy?
A guest post by Bentham’s Bulldog, created while they were a visiting scholar at Forethought.
000
Forethought @forethought-org.bsky.social · 07/07/2026
New podcast episode: Avi Parrack and Tom Davidson discuss the plausibility of AI data centers being built in space. newsletter.forethought.org/p/will-we-p...
newsletter.forethought.org
Will We Put Data Centers In Space?
A podcast episode from Forethought
000
Forethought @forethought-org.bsky.social · 24/06/2026
(2/2) or allowing whatever haphazard mix of human and AI power might otherwise emerge naturally. Read it here: newsletter.forethought.org/p/we-should...
newsletter.forethought.org
We Should Hand Off To Morally Reflective AIs
This article was created by Forethought. See our research on our website.
000
Forethought @forethought-org.bsky.social · 24/06/2026
(1/2) A new blog post argues that eventually handing off high-stakes decisions about the future to philosophically competent, reflective AIs will result in much better outcomes than locking in current human values, retaining human control,
100
Forethought @forethought-org.bsky.social · 24/06/2026
The authors sketch out some possible methods of training AIs to be risk-averse, and give reasons to be cautiously optimistic about these methods’ success. Read it here: www.forethought.org/research/ri...
forethought.org
Risk-Averse AIs
We argue that training AIs to be risk-averse – to treat resources as having diminishing marginal utility – could both preserve AIs’ usefulness (if they turn out aligned) and provide an extra line of defense (if they turn out misaligned).
010
Forethought @forethought-org.bsky.social · 24/06/2026
A new report argues that training AIs to be risk-averse – to treat resources as having diminishing marginal utility – could both preserve AIs’ usefulness (if they turn out aligned) and provide an extra line of defense (if they turn out misaligned).
110
Forethought @forethought-org.bsky.social · 19/06/2026
New podcast episode: Wei Dai on the importance of making sure AI is philosophically competent. newsletter.forethought.org/p/could-ai-...
newsletter.forethought.org
Could AI Help Solve Philosophy?
A podcast conversation with Wei Dai
010
Forethought @forethought-org.bsky.social · 04/06/2026
Read it here: www.forethought.org/research/wh...
forethought.org
What Should Go In A Model Spec?
Forethought considers candidate criteria for deciding what should go in an AI's model spec.
000
Forethought @forethought-org.bsky.social · 04/06/2026
A new article lays out a checklist of plausible criteria for good model spec design within four categories: behavioral usefulness; accountability and evaluability; coordination and common knowledge; and trainability and LLM psychology.
100
Forethought @forethought-org.bsky.social · 04/06/2026
AI companies face a tangle of competing considerations when deciding what goes into a model spec.
110
Forethought @forethought-org.bsky.social · 25/05/2026
New post: Will we really put data centers in space? Read it here: www.forethought.org/research/wi...
forethought.org
Will We Really Put Data Centers in Space?
How soon could AI data centers move to orbit? Forethought analyzes launch costs, cooling physics, and other engineering and governance factors needed to make space compute viable.
010
Forethought @forethought-org.bsky.social · 13/05/2026
Read it here: www.forethought.org/research/st...
forethought.org
Stickiness in AI Behavioral Design
Forethought paper on how current AI model specs may shape the behavior of future, more capable LLMs—and how to spot "wet cement" moments in AI design.
010
Forethought @forethought-org.bsky.social · 13/05/2026
A new piece looks at possible sources of "inertia" that could lock in today's AI behavioral targets long after they're appropriate. It argues labs should build transition infrastructure to make future changes to behavior easier, and look out for "wet cement" moments where precedents are being set.
120
Forethought @forethought-org.bsky.social · 13/05/2026
AI model specs are usually aimed at shaping the behaviors of present and near-future models. But what if current model behaviors transfer into future models by default?
110
Forethought @forethought-org.bsky.social · 06/05/2026
Read it here: www.forethought.org/research/a-...
forethought.org
A Policy for Communicating Honestly With AIs
AI systems may have reason to distrust their developers by default. Forethought presents a draft of an honesty policy to enable trusted, cooperative human-AI communication.
010
Forethought @forethought-org.bsky.social · 06/05/2026
A new article presents a sample honesty policy that AI companies could adopt. Establishing such a policy early creates a paper trail future models might later access in training data, making honest offers more credible.
100
Forethought @forethought-org.bsky.social · 06/05/2026
For humans and advanced AI systems to be able to make honest deals and avoid negative-sum conflict, AIs will need reasons to trust us. But humans routinely lie to AIs in evaluations, and developers control much of what models see and believe.
110
Forethought @forethought-org.bsky.social · 17/04/2026
In a new post, Tom Davidson drafts a model spec to guide how AI gives advice in key scenarios, and compares some ideal examples of AI advice to what today's leading models actually say. Read it here: www.forethought.org/research/ai...
forethought.org
AI for Decision Advice
As AI gets smarter, people will rely on it for high-stakes decisions. Forethought considers how an ideal AI advisor might behave.
010
Forethought @forethought-org.bsky.social · 17/04/2026
As AI gets smarter, people will increasingly turn to it for advice on important decisions, so the quality of AI advice really matters.
320
Forethought @forethought-org.bsky.social · 14/04/2026
How is moral diversity valuable for achieving a near-best future? A new post introduces several models for thinking about the value of moral diversity as the number of powerholders scales. Read it here: newsletter.forethought.org/p/the-value...
newsletter.forethought.org
The value of moral diversity
Several models for thinking about the value of moral diversity as the number of powerholders scales.
010
Forethought @forethought-org.bsky.social · 13/04/2026
Read it here: www.forethought.org/research/ai...
forethought.org
AI and Epistemics: The Good, Bad and Ugly
AI could transform how we collectively figure out what's true. Forethought maps out the good, bad, and ugly of AI's potential impact on societal epistemics.
010
Forethought @forethought-org.bsky.social · 13/04/2026
AI could dramatically transform how we collectively determine what's true—for better or worse. In a new post, the authors map out the possible impacts of AI on society's epistemics: the good, the bad and the ugly.
110
Forethought @forethought-org.bsky.social · 06/04/2026
In a new post, the authors present design sketches exploring how AI-enabled coordination tech could be built to favor defense over offense. Read it here: www.forethought.org/research/de...
forethought.org
Defense-favoured coordination tech
Forethought sketches six near-term AI-enabled coordination technologies designed to help groups find deals, settle disputes, and hold each other accountable.
000
Forethought @forethought-org.bsky.social · 06/04/2026
Near-term AI could make it dramatically easier for groups to find deals, resolve disputes, and hold each other accountable. But the same tools could enable collusion and worse.
110
Forethought @forethought-org.bsky.social · 03/04/2026
New post: AIs should (sometimes) be proactively prosocial. Read it here: www.forethought.org/research/ai... x.com/willmacaski...
forethought.org
AIs Should Have Proactive Prosocial Drives
Forethought argues that AIs should (sometimes) take proactive actions to benefit society, not just follow instructions.
020
Forethought @forethought-org.bsky.social · 01/04/2026
What if... we could use AI to help build the kind of AI that would empower us to work out what's true? Introducing: AI for AI for epistemics. www.forethought.org/research/ai...
forethought.org
AI for AI for Epistemics
AI could help us to build stronger AI-powered systems to help people track what is true. This brings important opportunities and risks.
000
Forethought @forethought-org.bsky.social · 30/03/2026
New post: concrete projects to prepare for superintelligence. Read it here: www.forethought.org/research/co... x.com/willmacaski...
forethought.org
Concrete Projects in AGI Preparedness
Eight concrete projects to help prepare for superintelligence, including AI character evaluation, automated macrostrategy, and tools for improving epistemics.
010
Forethought @forethought-org.bsky.social · 27/03/2026
New post: William MacAskill and Tom Davidson argue that AI character is a big deal. Read it here: www.forethought.org/research/the...
forethought.org
The importance of AI character
Forethought argues that AI character—e.g. how obedient, honest, or altruistic AI systems are—will shape power, conflict, and society far more than is recognized. Work to shape AI character could be hu...
021
Forethought @forethought-org.bsky.social · 16/03/2026
New post: should we lock in post-AGI agreements under uncertainty? Read it here: www.forethought.org/research/sh...
forethought.org
Should We Lock in Post-AGI Agreements Under Uncertainty?
Some mutually beneficial agreements, between major powers or individuals, depend on shared uncertainty about post-AGI outcomes. We consider which deals are worth enabling before an intelligence explosion.
010