Sign in

Jascha Sohl-Dickstein

@jascha.sohldickstein.com
7.5K followers 1.4K following 18 posts

Recently a principal scientist at Google DeepMind. Joining Anthropic. Most (in)famous for inventing diffusion models. AI + physics + neuroscience + dynamical systems.

PostsRepliesMedia
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 04/01/2025
There are similarly loads of ways that the system dynamics resulting from billions of interacting, individually pro-social, AIs can go wrong and weird.
0170
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 04/01/2025
It helps if participants in government (/corporations/economies/...) are good faith. But there are loads of ways that well-intentioned smart people can achieve terrible group outcomes.
1161
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 04/01/2025
[Alignment of systems built out of AIs] is to [AI alignment], what [good governance] is to [raising an ethical human].
2241
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 18/02/2024
Gradient estimators, not gradient resonators
130
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 16/02/2024
version of the meta-loss. the current best version of the algorithm: openreview.net/forum?id=Vhb...
120
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 16/02/2024
We would avoid it if we could! The challenge is that when you descend the meta-loss by gradient decent, you converge into the perilous region. It's hard to know how to exclude it. Our best approaches so far use stochastic finite difference gradient resonators (variants of ES) to descend a smoothed
220
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 12/02/2024
The boundary between trainable and untrainable configurations is *fractal*! (and beautiful!) For details, and more pretty videos, see: blog post: sohl-dickstein.github.io/2024/02/12/f... paper: arxiv.org/abs/2402.06184
sohl-dickstein.github.io
Neural network training makes beautiful fractals
This blog is intended to be a place to share ideas and results that are too weird, incomplete, or off-topic to turn into an academic paper, but that I think may be important. Let me know what you thin...
0284
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 12/02/2024
Have you ever done a dense grid search over neural network hyperparameters? Like a *really dense* grid search? It looks like this (!!). Blueish colors correspond to hyperparameters for which training converges, redish to those for which training diverges. Even better, a video: vimeo.com/903855670
Examples of fractals resulting from neural network training in a variety of experimental configurations
714031
Reposted by Jascha Sohl-Dickstein
Michael "Shapes Dude" Betancourt @betanalpha.bsky.social · 08/01/2024
The “principle of indifference” is often presented as an intuitively obvious motivation for specifying “non-informative” prior models. Unfortunately that intuition quickly falls apart in many common applications.  A long thread about applied probability theory!
1125
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 10/01/2024
All the appointments are filled. Will see how the meetings go, and evaluate doing this again. I'm looking forward to finding out what people are interested in!
210
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 10/01/2024
I created 11 meeting slots for this first round. I'll reply to this message when/if they've all filled up. calendar.app.google/ZdnWzeYw3qyR...
100
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 10/01/2024
I'm running an experiment, and holding some public office hours (inspired by Kyunghyun Cho doing something similar). Talk with me about anything! Ask for advice on your research or startup or career or I suppose personal life, brainstorm new research ideas, complain about mistakes I've made, ...
170
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 07/11/2023
Paper here: arxiv.org/abs/2311.02462 (Imagen is a Level 3 "Expert" Narrow AI)
000
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 07/11/2023
These Levels of AGI provide a rough framework to quantify the performance, generality, and autonomy of AGI models and their precursors. We hope they help compare models, assess risk, and measure progress along the path to AGI. (AlphaGo is a Level 4 "Virtuouso" Narrow AI)
110
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 07/11/2023
Levels of AGI: Operationalizing Progress on the Path to AGI Levels of Autonomous Driving are extremely useful, for communicating capabilities, setting regulation, and defining goals in self driving. We propose analogous Levels of *AGI*. (ChatGPT is a Level 1 "Emerging" AGI)
170
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 11/09/2023
AI can enable awesome (as in inspiring of awe) good in the world. We have amazing leverage as the people building it. We should use that leverage carefully.
010
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 11/09/2023
A theme is that we should worry about a *diversity* of risks. If we recognize e.g. only specific present harms, or only AGI misalignment risk, we will find our efforts overwhelmed by other types of AI-enabled disruption, and we won't be able to fix the problem we care about.
110
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 11/09/2023
My top fears include targeted manipulation of humans, autonomous weapons, massive job loss, AI-enabled surveillance and subjugation, widespread failure of societal mechanisms, extreme concentration of power, and loss of human control.
110
Jascha Sohl-Dickstein @jascha.sohldickstein.com · 11/09/2023
AI has the power to change the world in both wonderful and terrible ways. If we exercise care, the wonderful outcomes will be much more likely than the terrible ones. Towards that end, here is a brain dump of my thoughts about how AI might go wrong. sohl-dickstein.github.io/2023/09/10/d...
182