Sign in

Nenad Tomasev

@nenadtomasev.bsky.social
1.7K followers 202 following 70 posts

Developing AI responsibly. Senior Staff Research Scientist at Google DeepMind. Opinions are my own.

PostsRepliesMedia
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 29/10/2025
Great work from some of my colleagues, check this out! Such a creative application of generative AI! ♟️ www.chess.com/news/view/ai...
chess.com
DeepMind's AI Learns To Create Original Chess Puzzles, Praised By GMs
In a new study, researchers from Google DeepMind have created an AI system that is capable of generating creative chess puzzles, some of which impressed experts in chess compositions.
191
Reposted by Nenad Tomasev
Arvid Ågren @arvidagren.bsky.social · 14/09/2025
Take a sneak peak at The Paradox of the Organism at Google books books.google.com/books?hl=sv&... Don’t forget to pre-order!
books.google.com
The Paradox of the Organism
Leading evolutionary theorists and philosophers come together to understand how organisms persist, even as they’re riddled with internal conflict, from cancer cells to selfish genes.How does a vast me...
12610
Nenad Tomasev @nenadtomasev.bsky.social · 15/09/2025
Another concept that we explore is that of Mission Economies realized through steerable AI agent markets.
120
Nenad Tomasev @nenadtomasev.bsky.social · 15/09/2025
One of the concepts that we discuss is that of a hypothetical High-Frequency Negotiation (HFN) framework between Personal AI Assistants, on behalf of their users, and in line with individual preferences.
110
Nenad Tomasev @nenadtomasev.bsky.social · 15/09/2025
Happy to share a new preprint: Virtual Agent Economies arxiv.org/abs/2509.10147 where we discuss a number of possible frameworks for establishing steerable agent markets.
arxiv.org
Virtual Agent Economies
The rapid adoption of autonomous AI agents is giving rise to a new economic layer where agents transact and coordinate at scales and speeds beyond direct human oversight. We propose the "sandbox econo...
142
Nenad Tomasev @nenadtomasev.bsky.social · 09/05/2025
For those with interest in mental health and AI, and in particular on how potentially sensitive data gets collected and used there - there is now a Delphi survey that you can fill out to inform how this gets done, as a part of PARQAIR-MH project (www.parqair.org/home), via: redcap.link/2l91l043
parqair.org
000
Nenad Tomasev @nenadtomasev.bsky.social · 01/04/2025
Our work on concept discovery towards bridging the human-AI knowledge gap in AlphaZero has now been published in PNAS. As future AI systems become even more capable, we should be thinking of ways of utilizing them not only to perform tasks, but also to further our own knowledge and understanding.
095
Reposted by Nenad Tomasev
Ethan Mollick @emollick.bsky.social · 01/04/2025
This was cool: "Gemini 2.5, create a sim that is cross between Jujujajaki networks and cellular automata" Gemini: "What's a Jujujajaki network?" I paste in a paper. Gemini: "Got it, a dynamic network with local search & exploration." Worked in one shot. Me: "Make it nicer" Some cleverness here
3817
Nenad Tomasev @nenadtomasev.bsky.social · 26/03/2025
An interesting read on self-organizing systems, and how order and disorder may manifest differently at different scales: www.nature.com/articles/s44...
nature.com
Self-organizing systems: what, how, and why? - npj Complexity
npj Complexity - Self-organizing systems: what, how, and why?
040
Reposted by Nenad Tomasev
Sung Kim @sungkim.bsky.social · 05/03/2025
Dataset Distillation (2018/2020) They show that it is possible to compress 60,000 MNIST training images into just 10 synthetic distilled images (one per class) and achieve close to original performance with only a few gradient descent steps, given a fixed network initialization.
2265
Nenad Tomasev @nenadtomasev.bsky.social · 03/03/2025
I'm happy to advertise an upcoming Student Researcher position on my Agent Frontiers team here at the Google DeepMind Foundational Research Unit, aimed for a start date early in the summer (currently listed as late June, but obviously somewhat flexible).
1114
Reposted by Nenad Tomasev
Kevin Mitchell @wiringthebrain.bsky.social · 14/02/2025
Constrained roads to complex brains - Neural development and brain circuit evolution converged in birds and mammals www.science.org/doi/10.1126/...
science.org
Constrained roads to complex brains
Neural development and brain circuit evolution converged in birds and mammals
16925
Reposted by Nenad Tomasev
Matej Jusup @matejjusup.bsky.social · 14/02/2025
LLMs Mastering Board Games: ZurichNLP Meetup - Feb 20th! Excited to share insights from my student research at Google DeepMind at the upcoming ZurichNLP meetup! I'll present how we achieved high-level play in board games using LLMs with a search budget comparable to human chess grandmasters.
2153
Reposted by Nenad Tomasev
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 14/02/2025
Model-free deep RL algorithms like NFSP, PSRO, ESCHER, & R-NaD are tailor-made for games with hidden information (e.g. poker). We performed the largest-ever comparison of these algorithms. We find that they do not outperform generic policy gradient methods, such as PPO. arxiv.org/abs/2502.08938 1/N
39321
Reposted by Nenad Tomasev
Pablo Samuel Castro @pcastr.bsky.social · 10/02/2025
Can LLMs be used to discover interpretable models of human and animal behavior?🤔 Turns out: yes! Thrilled to share our latest preprint where we used FunSearch to automatically discover symbolic cognitive models of behavior. 1/12
313645
Reposted by Nenad Tomasev
Tim Roughgarden @timroughgarden.bsky.social · 09/02/2025
🙏🙏🙏 See also full set of video lectures at m.youtube.com/playlist?lis...
m.youtube.com
Algorithmic Game Theory (Stanford CS364A, Fall 2013) - YouTube
Course Web site: http://timroughgarden.org/f13/f13.html (includes lecture notes and homeworks). Course description: Broad survey of topics at the interface o...
1375
Reposted by Nenad Tomasev
Neil Traft @ntraft.bsky.social · 08/02/2025
No company in the self-driving industry has invested in self-play at scale. I've been dying for someone to finally do this—many others have felt the same. This landmark work finally shows the potential in this approach. This is a challenge to the industry.
2112
Nenad Tomasev @nenadtomasev.bsky.social · 07/02/2025
One of the things that I am most excited about is how we can use AI (and language models in particular) to help us generate new scientific hypotheses and accelerate scientific discovery.
160
Reposted by Nenad Tomasev
Jeff Dean @jeffdean.bsky.social · 05/02/2025
We launched a bunch of Gemini 2.0 models today. Compared to the 1.5 series models, each of the 2.0 models is generally better than the "one size up" model in the 1.5 series. 2.0 Flash & Flash-Lite set new standards in the quality/cost Pareto frontier. More details: blog.google/technology/g...
49414
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 17/01/2025
In December, I posted about our new paper on mastering board games using internal + external planning. 👇 Here's a talk now on Youtube about it given by my awesome colleague John Schultz! www.youtube.com/watch?v=JyxE...
youtube.com
John Schultz, DeepMind, Mastering Board Games by External and Internal Planning with Language Models
YouTube video by AI4All
13511
Reposted by Nenad Tomasev
Matej Jusup @matejjusup.bsky.social · 17/01/2025
John's talk is now available online! www.youtube.com/watch?v=JyxE...
youtube.com
John Schultz, DeepMind, Mastering Board Games by External and Internal Planning with Language Models
YouTube video by AI4All
0132
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 16/01/2025
Check out the Autonomous Agents for Social Good (AASG) Workshop taking place at #AAMAS 2025! Are you doing research on using autonomous agents or multi-agent systems to address social challenges to make the world a better place? Submission deadline: Feb. 4th. panosd.eu/aasg2025/
1162
Nenad Tomasev @nenadtomasev.bsky.social · 16/01/2025
For those interested in AI for improving mental health, registrations are still open for the next MEXA hackathon. www.linkedin.com/posts/mexaco...
linkedin.com
MEXA on LinkedIn: #mexa #hackathon #mentalhealth #ai #innovation
Have you registered for MEXA's next Virtual Hackathon, 𝙈𝙚𝙣𝙩𝙖𝙡 𝙃𝙚𝙖𝙡𝙩𝙝 𝙖𝙣𝙙 𝘼𝙄: 𝙄𝙣𝙩𝙚𝙧𝙫𝙚𝙣𝙩𝙞𝙤𝙣? Now is the time!⏰ Join us in…
030
Reposted by Nenad Tomasev
Matej Jusup @matejjusup.bsky.social · 13/01/2025
Join John's talk to get insights on our paper on mastering board games with language models!
061
Reposted by Nenad Tomasev
mr. TIM @timkellogg.me · 24/12/2024
A new paper dropped from DeepMind: Deliberation in Latent Space via Differentiable Cache Augmentation The trouble is, it's not very readable. I tried making a thread here, but it got far too long, so it's a blog now: timkellogg.me/blog/2024/12...
timkellogg.me
Explainer: Latent Space Experts - Tim Kellogg
6699
Nenad Tomasev @nenadtomasev.bsky.social · 18/12/2024
Now also out on Arxiv: www.arxiv.org/abs/2412.12119
arxiv.org
Mastering Board Games by External and Internal Planning with Language Models
While large language models perform well on a range of complex tasks (e.g., text generation, question answering, summarization), robust multi-step planning and reasoning remains a considerable challen...
1305
Reposted by Nenad Tomasev
Matej Jusup @matejjusup.bsky.social · 11/12/2024
I was honored to present our demo on “Mastering Chess With Language Models” at the @GoogleDeepMind booth at @NeurIPSConf right before an inspiring talk by Jeff Dean Thanks to everyone who dropped by and made it a memorable session! ♟️
2162
Reposted by Nenad Tomasev
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 09/12/2024
An updated intro to reinforcement learning by Kevin Murphy: arxiv.org/abs/2412.05265! Like their books, it covers a lot and is quite up to date with modern approaches. It also is pretty unique in coverage, I don't think a lot of this is synthesized anywhere else yet
arxiv.org
Reinforcement Learning: An Overview
This manuscript gives a big-picture, up-to-date overview of the field of (deep) reinforcement learning and sequential decision making, covering value-based RL, policy-gradient methods, model-based met...
926974
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 06/12/2024
If you will be at #NeurIPS2024 @neuripsconf.bsky.social and would like to come see our models in action, come say hi 👋 and check out our demo at the GDM booth! Wednesday, Dec. 11th @ 9:30-10:00. Lots of other great things to see as well! Check it out: 👇 deepmind.google/discover/blo...
1416
Nenad Tomasev @nenadtomasev.bsky.social · 06/12/2024
If you are as interested as we are in exploring how planning and reasoning with large language models can help master board games and vice versa, Join us this coming Wednesday in Vancouver at the GDM NeurIPS booth, where we will be showing a demo. deepmind.google/discover/eve...
deepmind.google
ICML 2024
1133
Nenad Tomasev @nenadtomasev.bsky.social · 06/12/2024
Now you can also play chess against Gemini - we've made one of the smaller models (without search) available in Gemini Advanced. The fun part? You can banter with Gemini as you play along, have it take on a personality of your choosing, or even get creative and poetic.
130
Reposted by Nenad Tomasev
Matej Jusup @matejjusup.bsky.social · 05/12/2024
The project I was quite honored to work on! ♟️ As a chess enthusiast, I was especially thrilled that an LLM equipped with an efficient MCTS reaches grandmaster-level strength with a search budget comparable to human players! 🤯 deepmind.google/research/pub...
161
Nenad Tomasev @nenadtomasev.bsky.social · 05/12/2024
I'm excited to share a new paper: "Mastering Board Games by External and Internal Planning with Language Models" storage.googleapis.com/deepmind-med... (also soon to be up on Arxiv, once it's been processed there)
storage.googleapis.com
47614
Nenad Tomasev @nenadtomasev.bsky.social · 03/12/2024
If you're planning on attending NeurIPS next week, and are interested in building advanced AI agents that use planning and reasoning for solving complex problems - I'll be at the GDM booth with my colleagues on the 11th where we'll be covering some of these topics.
3141
Reposted by Nenad Tomasev
Cole Hurwitz @colehurwitz.bsky.social · 02/12/2024
Ported over from X! What will a foundation model for the brain look like? 🧠 We argue that it must be able to solve a diverse set of tasks across multiple brain regions and animals. Check out our NeurIPS paper which introduces a multi-region, multi-animal, multi-task model arxiv.org/abs/2407.14668
18328
Nenad Tomasev @nenadtomasev.bsky.social · 03/12/2024
An interesting talk later today - highlighting the opportunities for acoustic AI in biodiversity monitoring, across a range of applications including coral reef health, population discovery and tracking threatened species. gdg.community.dev/events/detai...
gdg.community.dev
Acoustic AI in biodiversity monitoring | Google Developer Groups
Virtual Event - An overview of Google's efforts to develop flexible machine learning tools to help understand biodiversity through sounds recordings
030
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 26/11/2024
Check out this interview with @nenadtomasev.bsky.social about AlphaZero at the Chess World Championship! www.youtube.com/watch?v=ppiW...
youtube.com
Game 2 Commentary with GM David Howell and IM Jovanka Houska | FIDE World Championship Match 2024
YouTube video by FIDE chess
1132
Nenad Tomasev @nenadtomasev.bsky.social · 26/11/2024
A great new essay on AI for Science from our colleagues here: deepmind.google/public-polic...
deepmind.google
A new golden age of discovery
In this essay, we take a tour of how AI is transforming scientific disciplines from genomics to computer science to weather forecasting. Some scientists are training their own AI models, while...
0225
Nenad Tomasev @nenadtomasev.bsky.social · 25/11/2024
Super excited to be here for the world chess championship - and if game one is anything to go by, it's going to be quite an exciting match!
291
Nenad Tomasev @nenadtomasev.bsky.social · 19/11/2024
A couple of weeks ago, I had the pleasure of speaking at the 50th anniversary of the world computer chess championship icga.org?page_id=3957, hosted by ECAI and organized by the International Computer Games Association.
icga.org
2024 World Computer Chess Championships: The 50th Anniversary | ICGA
120
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 18/11/2024
Student Researcher positions in EMEA now accepting applications! Please repost. www.google.com/about/career...
google.com
Student Researcher, 2025 — Google Careers
0249
Reposted by Nenad Tomasev
Daniel Lowd @dlowd.com · 17/11/2024
Since this platform is finally attracting a critical mass of ML researchers, here's our recent work on prompt-based vulnerabilities of coding assistants: arxiv.org/abs/2407.11072 TL;DR — An attacker can convince your favorite LLM to suggest vulnerable code with just a minor change to the prompt!
arxiv.org
MaPPing Your Model: Assessing the Impact of Adversarial Attacks on LLM-based Programming Assistants
LLM-based programming assistants offer the promise of programming faster but with the risk of introducing more security vulnerabilities. Prior work has studied how LLMs could be maliciously fine-tuned...
421433
Reposted by Nenad Tomasev
NIHR Biomedical Research Centre: Maudsley @nihrmaudsleybrc.bsky.social · 14/11/2024
How can we integrate #AI ethically in healthcare? We're at Science Gallery London for a day of talks, we'll be hearing from experts across academia, government, medical sector, industry, and patient community, e.g. @saracerdas.bsky.social Former MEP, @nenadtomasev.bsky.social Google Deepmind
195
Reposted by Nenad Tomasev
Marc Lanctot @sharky6000.bsky.social · 14/11/2024
Zhang and Hardt '24 arxiv.org/abs/2405.01719. In this paper, the authors use social choice theory to show that as our benchmarks become more diverse, model ranking can become less stable. Evaluated over many (18!) different (LM) benchmarks.
arxiv.org
Inherent Trade-Offs between Diversity and Stability in Multi-Task Benchmarks
We examine multi-task benchmarks in machine learning through the lens of social choice theory. We draw an analogy between benchmarks and electoral systems, where models are candidates and tasks are vo...
1223
Reposted by Nenad Tomasev
David Pfau @davidpfau.com · 13/11/2024
One thing still missing here is good discussion of academic papers, so I guess I'll be the change I want to see in the world Really interesting results from Jacob Andreas's group showing great performance on ARC-AGI just from doing a few gradient descent steps at test time arxiv.org/abs/2411.07279
arxiv.org
The Surprising Effectiveness of Test-Time Training for Abstract Reasoning
Language models have shown impressive performance on tasks within their training distribution, but often struggle with novel problems requiring complex reasoning. We investigate the effectiveness of t...
718224
Reposted by Nenad Tomasev
Timothée Poisot @ctrlalttim.com · 12/11/2024
Our ability to monitor and protect biodiversity relies on how precisely we can predict where species are. But we haven't done a good job of measuring the uncertainty of these models. In a new preprint, I show how conformal prediction quantifies uncertainty in SDMs. A thread. 🧪🌏
ecoevorxiv.org
Conformal Prediction quantifies the uncertainty of Species Distribution Models
46410