Sign in

Edward Grefenstette

@egrefen.bsky.social
7.8K followers 100 following 67 posts

FR/US/GB AI/ML Person, Director of Research at Google DeepMind, Honorary Professor at UCL DARK, ELLIS Fellow. Ex Oxford CS, Meta AI, Cohere.

PostsRepliesMedia
Edward Grefenstette @egrefen.bsky.social · 21/07/2025
Do you have a PhD (or equivalent) or will have one in the coming months (i.e. 2-3 months away from graduating)? Do you want to help build open-ended agents that help humans do humans things better, rather than replace them? We're hiring 1-2 Research Scientists! Check the 🧵👇
3196
Edward Grefenstette @egrefen.bsky.social · 25/03/2025
FYI this posting for a research scientist position in the autonomous assistants team at Google DeepMind will be open for a little under a week, as of today. Please consider applying if you are interested and qualify. See post for details, or ask questions here.
053
Edward Grefenstette @egrefen.bsky.social · 18/03/2025
Our team in London is hiring a research scientist! If you want to come work with a wonderful group of researchers on investigating the frontiers of autonomous open-ended agents that help humans be better at doing things we love, come have a look. Link in post below 👇
2228
Edward Grefenstette @egrefen.bsky.social · 30/12/2024
🧵 As 2024 wraps up, please pardon my usual self-indulgence in tweeting about the year gone by. 🧵 This will be a reasonably short one... OR WILL IT? [1/17]
2170
Edward Grefenstette @egrefen.bsky.social · 24/12/2024
Merry Christmas (eve), you filthy animal(s).
1240
Edward Grefenstette @egrefen.bsky.social · 09/12/2024
Researchers: be constructively skeptical about LLMs. Find where they don't work by building with them. Find out if the failure is systemic or just transient. This way, you're best positioned to build what's next, or, if they keep working, to benefit from their growth.
3353
Edward Grefenstette @egrefen.bsky.social · 02/12/2024
Seek novelty in what you do, how you do it, and who you do it with. I feel part of happiness lies in committing to these things, but not obsessively overcommitting to just one of these things.
0220
Edward Grefenstette @egrefen.bsky.social · 25/11/2024
Multi-agent peeps: are there any *-MDP variants where there is more than one agent, but exactly one agent is acting on the environment at each time step? Not in the sense of "we take turns" (although I guess it's a special case) but more in the sense that the agents decide who gets to act...
380
Reposted by Edward Grefenstette
Max Bartolo @maxbartolo.bsky.social · 20/11/2024
🚨 LLMs can learn to reason from procedural knowledge in pretraining data! 🚨 I particularly enjoy research where the evidence contradicts our initial hypothesis. If you're interested in LLM reasoning, check out the 60+ pages of in-depth work at arxiv.org/abs/2411.12580
3677
Edward Grefenstette @egrefen.bsky.social · 20/11/2024
“LLMs can/can’t reason” — whatever you think, they clearly can solve some reasoning problems, but how do they learn to do this? Is the dependency on the training data measurable, relative to factual knowledge? Does this tell us something about their abilities? Find out here!
1250
Reposted by Edward Grefenstette
Laura @lauraruis.bsky.social · 20/11/2024
How do LLMs learn to reason from data? Are they ~retrieving the answers from parametric knowledge🦜? In our new preprint, we look at the pretraining data and find evidence against this: Procedural knowledge in pretraining drives LLM reasoning ⚙️🔢 🧵⬇️
36850139
Reposted by Edward Grefenstette
arxiv cs.CL @arxiv-cs-cl.bsky.social · 20/11/2024
Laura Ruis, Maximilian Mozes, Juhan Bae, Siddhartha Rao Kamalakara, Dwarak Talupuru, Acyr Locatelli, Robert Kirk, Tim Rockt\"aschel, Edward Grefenstette, Max Bartolo Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models arxiv.org/abs/2411.12580
0146
Edward Grefenstette @egrefen.bsky.social · 20/11/2024
Is there some way to stop Bluesky from popping a notification on my phone every time I get a follower?
390
Edward Grefenstette @egrefen.bsky.social · 19/11/2024
🌶️(?) take: Agents are somehow hot right because people realized that LLM output can be interpreted as a DSL which directs side effects in the world (e.g. tool calls) rather than just returning text in a chat/autocomplete sense. What are the open challenges? A 🧵... [1/11]
916531
Edward Grefenstette @egrefen.bsky.social · 18/11/2024
Is there some good way to selectively crosspost to X and Bluesky, e.g. draft a post somewhere central, and then just post to one/the other/both with a keypress or click? Obviously I can just copy/paste... maybe that's the easiest way.
360
Edward Grefenstette @egrefen.bsky.social · 18/11/2024
Deep down, everything is a minmax game. We'll get to AGI (whatever that means for you) by building better minmax objectives.
250
Edward Grefenstette @egrefen.bsky.social · 17/11/2024
So are we mainly shitposting here too or should I reserve Bluesky for balanced takes on ML and leave the spicy takes (mainly rage about politics) for Twitter?
18813
Edward Grefenstette @egrefen.bsky.social · 17/11/2024
What’s good, Bluesky?
2170
Edward Grefenstette @egrefen.bsky.social · 12/09/2023
🚨 JOB ALERT 🚨 We're hiring research scientists/engineers to conduct research on next-gen assistant technologies to power increasingly autonomous agents which strive to support humans Research Scientist: boards.greenhouse.io/deepmind/job... Research Engineer: boards.greenhouse.io/deepmind/job...
061
Edward Grefenstette @egrefen.bsky.social · 26/08/2023
Here we go again.
091