Sign in

Martin Klissarov

@martinklissarov.bsky.social
286 followers 111 following 18 posts

research @ Google DeepMind

PostsRepliesMedia
Martin Klissarov @martinklissarov.bsky.social · 27/06/2025
As AI agents face increasingly long and complex tasks, decomposing them into subtasks becomes increasingly appealing. But how do we discover such temporal structure? Hierarchical RL provides a natural formalism-yet many questions remain open. Here's our overview of the field🧵
13510
Reposted by Martin Klissarov
Edward Grefenstette @egrefen.bsky.social · 18/03/2025
Our team in London is hiring a research scientist! If you want to come work with a wonderful group of researchers on investigating the frontiers of autonomous open-ended agents that help humans be better at doing things we love, come have a look. Link in post below 👇
2228
Reposted by Martin Klissarov
Ulyana Piterbarg @upiter.bsky.social · 12/02/2025
Our paper showing that LMs benefit from human-like abstractions for code synthesis was accepted to ICLR! 🇸🇬 We show that order matters in code gen. -- casting code synthesis as a sequential edit problem by preprocessing examples in SFT data improves LM test-time scaling laws
1102
Martin Klissarov @martinklissarov.bsky.social · 04/02/2025
Can AI agents adapt zero-shot, to complex multi-step language instructions in open-ended environments? We present MaestroMotif, a method for skill design that produces highly capable and steerable hierarchical agents. Paper: arxiv.org/abs/2412.08542 Code: github.com/mklissa/maestromotif
1216
Reposted by Martin Klissarov
Devon Hjelm @devhje.bsky.social · 22/01/2025
Our paper on AI feedback was accepted to #ICLR2025 as a poster. Great work by @martinklissarov.bsky.social , @bmazoure.bsky.social , and Alex Toshev arxiv.org/abs/2410.05656
arxiv.org
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
Large pretrained models are showing increasingly better performance in reasoning and planning tasks across different modalities, opening the possibility to leverage them for complex sequential decisio...
095
Reposted by Martin Klissarov
Jack Parker-Holder @jparkerholder.bsky.social · 04/12/2024
Introducing 🧞Genie 2 🧞 - our most capable large-scale foundation world model, which can generate a diverse array of consistent worlds, playable for up to a minute. We believe Genie 2 could unlock the next wave of capabilities for embodied agents 🧠.
1523461