Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 01/07/2026I'm excited to share our latest work “Representation Learning Enables Scalable Multitask Deep Reinforcement Learning”. In this work, we revisit a fundamental question in reinforcement learning: What if representation learning was the key ingredient behind scalable RL? 1/🧵 15213
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 25/04/2026Still at #ICLR2026 and still into RL? Come talk to @johanobandoc.bsky.social @waltermayor.bsky.social and me at our poster today (Saturday)! 🗓️April 25, 3:15 PM -5:45 PM BRT 📍Pavilion 4· #️⃣ Poster #4515 292
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 17/03/2026New paper 🚨 "Stable Deep Reinforcement Learning via Isotropic Gaussian Representations" Deep RL suffers from unstable training, representation collapse, and neuron dormancy. We show that a simple geometric insight, isotropic Gaussian representations, can fix this. Here's how 👇 2295
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 10/12/2025This #NeurIPS2025 was tiring, but it was fantastic to connect with so many friends and colleagues! I was so busy I didn't get a chance to tweemote our papers at the conference, so I'll remedy that with this post-hoc thread: 👇🏾 131
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 20/10/2025🔊Simplicial Embeddings (SEMs) Improve Sample Efficiency in Actor-Critic Agents🔊 In our recent preprint we demonstrate that the use of well-structured representations (SEMs) can dramatically improve sample efficiency in RL agents. 1/X 1153
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 20/10/2025This work was led by @johanobandoc.bsky.social and @waltermayor.bsky.social , with @lavoiems.bsky.social , Scott Fujimoto and Aaron Courville. Read the paper at arxiv.org/abs/2510.13704 11/Xarxiv.orgSimplicial Embeddings Improve Sample Efficiency in Actor-Critic AgentsRecent works have proposed accelerating the wall-clock training time of actor-critic methods via the use of large-scale environment parallelization; unfortunately, these can sometimes still require la... 031
Reposted by Johan S Obando 👍🏽lasalaai.bsky.social @lasalaai.bsky.social · 04/10/2025📢 ¡Buenas noticias! Se extiende el período de inscripciones 🎉 No pierdas esta oportunidad de ser parte de un evento único que puede transformar tu futuro en la inteligencia artificial. 👉 Regístrate ahora y asegura tu lugar. lasala.ai#top 011
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 13/07/2025Thrilled to share our #ICML2025 paper “The Courage to Stop: Overcoming Sunk Cost Fallacy in Deep RL”, led by Jiashun Liu and with other great collaborators! We teach RL agents when to quit wasting effort, boosting efficiency with our proposed method LEAST. Here's the story 🧵👇🏾 1223
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 24/06/2025proud to share a survey of state representation learning in RL that my student ayoub echchahed and i prepared, that was just published on @tmlrorg.bsky.social ! this was the bulk of ayoub's masters thesis and he put a lot of work and care into it! a few details in thread below... 1/ 1216
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 05/06/2025The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks thrilled to share our #ICML2025 paper led by Walter Mayor & @johanobandoc.bsky.social , with Aaron Courville, where we explore how data collection affects agents in parallelized setups. 1/ 286
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 23/06/2025really excited about this new work we just put out, led by my students @roger-creus.bsky.social & @johanobandoc.bsky.social , where we examine the challenges of gradient propagation when scaling deep RL nets. roger & johan put in a lot of work and care in this work, check out more details in 🧵👇🏾 ! 032
Reposted by Johan S Obando 👍🏽Nenad Tomasev @nenadtomasev.bsky.social · 05/12/2024I'm excited to share a new paper: "Mastering Board Games by External and Internal Planning with Language Models" storage.googleapis.com/deepmind-med... (also soon to be up on Arxiv, once it's been processed there)storage.googleapis.com 47614
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 02/12/2024Everyone I spoke to at @rl-conference.bsky.social last summer agreed on it being one of the best conferences ever for an RL researcher... So many great RL-focused papers! CFP is out, send your work here! 14513
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 04/12/2024Post based on a talk I gave earlier this year at the AutoRL workshop in ICML, and leveraging two recent papers (1st with @joaogui1.bsky.social & @johanobandoc.bsky.social , 2nd with @jessefarebro.bsky.social ): 1-hparam transfer openreview.net/forum?id=szU... 2-CALE openreview.net/forum?id=vlU... 1151
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 04/12/2024📢 In Defense Of Atari 📢 New blog post in which I argue why the ALE is still a valuable resource for RL research! psc-g.github.io/posts/resear... 1418
Reposted by Johan S Obando 👍🏽Marzieh Fadaee @mziizm.bsky.social · 03/12/2024Good performance shouldn’t mean 'just in English' anymore 🪩 We provide a robust way to assess models with a new benchmark that captures in-language nuances and cultural contexts. 1182
Reposted by Johan S Obando 👍🏽Pablo Samuel Castro @pcastr.bsky.social · 28/11/2024Last year I gave a talk titled "From 'Bigger, Better, Faster' to 'Smaller, Sparser, Stranger'", which looked at the components that make up our BBF agent (arxiv.org/abs/2305.19452), highlighting some promising areas of research. Finally in blog form, have a read! psc-g.github.io/posts/resear... 36713
Reposted by Johan S Obando 👍🏽Nenad Tomasev @nenadtomasev.bsky.social · 25/11/2024While I will be talking about some of our research work, there are also several fun and engaging features for fans to explore blog.google/technology/a... as well as a new Kaggle challenge on developing efficient Chess AI www.kaggle.com/competitions... under resource constraints.lnkd.inLinkedInThis link will take you to a page that’s not on LinkedIn 031
Reposted by Johan S Obando 👍🏽Nenad Tomasev @nenadtomasev.bsky.social · 26/11/2024A great new essay on AI for Science from our colleagues here: deepmind.google/public-polic...deepmind.googleA new golden age of discoveryIn this essay, we take a tour of how AI is transforming scientific disciplines from genomics to computer science to weather forecasting. Some scientists are training their own AI models, while... 0225
Reposted by Johan S Obando 👍🏽David Abel @dabelcs.bsky.social · 22/11/2024RLDM will be held next year in Dublin! A reminder that the call for workshops is out: rldm.org/call-for-wor... The workshops are one of my favourite parts of the conference :) please get in touch if you have any questions!rldm.orgCALL FOR WORKSHOPS | RLDM 14215