Sign in

Pedro Santos

@pedrosantospps.bsky.social
20 followers 38 following 14 posts

PhD candidate at @istecnico.bsky.social working on sequential decision-making and reinforcement learning. ppsantos.github.io

PostsRepliesMedia
Pedro Santos @pedrosantospps.bsky.social · 28/08/2026
Last week, I attended the Reinforcement Learning Conference in Montreal to present our work, "Risk-Aware General-Utility Markov Decision Processes" (arxiv.org/abs/2607.09298). More details about our work are in the following blog post: ppsantos.github.io/posts/risk-a...
ppsantos.github.io
Bringing risk-awareness to general-utility MDPs
We motivate and explore risk-aware general-utility MDPs, where we aim to find an optimal policy with respect to a risk measure of the distribution of objective values induced by its interaction with t...
000
Pedro Santos @pedrosantospps.bsky.social · 10/02/2026
Our work, "Solving General-Utility Markov Decision Processes in the Single-Trial Regime with Online Planning", got accepted to ICLR 2026. arxiv.org/abs/2505.15782 1/N Joint work with Francisco S. Melo and Alberto Sardinha.
arxiv.org
Solving General-Utility Markov Decision Processes in the Single-Trial Regime with Online Planning
In this work, we contribute the first approach to solve infinite-horizon discounted general-utility Markov decision processes (GUMDPs) in the single-trial regime, i.e., when the agent's performance is...
110
Reposted by Pedro Santos
GAIPS Lab @gaipslab.bsky.social · 30/01/2026
Here’s Pedro at yet another international conference! 🙌✨ GAIPS member Pedro P. Santos presented “Centralized training with hybrid execution in multi-agent reinforcement learning via predictive observation imputation” at #AAAI2026, Singapore 🇸🇬 📄 Check out his paper: doi.org/10.1016/j.ar...
011
Reposted by Pedro Santos
GAIPS Lab @gaipslab.bsky.social · 03/10/2025
Here’s some photos of GAIPS member @pedrosantospps.bsky.social presenting his work on ICML 2025 in Vancouver and EWRL 2025 in Tübingen, Germany. His poster was selected as a "spotlight poster" (top 2.6% of the papers)! 🙌 Read his work here: icml.cc/virtual/2025...
011
Reposted by Pedro Santos
Mirco Mutti @mircomutti.bsky.social · 24/07/2025
Walking around posters at @icmlconf.bsky.social, I was happy to see some buzz around convex RL—a topic I’ve worked on and strongly believe in. Thought I’d share a few ICML papers on this direction. Let’s dive in👇 But first… what is convex RL? 🧵 1/n
151
Pedro Santos @pedrosantospps.bsky.social · 03/05/2025
Happy to share that our paper "The Number of Trials Matters in Infinite-Horizon General-Utility Markov Decision Processes" got accepted as a spotlight poster at the International Conference on Machine Learning (ICML).
251