Sign in

Théo Vincent

@theo-vincent.bsky.social
108 followers 258 following 53 posts

PhD student working on RL 🤖 @DFKI & @ias-tudarmstadt.bsky.social | Master MVA @ENS_ParisSaclay & ENPC 🎓 www.ias.informatik.tu-darmstadt.de/…

PostsRepliesMedia
Théo Vincent @theo-vincent.bsky.social · 23/08/2026
🥳 Congratulations!! It was a pleasure to work with Tim during his master's thesis! More to come soon⏳
000
Théo Vincent @theo-vincent.bsky.social · 21/08/2026
🏅Gradient Iterated Temporal-Difference Learning received Best Paper Award for empirical RL research @rl-conference.bsky.social🥳
110
Théo Vincent @theo-vincent.bsky.social · 15/05/2026
Why do we keep using semi-gradient methods when they can diverge?🤨 Gradient TD methods are often overlooked, while they have convergence guarantees! @rl-conference.bsky.social, we will present the first gradient TD method shown to be competitive against semi-gradient methods on deep RL benchmarks🏆
171
Théo Vincent @theo-vincent.bsky.social · 02/05/2026
Really excited about the talks, discussions, and reading the submitted works!
000
Reposted by Théo Vincent
RL in Big Worlds @rlcbigworlds.bsky.social · 01/05/2026
We are proud to have an amazing line-up of speakers! They will present their works, which incorporate the constraint that the world is bigger than the agent and impossible to anticipate, observe, or model perfectly. We are also looking forward to the panel discussion!
152
Théo Vincent @theo-vincent.bsky.social · 22/04/2026
I had an amazing time today visiting Professor Luiz Chaimowicz's lab in Belo Horizonte! I really enjoyed discovering the research being done at UFMG. It seems to be an amazing place to do some research! See you @iclr-conf.bsky.social in Rio
141
Théo Vincent @theo-vincent.bsky.social · 20/04/2026
I will be presenting 3 papers @iclr-conf.bsky.social this week 🇧🇷 Looking forward to some interesting exchanges!
141
Théo Vincent @theo-vincent.bsky.social · 15/04/2026
🌍World models can play an important role towards building general agents, but what should be their role in decision-making?🕹️ @joemwatson.bsky.social and I are organizing a Social @iclr-conf.bsky.social on this topic🎙️ 🗓️Feel free to join the conversation on Friday 24th April, at noon!!
030
Reposted by Théo Vincent
Intelligent Autonomous Systems @ias-tudarmstadt.bsky.social · 12/04/2026
Working with constrained agents in complex environments? Do not hesitate to submit your latest work to this workshop! See you in Montréal @rl-conference.bsky.social 🇨🇦
041
Reposted by Théo Vincent
Khurram Javed @khurramjaved.com · 11/04/2026
A bunch of us are organizing a workshop at RLC. If your goal is to develop algorithms that allow agents to learn from complex data streams without relying on human data and human designers, then this workshop would be a good fit.
051
Théo Vincent @theo-vincent.bsky.social · 11/04/2026
Really excited to organize this workshop! Many works overlook the complexity ratio between the agent and its environment, often leading to overpowered agents. If we want agents to learn continuously in the wild, we need to care about this ratio!
020
Reposted by Théo Vincent
RL in Big Worlds @rlcbigworlds.bsky.social · 11/04/2026
RL in Big Worlds is a workshop at @rl-conference.bsky.social about ideas that enable agents to achieve goals in environments vastly more complex than themselves. This requires giving agents the ability to learn continually and use approximate value functions, models, and policies effectively.
1107
Théo Vincent @theo-vincent.bsky.social · 22/03/2026
Benchmarking always takes a ton of time😮‍💨 and we often hear about it🗣️ But we rarely report the carbon footprint of experiments, which better reflects their weight! Here is the electricity emission of the experiments in each paper of my PhD👇
121
Reposted by Théo Vincent
Reinforcement Learning Conference @rl-conference.bsky.social · 12/02/2026
Quick reminder for everyone grinding on their RLC 2026 papers, only ~3 weeks to go! The submission site opens in just a few days (Feb 17). Deadlines: ⏳ March 1 (AoE): Abstract Submission ⏳ March 5 (AoE): Full Paper Submission Good luck with the final changes!
072
Reposted by Théo Vincent
Taylor W. Killian @twkillian.bsky.social · 13/02/2026
We're thrilled to share that the Call for Workshops for this year's @rl-conference.bsky.social is now live! As Workshop co-chair (alongside the wonderful Raksha Kumaraswamy and @claireve.bsky.social) we are looking forward to seeing the proposals for workshops that we receive. LINK IN NEXT POST
1115
Reposted by Théo Vincent
ahmed-hendawy.bsky.social @ahmed-hendawy.bsky.social · 11/02/2026
🧵 Accepted at @iclr-conf.bsky.social! Target networks stabilize bootstrapping in RL 🛡️ But induce slow-moving targets 🐢 Online networks adapt fast ⚡ But can diverge with function approximation 💥 𝗠𝗜𝗡𝗧𝗢 🌿 uses the online network 𝗼𝗻𝗹𝘆 𝗶𝗳 𝗶𝘁 𝗰𝗮𝗻 — yielding faster 𝘢𝘯𝘥 more stable RL. Here’s how 👇
1103
Reposted by Théo Vincent
Constantin Rothkopf @c-rothkopf.bsky.social · 08/02/2026
The Reinforcement Learning workshop at U Mannheim was a lot of fun and highly recommended if you are looking for an engaging exchange of ideas, thanks to the organizers: Leif Döring, @theo-vincent.bsky.social, @claireve.bsky.social, and Simon Weißmann! www.wim.uni-mannheim.de/doering/conf...
0132
Théo Vincent @theo-vincent.bsky.social · 05/02/2026
Should we use a target network in deep value-based RL?🤔 The answer has always been YES or NO, as there are pros and cons. @iclr-conf.bsky.social, I will present iS-QN, a method that lies in between this binary view, collecting the pros while reducing the cons🚀
1214
Reposted by Théo Vincent
lucas-schulze.bsky.social @lucas-schulze.bsky.social · 03/02/2026
🥳Our paper "Floating-Base Deep Lagrangian Networks (FeLaN)" has been accepted to #ICRA2026. FeLaN: a grey-box approach for physically consistent SysID of floating-base robots (humanoids, quadrupeds). 📄 arxiv.org/abs/2510.17270 💻 Soon! 🌐 schulze18.github.io/felan_website/
1103
Reposted by Théo Vincent
arXiv cs.LG Machine Learning @cslg-bot.bsky.social · 06/10/2025
Ahmed Hendawy, Henrik Metternich, Th\'eo Vincent, Mahdi Kallel, Jan Peters, Carlo D'Eramo: Use the Online Network If You Can: Towards Fast and Stable Reinforcement Learning arxiv.org/abs/2510.02590 arxiv.org/pdf/2510.02590 arxiv.org/html/2510.02590
011
Reposted by Théo Vincent
Claire Vernade @claireve.bsky.social · 02/12/2025
🎤 Announcing the 3rd workshop on Reinforcement Learning in Mannheim 🎤 We have an amazing lineup of speakers: @Mathieugeist, @gio_ramponi, Theresa Eimer, @SarahKeren_, @araffin2, @c_rothkopf, and @AdrienBolland ⏰ Friday 6th February 📍University of Mannheim
12210
Reposted by Théo Vincent
TMLR Published Papers @tmlr-pub.bsky.social · 27/10/2025
New #J2C Certification: Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning Théo Vincent, Daniel Palenicek, Boris Belousov, Jan Peters, Carlo D'Eramo openreview.net/forum?id=Lt2H8Bd8jF #reinforcement #iterative #iterations
021
Théo Vincent @theo-vincent.bsky.social · 19/09/2025
As usual, @ewrl18.bsky.social was a wonderful experience. I had the pleasure of presenting my research as a Contributed Talk 🎉 Special thanks to the organizers for making it happen!
182
Théo Vincent @theo-vincent.bsky.social · 04/08/2025
Looking forward to @rl-conference.bsky.social ! I will be presenting 4 posters. Feel free to come and exchange with me during the conference, at the Finding the Frame workshop, or at the Inductive Biases workshop🙂
020
Théo Vincent @theo-vincent.bsky.social · 19/07/2025
Had an amazing time presenting my research @cohereforai.bsky.social yesterday 🎤 In case you could not attend, feel free to check it out 👉 youtu.be/RCA22JWiiY8?...
youtu.be
Théo Vincent - Optimizing the Learning Trajectory of Reinforcement Learning Agents
YouTube video by Cohere
073
Théo Vincent @theo-vincent.bsky.social · 11/07/2025
🎤 Very excited to give a talk @cohereforai.bsky.social next week Friday 🎤 I will be presenting the research I have been working on for the last 2 years with Carlo D'Eramo, @jan-peters.bsky.social, and many more collaborators!
141
Reposted by Théo Vincent
Intelligent Autonomous Systems @ias-tudarmstadt.bsky.social · 12/06/2025
IAS is at RLDM 2025! We have many exiting works to share (see 👇), so come to our posters and talk to us!
443
Théo Vincent @theo-vincent.bsky.social · 12/06/2025
Sparse network -> sparse poster I will be presenting Eau De Q-Network today @rldmdublin2025.bsky.social Feel free to come and exchange at Poster #28 🎤 bsky.app/profile/theo...
010
Reposted by Théo Vincent
timschneider94.bsky.social @timschneider94.bsky.social · 12/06/2025
Excited to present our latest work at RLDM 2025! If you’re curious about tactile sensing, active perception, or RL in robotics, stop by my poster. Here’s what we’ve been up to: 🧵 #Robotics #TactileSensing #ReinforcementLearning #Transformers #ActivePerception @ias-tudarmstadt.bsky.social
192
Théo Vincent @theo-vincent.bsky.social · 09/06/2025
Very excited to present 🎉Eau De Q-Network🎉 on Thursday @rldmdublin2025.bsky.social Poster #28 🔍Eau De Q-Network gradually prunes the network weights at the agent's learning pace, ultimately reaching a final sparsity level that is discovered by the algorithm!🔎 👉📰 arxiv.org/pdf/2503.01437
1112
Reposted by Théo Vincent
Snehal Jauhri @snehaljauhri.bsky.social · 06/04/2025
Learn more at the workshop website: egoact.github.io/rss2025 Happy to be organizing this with @georgiachal.bsky.social, Yu Xiang, @danfei.bsky.social and @galasso.bsky.social!
egoact.github.io
Web home for EgoAct: 1st Workshop on Egocentric Perception and Action for Robot Learning @ RSS2025
032
Théo Vincent @theo-vincent.bsky.social · 24/03/2025
Very happy to announce that iterated Q-Network (i-QN) has been published in TMLR 🎉 i-QN learns several Bellman iterations in parallel instead of learning them sequentially via repeated target updates ✨ This directly translates to performance improvements on the Atari and MuJoCo benchmarks 🚀
061
Reposted by Théo Vincent
TMLR Published Papers @tmlr-pub.bsky.social · 23/02/2025
Iterated $Q$-Network: Beyond One-Step Bellman Updates in Deep Reinforcement Learning Théo Vincent, Daniel Palenicek, Boris Belousov, Jan Peters, Carlo D'Eramo Action editor: Pablo Castro openreview.net/forum?id=Lt2H8Bd8jF #reinforcement #iterative #iterations
032
Reposted by Théo Vincent
Eugene Vinitsky 🍒 @eugenevinitsky.bsky.social · 18/11/2024
The RL (and some non-RL folks) starter pack is almost full. Pretty clear that the academic move here has succeeded go.bsky.app/3WPHcHg
1210632