Reposted by Giovanni Monea

We recently pushed an update to our in-context RL paper. Usually, updates don't justify a post, but this one is exceptionally contentful -> 🧵
tl;dr: all the findings are stronger, and the behaviors are super cool!
arxiv.org/abs/2410.05362