Reposted by Liv d'AlibertiMarcel Hussing @marcelhussing.bsky.social · 22/05/2026🚨 New Preprint Alert: Behavior-Consistent Deep Reinforcement Learning 🚨 TLDR: We introduce an approach that achieves behavioral similarity across independent algorithm executions in continuous state-action space deep RL. 1294
Reposted by Liv d'AlibertiManoel Horta Ribeiro @manoelhortaribeiro.bsky.social · 05/01/2026Do reasoning models have real “Aha!” moments—mid-chain realizations where they intrinsically self-correct? In a new pre-print, “The Illusion of Insight in Reasoning Models," led by @liv-daliberti.bsky.social we provide strong evidence that they do not! 📜: arxiv.org/abs/2601.00514 710516