Sign in

Liv d'Aliberti

@liv-daliberti.bsky.social
35 followers 5 following 1 posts

AI Researcher

PostsRepliesMedia
Reposted by Liv d'Aliberti
Marcel Hussing @marcelhussing.bsky.social · 22/05/2026
🚨 New Preprint Alert: Behavior-Consistent Deep Reinforcement Learning 🚨 TLDR: We introduce an approach that achieves behavioral similarity across independent algorithm executions in continuous state-action space deep RL.
1294
Reposted by Liv d'Aliberti
Manoel Horta Ribeiro @manoelhortaribeiro.bsky.social · 05/01/2026
Do reasoning models have real “Aha!” moments—mid-chain realizations where they intrinsically self-correct? In a new pre-print, “The Illusion of Insight in Reasoning Models," led by @liv-daliberti.bsky.social we provide strong evidence that they do not! 📜: arxiv.org/abs/2601.00514
710516