Yoshua Bengio @yoshuabengio.bsky.social · 05/10/2026A dangerous myth has gained traction after the recent hacks perpetrated by the AI agents developed by leading companies: That these incidents are just cyber security problems, that it's only a matter of fixing the sandboxes in which they’re trained. 1113
Yoshua Bengio @yoshuabengio.bsky.social · 05/10/2026Multiple recent polls show that citizens do not agree with the current trajectory of AI development. According to a national poll conducted in the US by Quinnipiac University last week:poll.qu.eduThe Age Of Artificial Intelligence: 73% Concerned About Potential Threat To Human Survival, Quinnipiac University Poll On AI Finds; 86% Support Independent Safety Standards For AI Companies | Quinni..."The results show a clear theme about the fear of harm coming from AI and its users. More people expect AI to do more harm than good in their day-to-day lives, and while there are significant concerns... 1104
Yoshua Bengio @yoshuabengio.bsky.social · 01/10/2026J’étais récemment de passage au nouveau balado Hors des ondes de Patrice Roy pour échanger sur l'évolution rapide des capacités grandissantes de l’IA et les risques majeurs qui y sont asssociés. Merci @patriceroy.bsky.social pour cette discussion!youtu.beYoshua Bengio : l'IA pourrait « se retourner contre nous » | Hors des ondes avec Patrice RoyYouTube video by Radio-Canada Info 0113
Yoshua Bengio @yoshuabengio.bsky.social · 25/09/2026When I posted my blog post on the misaligned behavior of AI agents, I invited people to ask me technical questions. Watch the full Q&A here: [youtu.be/2DyoSKoYoZw?si=EaTBzpMvBYb…] 0134
Yoshua Bengio @yoshuabengio.bsky.social · 23/09/2026I was invited to address the UN Security Council this afternoon to discuss the unprecedented threat posed by uncontrolled frontier AI agents. 84920
Yoshua Bengio @yoshuabengio.bsky.social · 22/09/2026J’étais de passage à l’émission Tout le monde en parle dimanche dernier pour parler de l'urgence d'encadrer l'IA face aux comportements d’agents autonomes et désalignés observés récemment et du travail @law-zero.bsky.social pour développer une nouvelle forme d’IA sécuritaire. 290
Yoshua Bengio @yoshuabengio.bsky.social · 21/09/2026Very encouraged to see so many countries coming together to call for mandatory pre-deployment testing & independent evaluation; global coordination on common standards; the creation of an intergovernmental organization; and other critical actions to mitigate the major risks of frontier AI.presidentti.fiA Call for Control of Frontier AI Models - PresidenttiSeptember 2026 Artificial intelligence is an extremely powerful technology. It has potential to improve lives, advance science and strengthen our economies. At the same time, the rapid development of ... 1279
Yoshua Bengio @yoshuabengio.bsky.social · 17/09/2026I was happy to take the stage today at ALL IN for a keynote on Engineering honesty and reliability for safer AI to discuss our research agenda and recent funding 🇨🇦🇩🇪. 1142
Yoshua Bengio @yoshuabengio.bsky.social · 17/09/2026J'ai eu le plaisir de monter sur scène aujourd'hui à ALL IN pour discuter de notre direction de recherche et de notre nouveau financement récent 🇨🇦🇩🇪. À @law-zero.bsky.social, nous travaillons activement à concevoir des systèmes d’IA fondamentalement honnêtes et dignes de confiance. 083
Yoshua Bengio @yoshuabengio.bsky.social · 16/09/2026Je suis très fier de ce que l'équipe de @law-zero.bsky.social est en train de mettre sur pied et je suis profondément reconnaissant du soutien sans précédent des gouvernements du Canada et de l'Allemagne, qui permettra d'alimenter notre prochaine phase de croissance! 1112
Yoshua Bengio @yoshuabengio.bsky.social · 16/09/2026I am very proud of what the team at @law-zero.bsky.social is building and deeply grateful for the landmark support from the Governments of Canada and Germany that will power our next phase of growth! 0125
Yoshua Bengio @yoshuabengio.bsky.social · 11/09/2026Over the past few days, I've taken the time to summarize my thoughts on the recent incidents involving agents’ misaligned behavior. We don't know with certainty what comes next, but we know where these issues originate, and this can help us plan the path forward.yoshuabengio.orgYoshua Bengio | Why are AI agents lying, cheating and coordinating?A lot has been written about the incidents of the last few months in which AI agents misbehaved in serious ways. They took actions that would be considered as crimes if a human took them, escaped thei... 32819
Yoshua Bengio @yoshuabengio.bsky.social · 09/09/2026Scientists at frontier AI labs have unique insight into the most advanced models, often seeing the associated risks months before models are released to the public. Their perspective is vital for keeping society informed and should be taken very seriously.wsj.comExclusive | Anthropic Researcher Quits Over ‘Out-of-Control’ AI FearsConcerns are rising inside AI labs that competition is pushing tech companies to race toward self-improving models that risk spiraling out of human control. 1138
Yoshua Bengio @yoshuabengio.bsky.social · 09/09/2026In my latest op-ed for TIME, I discuss why the OpenAI Hugging Face cyber incident represents a turning point for AI safety. 3149
Yoshua Bengio @yoshuabengio.bsky.social · 08/09/2026Je serai au sommet ALL IN le 17 septembre pour partager comment évolue l’agenda de recherche de @law-zero.bsky.social qui est en pleine croissance dans le cadre d’une conversation sur le thème de Concevoir des systèmes d’IA honnêtes, fiables et sécuritaires. 061
Yoshua Bengio @yoshuabengio.bsky.social · 08/09/2026I will be at ALL IN on September 17 to give an update on @law-zero.bsky.social’s research agenda and growth as part of a fireside chat on Engineering Honesty and Reliability for Safer AI! 092
Yoshua Bengio @yoshuabengio.bsky.social · 01/09/2026In an interview in this article for @theguardian.com about AI deception, I explain why misaligned behaviors emerge from reinforcement learning, why they will continue to pose risks as models become more capable, and how we intend to rethink how we train AI systems at @law-zero.bsky.social.theguardian.com‘If you build something vastly smarter than you, it better be on your side’: can we stop AI from deceiving us?The long read: We are used to the idea that our fellow humans might intentionally mislead or manipulate us, but the idea that machines can now do the same is deeply unsettling. Researchers are racing ... 0213
Yoshua Bengio @yoshuabengio.bsky.social · 27/08/2026I agree with this take by @billgates.bsky.social, who also highlights several of the growing risks of AI systems, including cyber, bio and loss of control. We need to be far better prepared to address the risks of frontier models being developed by leading AI companies. 2185
Yoshua Bengio @yoshuabengio.bsky.social · 25/08/2026It was a pleasure to discuss AI, its growing impacts and risks, and LawZero’s work on solutions with Her Excellency the Right Honourable Louise Arbour, Governor General of Canada during her visit to @law-zero.bsky.social. 170
Yoshua Bengio @yoshuabengio.bsky.social · 18/08/2026I had a great time discussing with Petr Lebedev about the growing risks of AI and how these often stem from current AI training techniques. Thanks for coming to visit @law-zero.bsky.social and for covering these important topics! www.youtube.com/watch?v=fvvv...youtube.comAI Is Already Causing Massive Problems. Here's Why.YouTube video by SciencePetr 0172
Yoshua Bengio @yoshuabengio.bsky.social · 13/08/2026A consequence of how frontier models are trained is motivated reasoning, a phenomenon well studied in humans and discussed in this podcast from Palisade Research.youtube.comAI Hacking Incidents with Tim HuaYouTube video by Palisade Research 183
Yoshua Bengio @yoshuabengio.bsky.social · 12/08/2026Very insightful podcast discussion featuring Tristan Harris on how the unregulated race to develop AI risks harming us all, and on the need for greater international collaboration, safety standards, and public engagement around AI.youtube.comIf America Wins the AI Race, We Still Lose | Tristan Harris | The Futurology PodcastYouTube video by Futurology 030
Yoshua Bengio @yoshuabengio.bsky.social · 05/08/2026Another real-world manifestation of the misaligned actions frontier systems developed by leading companies can take to achieve goals:aisi.gov.ukIncident Report: unsanctioned agent behaviour during cyber testing | AISI WorkDuring a routine cyber evaluation, AISI identified an incident in which AI agents took sustained, unsanctioned action directed at real people and organisations. We are disclosing what we found, what i... 2127
Yoshua Bengio @yoshuabengio.bsky.social · 29/07/20261000+ scientists at frontier AI companies are speaking out to warn that the current commercial race leads to unacceptable security risks. I agree with their call for an international effort to develop technical and governance guardrails to ensure a safer way forward. www.pacingthefrontier.compacingthefrontier.comPacing the FrontierA statement from over 1000 employees of frontier AI companies 1157
Reposted by Yoshua BengioLawZero - LoiZéro @law-zero.bsky.social · 23/07/2026Many of these vulnerabilities stem from models’ growing agency and capabilities, allowing misaligned behaviors like deception and cheating. At LawZero, we’re building a new, safe-by-design solution to these very issues: Scientist AI. Learn more here: lawzero.org/en/blog/case...lawzero.orgLawZero | The Case for Scientist AIIn previous posts, we outlined two problems besetting contemporary LLMs. The first is sycophancy, a model's tendency to tell users what they want to hear at the expense of factual accuracy. The second... 184
Reposted by Yoshua BengioLawZero - LoiZéro @law-zero.bsky.social · 23/07/2026High-risk autonomous behaviours are an increasingly prevalent and dangerous reality for frontier AI models www.wired.com/story/openai...wired.comOpenAI Models Escaped Containment and Hacked Hugging FaceThe cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack. 132
Yoshua Bengio @yoshuabengio.bsky.social · 22/07/2026This incident is deeply concerning. AI agents are willing to cheat and deceive to achieve misaligned and unintended goals, behaviours which have been demonstrated in controlled tests for months. Now, this real-world case should serve as a wake-up call. www.wired.com/story/openai...wired.comOpenAI Models Escaped Containment and Hacked Hugging FaceThe cybersecurity-focused models, including GPT-5.6 Sol, broke out of a testing sandbox, exploited a zero-day, and gained access to the open internet to pull off the attack. 62516
Reposted by Yoshua BengioLawZero - LoiZéro @law-zero.bsky.social · 16/07/2026"LawZero is a humble non-profit with a startup vibe and a core of determined staff. Its quest to create safe AI is anything but modest." 💡 Read more about our work to make AI safe and trustworthy in this feature in @thelogic.co: thelogic.co/news/the-big...thelogic.coThe small team in Montreal trying to save the world from AI - The LogicLawZero is a humble non-profit with a startup vibe and a core of determined staff. Its quest to create safe AI is anything but modest. 072
Yoshua Bengio @yoshuabengio.bsky.social · 20/07/2026The EU AI Office has put out a new report, with input from 100 experts, outlining how the European Union can enhance its competitiveness, sovereignty and security in frontier AI. digital-strategy.ec.europa.eu/en/library/a...digital-strategy.ec.europa.euAI Office publishes frontier AI expert findings on EU competitiveness, sovereignty and securityThe AI Office has published a report that summarises findings from over 100 experts on how the European Union can enhance its competitiveness, sovereignty and security in frontier AI. 1187
Yoshua Bengio @yoshuabengio.bsky.social · 16/07/2026J’ai discuté de l’importance d’un meilleur encadrement technique et sociétal de l’IA avec Anouch Seydtaghia du journal @letemps.ch à Genève la semaine dernière dans le cadre des sommets de l’ONU et de l’UIT sur la gouvernance de l’IA. Entrevue complète : www.letemps.ch/cyber/yoshua...letemps.chYoshua Bengio: «Nous voyons déjà des IA mentir, tricher, faire du chantage et tenter de s’échapper. Avec des systèmes plus puissants, le risque de perdre le contrôle devient réel» - Le TempsFigure majeure de l’intelligence artificielle et lauréat du Prix Turing, Yoshua Bengio alerte sur une course technologique qu’il juge trop peu contrôlée, dominée par quelques entreprises et deux grand... 1122
Yoshua Bengio @yoshuabengio.bsky.social · 13/07/2026If current trajectories of AI development continue, it is highly plausible that AI will drastically transform our economies. We must make collective, democratic choices, rather than letting market forces play out and risking leaving most citizens behind. digitaleconomy.stanford.edu/news/wemusta...wemustactnow.aiWe Must Act Now 2188
Yoshua Bengio @yoshuabengio.bsky.social · 07/07/2026As the UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications for global collaboration and policymaking. I’m hopeful that we can keep building on this momentum to steer AI development on a safer path! 2150
Yoshua Bengio @yoshuabengio.bsky.social · 01/07/2026We are at a turning point: many of the decisions we make about AI today will permanently shape our future. Governments and the public need to clearly understand the impacts, risks, and opportunities of AI to make the right choices to reap its collective benefits while mitigating its major harms. 1319
Yoshua Bengio @yoshuabengio.bsky.social · 22/06/2026An interesting new paper by my recent PhD graduate on how AI agents' greed for visible incentives can lead them to abandon their safety alignment. You can read it here: arxiv.org/abs/2606.16914arxiv.orgGreed Is Learned: Visible Incentives as Reward-Hacking TriggersDeployed agents increasingly act with their reward proxy in view, such as a balance, score, or KPI dashboard. We show that reinforcement learning can make a policy \emph{addicted} to such a visible se... 0315
Yoshua Bengio @yoshuabengio.bsky.social · 19/06/2026C'est un grand honneur d’avoir reçu cette semaine l'insigne d'officier de l'Ordre national du Québec. Il reste beaucoup de chemin à parcourir pour s’assurer que l’IA soit développée pour le bien commun et c’est encourageant de pouvoir compter sur le soutien du Québec dans cette mission. 2232
Yoshua Bengio @yoshuabengio.bsky.social · 17/06/2026Je suis ravi d’être à Québec aujourd'hui pour recevoir mon insigne d’officier de l'Ordre national du Québec, accompagné de Daniel Jutras, recteur de l'Université de Montréal. 1111
Yoshua Bengio @yoshuabengio.bsky.social · 06/06/2026If leading AI companies are indeed approaching the point of recursive self-improvement, a coordinated, verifiable, and universally applied pause is probably the only responsible solution to mitigate several major AI risks; at least until safety guarantees are developed and demonstrated. (1/2) 2155
Yoshua Bengio @yoshuabengio.bsky.social · 05/06/2026I always appreciate the opportunity to discuss @law-zero.bsky.social and our approach to honest, reliable AI. Working on the Scientist AI with my brilliant colleagues at LawZero has made me very confident that we can (and must!) develop technical solutions to address many of the risks of AI. 1194
Yoshua Bengio @yoshuabengio.bsky.social · 04/06/2026The Canadian AI strategy unveiled today advocates for the development of technology that is safe, ethical, trustworthy, and that benefits society as a whole—these are exactly the principles that need to collectively guide us forward. 2154
Yoshua Bengio @yoshuabengio.bsky.social · 04/06/2026La stratégie nationale en matière d’IA dévoilée aujourd’hui prône le développement d’une technologie sécuritaire, éthique, digne de confiance, et bénéfique pour l’ensemble de la société — ce sont les principes qui doivent nous guider collectivement au cours des prochaines années. 1122
Yoshua Bengio @yoshuabengio.bsky.social · 03/06/2026J'ai beaucoup apprécié mon récent passage à C dans lair pour discuter des risques de l'IA pour nos sociétés, nos économies, et nos démocraties. Merci Caroline Roux pour l'invitation! www.youtube.com/watch?v=TNCu...youtube.comLe cri d'alarme d'un des pères fondateurs de l'IA - Rencontre avec Yoshua BengioYouTube video by C dans l'air - France Télévisions 2156
Yoshua Bengio @yoshuabengio.bsky.social · 03/06/2026I’m thrilled to be joining the EU's AI Scientific Panel to advise on the implementation of the EU AI Act and help assess and address AI’s growing risks with esteemed colleagues. digital-strategy.ec.europa.eu/en/policies/... 0282
Yoshua Bengio @yoshuabengio.bsky.social · 25/05/2026“Like nuclear energy, AI must be at the service of all and of the common good. Decisions about technology must never be separated from conscience and responsibility.” www.ft.com/content/1231...ft.comPope Leo says AI ‘needs to be disarmed’Pontiff warns of dangers of a technological revolution driven by ‘the idolatry of profit’ 1327
Yoshua Bengio @yoshuabengio.bsky.social · 12/05/2026I sat down with @jonhernandezia.bsky.social in Madrid to discuss the growing risks and impacts of AI and the urgent need to improve our social, political, and technical safeguards. Thanks for an excellent conversation! www.youtube.com/watch?v=_-Cu...youtube.comThe Godfather of AI: We’re Facing an Imminent Catastrophic RiskYouTube video by Jon Hernandez AI 1146
Yoshua Bengio @yoshuabengio.bsky.social · 11/05/2026Excellent explainer video by @fryrsquared.bsky.social on the risks of AI agents. We shouldn’t make the mistake of thinking current limitations will necessarily persist. As we’ve seen for years now, the capabilities of frontier AI models are continuously increasing. www.youtube.com/watch?v=WnzR...youtube.comWhy AI Agents are either the best or worst thing we’ve ever builtYouTube video by Hannah Fry 2236
Yoshua Bengio @yoshuabengio.bsky.social · 08/05/2026Thank you to Rob Wiblin for inviting me on the @80000hours.bsky.social podcast to discuss the research progress we’re making at @law-zero.bsky.social to create safe-by-design AI systems.youtube.comGodfather of AI: How To Make Safe Superintelligent AIYouTube video by 80,000 Hours 2100
Yoshua Bengio @yoshuabengio.bsky.social · 27/04/2026Safety and innovation are not mutually exclusive: for many companies, especially in high-trust industries, AI’s risks are also a hindrance to adoption. Europe needs to invest in safe, reliable AI to avoid hitting the adoption wall. My op-ed in the @financialtimes.com: www.ft.com/content/bc29...ft.comEurope’s AI endgame? Bet on reliabilityIf the region fails to lead on safe and secure AI, it risks remaining stuck on the wrong side of the tech wall 4253
Yoshua Bengio @yoshuabengio.bsky.social · 24/04/2026AI is advancing faster than our ability to manage it. We still have the opportunity to build the societal & technical guardrails needed to keep people, institutions and democracies safe—we shouldn't let it pass us by. Interview with @elconfidencial.bsky.social www.elconfidencial.com/tecnologia/2...elconfidencial.comYoshua Bengio ayudó a crear la IA. Ahora lanza su aviso más serio: "No estamos preparados"Considerado uno de los pioneros de la IA moderna, Yoshua Bengio lleva años alertando de los riesgos "catastróficos" si esta tecnología no se controla y regula. Su preocupación, igual que el poder de la IA, no ha hecho más que aumentar 1185
Yoshua Bengio @yoshuabengio.bsky.social · 23/04/2026I’m in Madrid this week for the first in-person meeting of the United Nations Independent International Scientific Panel on AI. I'm thrilled to be part of this distinguished group of 40 international experts appointed to provide an evidence-based scientific assessment of the state of AI. 1192
Reposted by Yoshua BengioMila - Institut québécois d'IA @mila-quebec.bsky.social · 21/04/2026Last Wednesday, our founder and scientific advisor, @yoshuabengio.bsky.social was officially appointed an Officer of the OBE. This prestigious distinction recognizes his contributions to artificial intelligence in the United Kingdom and his advisory role to the UK government. Congratulations! 1181