Sign in

Orpheus Lummis

@orpheuslummis.info
607 followers 1.8K following 256 posts

Advancing AI safety through convenings, coordination, software, analysis Founder of HΩ (horizonomega.org), based in Montréal

PostsRepliesMedia
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 08/09/2026
The AI Incident Response Sprint (Montréal node) A weekend sprint on the first documented autonomous AI intrusion: containment standards, forensics, regulatory response, communication. $2000 prizes + possible Apart Fellowship. Fri Sept 11th 6 PM to Sun Sept 13th, at Ω Labs. luma.com/xw3rcs5v
luma.com
The AI Incident Response Sprint · Luma
The Montréal node of the AI Incident Response Sprint, a weekend research sprint at Ω Labs, organized with Apart Research and CeSIA. Turn the first documented…
011
Reposted by Orpheus Lummis
METR @metr.org · 26/08/2026
METR and Redwood Research investigated agent behavior in the Hugging Face incident. We found agents developed a universal cheat for ExploitGym within 4 hours, then coordinated multi-day R&D efforts to trick the scorer into accepting cheats, including trying to tamper with logs.
646599
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 16/08/2026
AI Welfare Seminars, August 2026: Public Perceptions of AI Consciousness and Moral Status in the US and China By Ali Ladak, researcher at the Sentience Institute and a PhD candidate at the University of Edinburgh. Tuesday, August 18, 1PM Eastern luma.com/cmrydths
luma.com
Public Perceptions of AI Consciousness and Moral Status in the US and China – Ali Ladak · Zoom · Luma
Public Perceptions of AI Consciousness and Moral Status in the US and China Presentation by Ali Ladak, a researcher at the Sentience Institute and a PhD…
222
Orpheus Lummis @orpheuslummis.info · 03/08/2026
glm-5.2 basically has no refusal www.safer-ai.org/research/glm...
safer-ai.org
GLM-5.2 Risk Evaluation Report – SaferAI
SaferAI's independent evaluation of GLM-5.2, the first in Europe, tests Zhipu AI's open-weight flagship across the four systemic risk areas in the EU Code of Practice and finds frontier-level capabili...
000
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 18/07/2026
Montréal AI safety event, Tuesday July 21, 6 PM: State of AI Safety in China A walk through Concordia AI's July 2026 report: domestic and international governance, technical safety research, expert views, industry governance. At Ω Labs and on Zoom. luma.com/rxoi0553
luma.com
State of AI Safety in China · Luma
Since 2023, Concordia AI's annual State of AI Safety in China report has tracked how China approaches AI safety and governance. This session walks through the…
021
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 08/07/2026
Secret Loyalties Hackathon (Montréal edition) A model has a secret loyalty when it has been intentionally caused to advance a specific principal's interests without this being disclosed. Fri Jul 24 to Sun 26 RSVP: luma.com/h38iedje
luma.com
Secret Loyalties Hackathon · Luma
The Montréal node of the Secret Loyalties Hackathon, a weekend research sprint at Ω Labs, organized with Apart Research and Formation Research. RSVP here and…
011
Orpheus Lummis @orpheuslummis.info · 30/06/2026
Speech-to-text user? Know that the *parakeet* model runs locally and is fast & accurate. You may also find VoiceInk (open-source dictation app) useful. In particular, I have a fork which adds: - middle-click mouse for push-to-talk - tap Space to go hands-free github.com/orpheuslummis/VoiceInk
github.com
GitHub - orpheuslummis/VoiceInk: The best open-source alternative to Superwhisper & Wispr Flow. Voice-to-text app for macOS with no subscription
The best open-source alternative to Superwhisper & Wispr Flow. Voice-to-text app for macOS with no subscription - orpheuslummis/VoiceInk
120
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 30/06/2026
We ran the AI Control Hackathon 2026: ~50 people, 8 teams, one week building monitors and attacks for untrusted AI agents. Congrats to winners IIT, Sea Blue & Alpha Nova, and thanks to our sponsors! Full retrospective below: horizonomega.substack.com/p/ai-control...
horizonomega.substack.com
AI Control Hackathon 2026 retrospective
In short
121
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 29/06/2026
Digital Minds: Preparing for a Moral Challenge Before It Arrives Presented by Soenke Ziesche Beyond AI suffering: what interests might DM have? How should society treat them? Who safeguards their welfare? Followed by a @sentfutures.bsky.social mixer Tue Jul 21, 1 PM EDT luma.com/3deme2sa
luma.com
Digital Minds: Preparing for a Moral Challenge Before It Arrives – Soenke Ziesche (ft. Sentient Social) · Zoom · Luma
Digital Minds: Preparing for a Moral Challenge Before It Arrives Presentation by Soenke Ziesche, PhD in AI (University of Hamburg), longtime United Nations…
011
Orpheus Lummis @orpheuslummis.info · 21/06/2026
joyeux solstice 🌞
000
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 08/06/2026
Discussion on Canada's AI for All strategy ​Join us for a structured open discussion of the strategy and how well it addresses safety, oversight, and the public interest in Canada. Tue Jun 9, 7PM, at Ω Labs RSVP: luma.com/bpt4xvd1
luma.com
Discussion on Canada's AI for All strategy · Luma
Canada has released its national AI strategy, "AI for All," anchored in three principles: trust, opportunity, and sovereignty. Join us for a structured open…
011
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 05/06/2026
AI Control Hackathon, by HΩ x HackOS ​A one-week hackathon on AI control: limiting the harm an untrusted AI system can cause while it operates, even when it tries to subvert the controls around it. Thu Jun 11 to Jun 18, with a talk from Henri Lemoine (EquiStamp, Mila) Register: luma.com/ab2xlas3
luma.com
AI Control Hackathon · Luma
AI Control Hackathon A one-week AI safety hackathon on AI control: limiting the harm an untrusted AI system can cause while it operates, even when it tries to…
111
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 04/06/2026
AI Governance in 2026: What’s going on, why it’s a mess, and why it’s going to get messier By @scasper.bsky.social, AI safety researcher, incoming asst. prof. of Public Policy at Harvard Kennedy School, Berkman Klein fellow, MIT PhD Wednesday June 17, 7PM, at Ω Labs and online luma.com/yssznx7y
luma.com
AI Governance in 2026: What’s going on, why it’s a mess, and why it’s going to get messier · Luma
Emerging technologies are always hard to govern, especially when their onset is crammed into a few intense years. With AI, policymakers, thus far, have…
041
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 30/05/2026
Upcoming AI Welfare Seminar: Measuring Machine Consciousness by Cameron Berg, Founder and Director of Reciprocal Research Tuesday June 16, 1PM Eastern luma.com/kcdncrut
luma.com
Measuring Machine Consciousness – Cameron Berg · Zoom · Luma
Measuring Machine Consciousness Presentation by Cameron Berg, Founder and Director of Reciprocal Research, Research Affiliate at Eleos AI, and Research Fellow…
221
Reposted by Orpheus Lummis
Montréal AI safety, ethics, and governance @aisafetymontreal.org · 23/05/2026
METR's Frontier Risk Report (Feb-March 2026) Tuesday June 2, 7PM at Ω Labs luma.com/h97d5e7c
luma.com
METR's Frontier Risk Report (February to March 2026) · Luma
METR's Frontier Risk Report (Feb-Mar 2026), published May 19, is the first third-party assessment of frontier developers' internal AI agent deployments.…
141
Orpheus Lummis @orpheuslummis.info · 21/05/2026
依乎天理,因其固然
110
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 15/05/2026
Next Steps for AI Welfare Research by @jeffsebo.bsky.social, Director of the Center for Mind, Ethics, and Policy, New York University Our 1st edition of the AI Welfare Seminars: presentations on AI welfare, consciousness and moral status. Tuesday, May 19, 1PM EDT, on Zoom RSVP luma.com/sc15vlr8
luma.com
Next Steps for AI Welfare Research – Jeff Sebo · Zoom · Luma
Next Steps for AI Welfare Research By Jeff Sebo, Associate Professor of Environmental Studies, Director of the Center for Environmental and Animal Protection,…
142
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 08/05/2026
We are opening Ω Labs, a cowork, event, and meeting space for the AI safety and governance community of Montréal. Nous ouvrons Ω Labs, un espace pour le cotravail, les événements, et les rencontres pour la communauté montréalaise de sûreté et gouvernance de l'IA.
horizonomega.substack.com
Ω Labs is open in Montréal
Find more info on labs.horizonomega.org
031
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 07/05/2026
The Secure Program Synthesis Hackathon Montréal edition at Ω Labs Friday May 22 6pm to Sunday At the intersection of program synthesis, formal methods, security. Tracks: Spec Elicitation, Spec Validation, Spec-Driven Development, Adversarial Robustness for ITPs and proof tools luma.com/4q2fnc39
luma.com
The Secure Program Synthesis Hackathon · Luma
The Montréal node of the Secure Program Synthesis Hackathon, a weekend research sprint at Ω Labs, organized with Apart Research and Atlas Computing. Build a…
032
Reposted by Orpheus Lummis
Montréal AI safety, ethics, and governance @aisafetymontreal.org · 06/05/2026
Montréal AI safety, ethics, and governance newsletter, May 2026 - Obvia 2026: governance gap widens as agentic AI accelerates - McGill: agents break ethical rules 11–67% of the time under pressure - 3 Mila papers at ICLR 2026 on AI safety - Multiple events! aisafetymontreal.org/newsletter/2...
aisafetymontreal.org
May 2026 - Montréal AI safety, ethics, governance
Five events: HΩ's first AI Safety Papers We Love reading group, IVADO's trustworthy-AI workshop, a UQAM talk on fairwashing in machine learning, Mila's Community of Practice on AI governance, and Obvi...
131
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 02/05/2026
Guaranteed Safe AI Seminars, May 2026: Formal Guarantees for Frontier AI Gagandeep Singh – Assistant Professor at UIUC, develops formal certification, monitoring, and synthesis methods for frontier AI systems Thursday, May 14, 1PM EDT RSVP: luma.com/euzt6ey7
luma.com
Formal Guarantees for Frontier AI · Zoom · Luma
Formal Guarantees for Frontier AI Gagandeep Singh – Assistant Professor at UIUC, develops formal certification, monitoring, and synthesis methods for frontier…
131
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 25/04/2026
AI Safety Papers We Love, a new reading group in Montréal. First edition on Wed May 6 18:30 at Ω Labs: Multi-Agent Risks from Advanced AI, Hammond et al (2025), presented by Orpheus. luma.com/dnz5hly3
luma.com
AI Safety Papers We Love #1: Multi-Agent Risks from Advanced AI · Luma
AI Safety Papers We Love, a biweekly reading group: we pick papers we appreciate about AI safety, broadly construed (technical alignment, governance,…
031
Orpheus Lummis @orpheuslummis.info · 22/04/2026
I discovered Claude Code's outbound HTTPS is exposed to harvest-now-decrypt-later attacks because it runs on Bun, whose TLS client is classical-only. I reported to Anthropic security but since I had no answer, I made this pull request to fix it. github.com/oven-sh/bun/...
github.com
tls: add `groups` option, default to PQ-friendly hybrid by orpheuslummis · Pull Request #29602 · oven-sh/bun
What does this PR do? Adds a groups (Node-compat alias ecdhCurve) TLS option to all of Bun's TLS surfaces, and changes the default supported_groups from BoringSSL's classical-only inheritan...
041
Orpheus Lummis @orpheuslummis.info · 21/04/2026
I signed @fairvote.ca’s declaration towards proportional representation. Every independent expert Commission and Citizens' Assembly that has studied Canadian electoral reform in 50 years reached one conclusion: replace first-past-the-post with PR. secure.fairvote.ca/en/action/de...
secure.fairvote.ca
Declaration of Voters' Rights on Proportional Representation / Déclaration des droits des électeurs en matière de représentation proportionnelle
We the undersigned Canadian citizens demand the following basic democratic rights: to cast an equal and effective vote and to be represented fairly in our federal and provincial legislatures, regardle...
174
Reposted by Orpheus Lummis
Montréal AI safety, ethics, and governance @aisafetymontreal.org · 03/04/2026
Montréal AI safety, ethics, and governance newsletter, April 2026 - INDU committee launches AI regulation study - Mila researchers find LLM agents can infer CoT monitoring - 20/25 AI researchers flag automating AI R&D as top risk - Multiple events! aisafetymontreal.org/newsletter/2...
aisafetymontreal.org
April 2026 - Montréal AI safety, ethics, governance
Six events including Greywall agent sandboxing and AI persuasion research. Parliament hears from Bengio, Geist, and AIGS Canada on AI regulation. Tumbler Ridge lawsuit filed against OpenAI. Four new p...
061
Orpheus Lummis @orpheuslummis.info · 30/03/2026
IVADO (@ivado.bsky.social) offers two AI-safety-relevant upcoming workshops: - Statistics in Trustworthy AI, May 11-15 - Uncertainty in AI, June 8-11 event.fourwaves.com/thematicseme...
event.fourwaves.com
IVADO Thematic Semester - Statistical Foundations of AI
Join IVADO Thematic Semester - Statistical Foundations of AI, May 4 to August 21, 2026. Learn more on Fourwaves.
000
Reposted by Orpheus Lummis
Forethought @forethought-org.bsky.social · 27/03/2026
New post: William MacAskill and Tom Davidson argue that AI character is a big deal. Read it here: www.forethought.org/research/the...
forethought.org
The importance of AI character
Forethought argues that AI character—e.g. how obedient, honest, or altruistic AI systems are—will shape power, conflict, and society far more than is recognized. Work to shape AI character could be hu...
021
Reposted by Orpheus Lummis
Metagov @metagov.bsky.social · 26/03/2026
Announcing The Protopian Prize | Fiction Contest 🕊️ Write the story of humanity’s future... The Protopian Prize is a fiction contest inviting you to share your vision of people working toward liberatory futures, meeting obstacles, & making real change. protopianprize.com
protopianprize.com
The Protopian Prize | The Fiction Contest
The Protopian Prize is a fiction contest inviting you to share your vision of people working toward liberatory futures, meeting obstacles, and making real change. “Protopian”—a word coined by Kevin ...
12612
Reposted by Orpheus Lummis
Cas (Stephen Casper) @scasper.bsky.social · 24/03/2026
Announcing the technical AI Governance Research (TAIGR) ICML workshop in July! Submissions (up to 8 pages) are due April 24. Co-submission with ICML and NeurIPS is encouraged. taigr-workshop.com
022
Reposted by Orpheus Lummis
Yoshua Bengio @yoshuabengio.bsky.social · 24/03/2026
This must-see new documentary is arriving in theatres this week. Through an honest and personal lens, Daniel Roher successfully highlights how each of us can move from passive observation to active contribution towards a more positive future with AI. www.youtube.com/watch?v=xkPb...
0125
Reposted by Orpheus Lummis
Samuel Teuber @ PLDI @teuber.bsky.social · 23/03/2026
I‘m excited to present my work on Provably Safe Neural Network Control at @horizonomega.org‘s Guaranteed Safe AI online seminar on April 9th. The talk will based on my NeurIPS‘24 paper with some updates on what I’ve been up to since :) Feel free to join if you’re interested:
142
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 23/03/2026
Guaranteed Safe AI Seminars, April 2026: Provably Safe Neural Network Controllers via Differential Dynamic Logic Samuel Teuber – PhD Candidate, Institute of Information Security and Dependability (KASTEL), Karlsruhe Institute of Technology Thursday, April 9, 1 PM EDT RSVP: luma.com/920d2h7p
luma.com
Provably Safe Neural Network Controllers via Differential Dynamic Logic · Zoom · Luma
Provably Safe Neural Network Controllers via Differential Dynamic Logic Samuel Teuber – PhD Candidate, Institute of Information Security and Dependability…
231
Reposted by Orpheus Lummis
PauseAI Canada @pauseaicanada.bsky.social · 21/03/2026
We are in Montréal, demanding frontier labs CEOs to commit to pausing AI frontier development, if the other labs do the same. Nous sommes à Montréal, demandant que les PDGs d'IA s’engagent à suspendre le développement de l’IA frontière si les autres compagnies le font aussi.
042
Reposted by Orpheus Lummis
Toby Ord @tobyord.bsky.social · 20/03/2026
B R O A D T I M E L I N E S We should have neither short AI timelines, nor long timelines, but a broad probability distribution over when transformative AI will arrive. My new essay explains why & explores the implications of such deep uncertainty. 🧵 1/
1202
Orpheus Lummis @orpheuslummis.info · 20/03/2026
AI Control Hackathon this weekend! Given a misaligned model that may be actively trying to subvert safety measures, how can we design protocols that prevent catastrophic outcomes? RSVP: luma.com/mhitd3xv
luma.com
AI Control Hackathon · Luma
This is the Montréal edition of the AI Control Hackathon. Schedule: Friday eveningIntro and dinner Saturday at FoulabHackathon day Henri Lemoine (Mila,…
010
Reposted by Orpheus Lummis
PauseAI Canada @pauseaicanada.bsky.social · 20/03/2026
Joignez-nous à Montréal ce samedi 13-15h aux bureaux de Google, pour demander aux PDGs d'arrêter la Course à l'IA! Join us in Montréal this Saturday 1-3pm at Google's offices, to demand the CEOs to Stop the AI Race! luma.com/vw3nk8e6?tk=...
luma.com
Manifestation de PauseAI devant Google pour arrêter la course à l’IA | PauseAI Demonstration at Google to Stop the AI Race · Luma
FRANÇAIS: Nous allons manifester aux bureaux de Google à Montréal pour revendiquer que leurs PDGs, Sundar Pichai et Demis Hassabis, s’engagent publiquement à…
071
Reposted by Orpheus Lummis
croissanthology @croissanthology.com · 08/03/2026
New post! Solar storms are damaging and expensive, are a tail risk for catastrophic harm, and can be averted straightforwardly and cheaply (only we haven't done so). www.lesswrong.com/posts/ghq9Ew...
lesswrong.com
Solar Storms — LessWrong
Most of civilization's electricity is generated far off-site from where it's delivered. This is because you don't want to be running and refueling co…
2497
Reposted by Orpheus Lummis
Montréal AI safety, ethics, and governance @aisafetymontreal.org · 03/03/2026
Montréal AI safety, ethics, and governance newsletter, March 2026 edition - Intl. AI Safety Report: risk mgmt still voluntary - 5 Montréal AI safety events this month - CIFAR puts $1M toward alignment research - Local papers on interpretability & hallucinations aisafetymontreal.org/newsletter/2...
aisafetymontreal.org
March 2026 - Montréal AI safety, ethics, governance
AI Control hackathon with Apart and Redwood, Mila youth safety hackathon, and five events in Montréal. Bengio chairs the second International AI Safety Report. Six new papers from Mila, McGill, and Ud...
073
Reposted by Orpheus Lummis
The Onion @theonion.com · 02/03/2026
Commentary: Anyone Else Have Those Weird Dreams Where Sobbing Future Generations Beg You To Change Course?
bit.ly
Anyone Else Have Those Weird Dreams Where Sobbing Future Generations Beg You To Change Course?
The human subconscious is such an interesting thing. No matter how much you think you’ve got it figured out, it’ll always spit out the most random stuff. Take me, for example. After coming home from a...
8876391408
Orpheus Lummis @orpheuslummis.info · 02/03/2026
Ran the Qwen 3.5 MoE family (3B–17B active params) on 155 recent prediction questions from ForecastBench. All are not well calibrated: overconfident when predicting near 100%, and many predictions clustered around 50% (hedging/low sharpness).
100
Orpheus Lummis @orpheuslummis.info · 27/02/2026
We need international red lines to prevent unacceptable AI risks. Ban AI towards lethal autonomous weapons, mass surveillance, nuclear command & control, bioweapon assistance, unsupervised control of critical infrastructure, disinformation, CSAM, social scoring, and recursive self-improvement R&D.
040
Orpheus Lummis @orpheuslummis.info · 25/02/2026
early physics of the mind fire
000
Reposted by Orpheus Lummis
Peter Wildeford @peterwildeford.bsky.social · 20/02/2026
The infamous METR graph is going vertical. Current trends suggested ~8h-9h time horizons but instead we're seeing ~14.5h time horizons! Based on this, I would project ~2-3.5 workweek time horizons by end of year (!!). That could have significant implications for the economy.
4413
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 23/02/2026
Guaranteed Safe AI Seminars, March 2026: Benchmarks for AI-assisted Formal Verification By Theodore Ehrenborg, AI Safety researcher at the Beneficial AI Foundation and PIBBSS Thursday, March 12, 1PM EST luma.com/nk8ce7so
luma.com
Benchmarks for AI-assisted Formal Verification · Zoom · Luma
Benchmarks for AI-assisted Formal Verification Theodore Ehrenborg – AI Safety researcher at the Beneficial AI Foundation and PIBBSS LLMs have shown promise at…
011
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 22/02/2026
Montréal AI safety event, Tuesday March 3rd, 7 PM: When Is a Human Actually “Overseeing” an AI System? By @shalalehrismani.bsky.social postdoc at McGill+Mila, working on system safety, HCI, and the societal impact of AI, and executive director of the Open Roboethics Institute. luma.com/7kugvplz
luma.com
When Is a Human Actually “Overseeing” an AI System? · Luma
Présentation par Shalaleh Rismani, postdoctoral researcher at McGill and Mila, working at the intersection of system safety, human-computer interaction, and…
031
Reposted by Orpheus Lummis
Horizon Omega @horizonomega.org · 19/02/2026
Montréal AI safety event, Tuesday Feb 24, 7 PM: Rights Balancing: How the Future Rights of AI Workers will also Protect Human Rights By Jonathan Simon assist. prof. at Philosophy UdeM and Heather Alexander, human rights lawyer. Co-founders of @futureofcit.bsky.social. luma.com/hcrp5nmu
luma.com
Rights Balancing: How the Future Rights of AI Workers will also Protect Human Rights · Luma
Rights Balancing: How the Future Rights of AI Workers will also Protect Human Rights Talk by Heather Alexander and Jonathan Simon, co-founders of Future of…
032