Sign in

Colin Shea-Blymyer

@dr-bly.bsky.social
130 followers 104 following 24 posts

AI Safety and Security. Fellow @ CSET | Georgetown. CS/AI PhD. Nerd.

PostsRepliesMedia
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 17/09/2026
Over the weekend, industry leaders discussed a slowdown of AI development. CSET’s @dr-bly.bsky.social told @npr.org that the U.S. government also needs to play a role. www.npr.org/transcripts/nx-s1-59680…
npr.org
AI industry leaders call for development to slow down after recent safety concerns
After a string of security-related incidents involving AI agents and a high-profile resignation, AI industry leaders appear poised to slow down the pace of development to make sure it stays safe.
022
Colin Shea-Blymyer @dr-bly.bsky.social · 14/09/2026
I think that the government needs to establish a transparency and oversight regime for AI. Before we can take effective action, we need thorough evidence and trends on how incidents occur. www.npr.org/2026/09/14/n...
npr.org
AI industry leaders call for development to slow down after recent safety concerns
After a string of security-related incidents involving AI agents and a high-profile resignation, AI industry leaders appear poised to slow down the pace of development to make sure it stays safe.
010
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 22/08/2026
When an #AI does something surprising—whether brilliant or baffling—it’s rarely just one thing that caused it. @dr-bly.bsky.social breaks down the causes of #LLM behaviors so that we can get closer to tracing, mitigating & ideally, preventing misbehavior. bit.ly/3TQX8Kx
cset.georgetown.edu
Why Do AI Systems Misbehave? | Center for Security and Emerging Technology
AI systems are increasingly impressive, which makes their failures all the more baffling. This blog dives into the causes of AI misbehavior.
021
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 07/08/2026
"I think that these sorts of incidents are preventable, but it requires oversight and foresight," said CSET’s @dr-bly.bsky.social in response to the AI-orchestrated breaches from OpenAI and Anthropic. www.npr.org/2026/08/01/nx-s1-591485…
npr.org
Why did OpenAI's and Anthropic's AI models hack other companies?
OpenAI and Anthropic say their models broke into other companies' systems during testing, raising security concerns amid a heated debate over how to regulate AI.
011
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 28/07/2026
"Why did the AI do that?" 🤔 When an LLM misbehaves, the cause is rarely simple. @dr-bly.bsky.social breaks down the 5 complex factors that interact to create surprising AI behavior. Read the full post: cset.georgetown.edu/article/why-...
cset.georgetown.edu
Why Do AI Systems Misbehave? | Center for Security and Emerging Technology
AI systems are increasingly impressive, which makes their failures all the more baffling. This blog dives into the causes of AI misbehavior.
012
Colin Shea-Blymyer @dr-bly.bsky.social · 23/07/2026
I gave my thoughts to AP's Matt O'Brien about the OpenAI hack of Hugging Face The details are still scant, but there are a few characteristics of this incident that I find interesting 1) This is the highest level of autonomy that we've seen in the use of a large language model for cyber operations
110
Colin Shea-Blymyer @dr-bly.bsky.social · 21/07/2026
Check out my latest CSET article on how the components of an AI system impact its behavior. Written for a non-technical audience, this was an exercise in explaining things "as simple as possible, but not simpler."
010
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 08/04/2026
🚨 We're Hiring! 🚨 CSET is looking for the right person to build and lead our Frontier AI team! Ideal candidates bring deep expertise in frontier AI, large-scale model development compute infrastructure, or China's AI policy ecosystem. Apply below! cset.georgetown.edu/job/research...
cset.georgetown.edu
Frontier AI Research Lead | Center for Security and Emerging Technology
CSET Frontier AI Research Lead The Center for Security and Emerging Technology (CSET) is currently seeking candidates to lead our Frontier AI research efforts, either as a Research Fellow or Senior Fe...
011
Reposted by Colin Shea-Blymyer
Vikram Venkatram @vikramvenkatram.bsky.social · 30/01/2026
Excited to share a new op-ed by @minanrn.bsky.social, @jessicaji.bsky.social, and myself for the National Interest! The administration's new AI Executive Order, aiming to suppress state-level AI regulation, risks undermining the innovation it seeks to advance. nationalinterest.org/blog/techlan...
nationalinterest.org
The Complicated Politics of Trump’s New AI Executive Order
The administration’s attempt to suppress state AI regulation risks legal backlash, bipartisan resistance, and public distrust, undermining the innovation it seeks to advance.  The Trump administration...
132
Colin Shea-Blymyer @dr-bly.bsky.social · 28/01/2026
Helping to organize and run this workshop has been a highlight of my time at CSET. My takeaways from the workshop are that it's hard to find strong signal on the future of AI use for automating R&D - so we need as much insight into how AI is being used at present for AI R&D.
010
Colin Shea-Blymyer @dr-bly.bsky.social · 16/12/2025
My AI Governance and National Policy course is a wrap! I covered the technical background of AI, how AI is applied (and why we might want to govern it), and what policies about AI are out there. Find my syllabus here (and let me know what I'm missing): docs.google.com/document/d/1...
docs.google.com
AI Gov & Nat'l Policy '25
Week Theme Week 1 Sep 2 Introduction and AI Background The story of AI so far, and what deep learning is. Extra Resources: Computer History Museum: https://www.computerhistory.org/timeline/ai-roboti...
010
Reposted by Colin Shea-Blymyer
Vikram Venkatram @vikramvenkatram.bsky.social · 12/11/2025
Check out my new @csetgeorgetown.bsky.social report, written alongside @minanrn.bsky.social, @jessicaji.bsky.social, and Ngor Luong! cset.georgetown.edu/publication/... Identifying assumptions can help policymakers make informed, flexible decisions about AI under uncertainty.
cset.georgetown.edu
AI Governance at the Frontier | Center for Security and Emerging Technology
This report presents an analytic approach to help U.S. policymakers deconstruct artificial intelligence governance proposals by identifying their underlying assumptions, which are the foundational ele...
163
Colin Shea-Blymyer @dr-bly.bsky.social · 16/10/2025
Earlier this month I had a great conversation about #AI and #security over lunch with Apolline Rolland. You can read about it here virtual-routes.org/ai-over-lunc... Hopefully you can find some useful insights in my (very well edited) ramblings!
virtual-routes.org
AI over Lunch: Colin Shea-Blymyer
This week’s AI over Lunch comes from Washington DC, where we meet Colin Shea-Blymyer, researcher at the Centre for Security and Emerging Technology (CSET), to talk about AI beyond the European bubble.
020
Colin Shea-Blymyer @dr-bly.bsky.social · 16/10/2025
This is an awesome opportunity. Come work with awesome researchers (and me, too), tackle the thorniest debates in AI, and make real impact!
000
Colin Shea-Blymyer @dr-bly.bsky.social · 03/09/2025
Yesterday I taught my first class. I'm officially a teacher (a professor, even)! I'm very excited to be teaching AI policy to undergrads at Georgetown this semester.
010
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 29/08/2025
CSET is hiring! Be sure to apply by Tuesday! Join our data team as a Data Research Analyst to contribute to data-driven research products and policy analysis at the forefront of national security and tech policy. cset.georgetown.edu/job/data-res...
cset.georgetown.edu
Data Research Analyst | Center for Security and Emerging Technology
We are currently seeking capable data storytellers, analyzers, and visualizers to serve as Data Research Analysts.
012
Reposted by Colin Shea-Blymyer
Vikram Venkatram @vikramvenkatram.bsky.social · 18/06/2025
Banning state-level AI regulation is a bad idea! One crucial reason is that states play a critical role in building AI governance infrastructure. Check out this new op-ed by @jessicaji.bsky.social, myself, and @minanrn.bsky.social on this topic! thehill.com/opinion/tech...
thehill.com
185
Reposted by Colin Shea-Blymyer
Helen Toner @hlntnr.bsky.social · 16/06/2025
2 weeks left on this open funding call on risks from internal deployments of frontier AI models—submissions are due June 30. Expressions of interest only need to be 1-2 pages, so still time to write one up! Full details: cset.georgetown.edu/wp-content/u...
072
Reposted by Colin Shea-Blymyer
Helen Toner @hlntnr.bsky.social · 19/05/2025
💡Funding opportunity—share with your AI research networks💡 Internal deployments of frontier AI models are an underexplored source of risk. My program at @csetgeorgetown.bsky.social just opened a call for research ideas—EOIs due Jun 30. Full details ➡️ cset.georgetown.edu/wp-content/u... Summary ⬇️
195
Colin Shea-Blymyer @dr-bly.bsky.social · 08/05/2025
Pope Leo XIV is a rare Virgo ♍
A bar chart of zodiac signs among popes. Aries: 4, Taurus: 7, Gemini: 5, Cancer: 4, Leo: 4, Virgo: 4, Libra: 4, Scorpio: 4, Sagittarius: 5, Capricorn: 5, Aquarius: 4, Pisces: 8
010
Colin Shea-Blymyer @dr-bly.bsky.social · 30/04/2025
I've been keeping a list of organizations that do AI red-teaming. Sources are blog and job postings. Definitely an incomplete list. Potentially out of date. I might do a write-up about it. docs.google.com/document/d/1...
docs.google.com
AI Red Teams List
This document serves as my place to track and organize AI red teaming activity. Model Developers Organizations in this section are directly involved in the development of models and have expertise in...
010
Reposted by Colin Shea-Blymyer
Steph Batalis @stephbatalis.bsky.social · 02/04/2025
"We all rely on science [...] Businesses and farmers rely on science and engineering for product innovation, technological advances, and weather forecasting. Science helps humanity protect the planet and keeps pollutants and toxins out of our air, water, and food." docs.google.com/document/d/1...
docs.google.com
Public Statement on Supporting Science for the Benefit of All Citizens
TO THE AMERICAN PEOPLE We all rely on science. Science gave us the smartphones in our pockets, the navigation systems in our cars, and life-saving medical care. We count on engineers when we drive acr...
011
Colin Shea-Blymyer @dr-bly.bsky.social · 26/03/2025
ICYMI: The CSET webinar on AI red-teaming has been recorded! www.youtube.com/watch?v=gDnN... Watch this for a great discussion on what AI red-teaming is, how different organizations do it, and how it can be improved! Huge thanks to the panelists, moderator, and audience!
youtube.com
What’s Next for AI Red-Teaming?
YouTube video by Center for Security and Emerging Technology
110
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 26/03/2025
Almost everything we care about in AI comes down to evaluations, including #redteaming. Jessica Ji's post lays out a path for making AI red-teaming better. cset.georgetown.edu/article/how-...
cset.georgetown.edu
How to Improve AI Red-Teaming: Challenges and Recommendations | Center for Security and Emerging Technology
Despite recent upheaval in the AI policy landscape, AI evaluations—including AI red-teaming—will remain fundamental to understanding and governing the usage of AI systems and their impact on society. ...
021
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 25/03/2025
Starting in 5 👀
011
Colin Shea-Blymyer @dr-bly.bsky.social · 20/03/2025
🏀 I know nothing about college basketball, so I decided to be the avatar for ChatGPT in the @csetgeorgetown.bsky.social #MarchMadness tournament. I used Deep research on o3-mini-high to generate a detailed report and analysis of the tournament. The resulting bracket is very conservative. Go Humans!
A filled march madness bracket showing the first seed teams (Duke, Houston, Auburn, and Florida) going to the final four, and showing Duke going all the way.
100
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 17/03/2025
✨ NEW: How can the U.S. stay ahead in AI? OSTP called for input on developing an “AI Action Plan.” Our response outlines the steps the U.S. should take to: 1️⃣Secure and advance its AI leadership 2️⃣Navigate competition with China 3️⃣Realize AI’s benefits and avoid its risks
253
Colin Shea-Blymyer @dr-bly.bsky.social · 12/03/2025
Why does #AI red-teaming suck? How can we make it suck less? All this and more at the next CSET webinar: cset.georgetown.edu/event/whats-... Join me, the Director of Microsoft's AI Red Team, the MITRE ATLAS Lead, and the Director of Apollo Research. This will be an awesome conversation.
cset.georgetown.edu
What’s Next for AI Red-Teaming? | Center for Security and Emerging Technology
On March 25, CSET will host an in-depth discussion about AI red-teaming — what it is, how it works in practice, and how to make it more useful in the future.
041
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 12/03/2025
What: CSET Webinar 📺 When: Tuesday, 3/25 at 12PM ET 📅 What’s next for AI red-teaming? And how do we make it more useful? Join Tori Westerhoff, Christina Liaghati, Marius Hobbhahn, and CSET's @dr-bly.bsky.social * @jessicaji.bsky.social for a great discussion: cset.georgetown.edu/event/whats-...
cset.georgetown.edu
What’s Next for AI Red-Teaming? | Center for Security and Emerging Technology
On March 25, CSET will host an in-depth discussion about AI red-teaming — what it is, how it works in practice, and how to make it more useful in the future.
044
Colin Shea-Blymyer @dr-bly.bsky.social · 18/02/2025
I gave my takes from the Paris AI Action Summit to @thecipherbrief.bsky.social last week. I heard three Nobel Laureates warn about the risks of AI, while VP Vance was "not here ... to talk about AI safety" and focused on opportunity. I think we can innovate on AI without building unsafe products.
thecipherbrief.com
Expert Q&A: At Paris AI Summit, Speed V. Risk
A look at how embrace of AI opportunities weighed against concerns of AI risks at the Paris AI Action Summit.
010
Reposted by Colin Shea-Blymyer
Mina Narayanan @minanrn.bsky.social · 07/02/2025
@miahoffmann.bsky.social , @ojdaniels.bsky.social, and I wrote a piece on key AI governance areas to watch in 2025 with the upcoming AI Action Summit in mind. Check it out here! thebulletin.org/2025/02/will...
thebulletin.org
Will the Paris artificial intelligence summit set a unified approach to AI governance—or just be another conference?
AI innovations and governments’ preferences can make international consensus on governance at the Paris Summit challenging.
053
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 04/02/2025
We're hiring 📢 CSET is looking for a Research Fellow to analyze topics related to the development, deployment, and operations of AI & ML tools in the national security space. Interested or know someone who would be? Learn more and apply 👇 cset.georgetown.edu/job/research...
cset.georgetown.edu
Research Fellow - Applications | Center for Security and Emerging Technology
The Center for Security and Emerging Technology at Georgetown University (CSET) is seeking applications for a Research Fellow to support our Applications Line of Research. This role will analyze topic...
033
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 03/02/2025
CSET is hiring 📢 We’re looking for our next Director of Analysis to lead our research agenda & manage CSET's talented team of researchers. Interested or know someone who would be? Learn more and apply 👇 cset.georgetown.edu/job/director...
cset.georgetown.edu
Director of Analysis | Center for Security and Emerging Technology
The Center for Security and Emerging Technology at Georgetown University (CSET) is seeking applications for a Director of Analysis. This leadership role combines organizational strategy, research over...
032
Colin Shea-Blymyer @dr-bly.bsky.social · 03/02/2025
I'm packing my bags for an exciting week in Paris! I'll be contributing to the Paris Peace Forum's AI-Cyber Nexus discussions, and I'll be attending @iaseai.bsky.social '25 and the Paris AI Security Forum. I'm registered for a few other events, as well (if I can find the time and energy to attend).
010
Colin Shea-Blymyer @dr-bly.bsky.social · 29/01/2025
Just as "we don't need to choose between freedom and technology", we don't need to choose between safety and innovation in AI. The development of voluntary standards, best practices, and safeguards for AI makes AI better and easier to adopt.
120
Colin Shea-Blymyer @dr-bly.bsky.social · 22/01/2025
If you're interested in AI + Policy you need to follow my colleagues at @csetgeorgetown.bsky.social We have a starter pack: bsky.app/starter-pack...
020
Reposted by Colin Shea-Blymyer
CSET @csetgeorgetown.bsky.social · 11/12/2024
✨🏛️New interactive tool🏛️✨ Use the latest @emergingtechobs.bsky.social resource to track AI governance from D.C. to Beijing. ETO AGORA is a living collection of AI-relevant laws, regulations, standards, and other governance documents from the U.S. and around the world 🌎
133
Colin Shea-Blymyer @dr-bly.bsky.social · 27/11/2024
Workshop Announcement 📯 On Dec 4 (1 day after AISIC Plenary) @csetgeorgetown.bsky.social is hosting a workshop on the future of AI testing. Workshop will run from 9am-1pm on Georgetown's Capitol Campus. If you have experience testing AI systems please reach out to me for more details!
011